edge-computing

Tag

Cards List
#edge-computing

5x faster Edge Functions: V8 isolates to Firecracker MicroVMs

Hacker News Top ↗ · 6h ago Cached

Netlify rebuilt Edge Functions infrastructure on Firecracker MicroVMs (with Unikraft), cutting warm-invocation latency from 25–40ms to ~5–6ms at p50 — roughly 5x faster — while improving security and reliability without changing how developers write functions.

0 favorites 0 likes
#edge-computing

Encoder-Sharing Hierarchical Federated Multi-Task Learning for VANETs

arXiv cs.LG ↗ · 21h ago Cached

This paper proposes EN-HMTFL, an encoder-sharing hierarchical multi-task federated learning framework for vehicular ad hoc networks that lets vehicles train heterogeneous perception tasks collaboratively via a shared encoder while keeping task-specific decoders local, improving accuracy by up to 24% and reducing communication rounds by up to 28.8%.

0 favorites 0 likes
#edge-computing

d1 (2 minute read)

TLDR AI ↗ · yesterday Cached

Liquid AI and Artificial Analysis release Pipette, an open-source benchmarking suite for on-device AI that measures quality, speed, latency, and memory use across models, quantizations, runtimes, and devices, shipping with 10k+ verified results and clients for macOS, Windows, iOS, and Android.

0 favorites 0 likes
#edge-computing

Forlinx 20-TOPS M.2 AI accelerator supports PCIe cascading for local LLM inference

Reddit r/LocalLLaMA ↗ · yesterday

Forlinx Embedded has launched an M.2 AI accelerator card based on Rockchip's RK1820 and RK1828 processors, providing 20 TOPS of INT8 performance for local LLM inference on embedded Linux and Android systems.

0 favorites 0 likes
#edge-computing

HybridInfer: Thermal-Aware Reinforcement-Learning Tier Routing for On-Device, Edge, and Cloud LLM Inference

arXiv cs.LG ↗ · yesterday Cached

HybridInfer introduces a thermal-aware reinforcement-learning router for multi-tier LLM inference, achieving higher quality and cost efficiency than heuristics on real Android devices.

0 favorites 0 likes
#edge-computing

PTC-Decoder: Towards Intelligent SLMs on Offline Resource-Constrained Edge Devices

arXiv cs.AI ↗ · 2d ago Cached

PTC-Decoder is a training-free, plug-and-play framework that improves small language models on offline, resource-constrained edge devices by enforcing plan adherence through token-level constraints, enhancing step-level reliability in multi-step agent tasks.

0 favorites 0 likes
#edge-computing

@MaximeRivest: Here is when and how to finetune a specialized decision model (aka classifier) that is better then Opus, faster then Je…

X AI KOLs Timeline ↗ · 3d ago Cached

The article explains when and how to finetune a specialized decision model classifier that outperforms Opus, is faster than Jev, and runs on user devices.

0 favorites 0 likes
#edge-computing

@mylifcc: Cloudflare is shaking up backend engineers' jobs this time Workers can now directly accept TCP connections and even run…

X AI KOLs Timeline ↗ · 4d ago

Cloudflare has updated its Workers service to directly accept TCP connections and run gRPC, enabling backend workloads that previously required dedicated servers to be deployed on its global edge network.

0 favorites 0 likes
#edge-computing

Ling Tiny 3.0 is a glimpse of the future

Reddit r/LocalLLaMA ↗ · 5d ago

The author runs the Ling Tiny 3.0 AI model on a 2017 laptop without GPU, achieving 10 tokens per second and completing tasks like code generation, showcasing the potential for edge intelligence on existing hardware.

0 favorites 0 likes
#edge-computing

@PrajwalTomar_: Stop what you are doing and read this. I built an AI for a factory floor that works with the wifi OFF and honestly it's…

X AI KOLs Timeline ↗ · 6d ago Cached

An individual built an AI system for factory floors that operates offline to assist technicians by providing historical solutions for machine error codes.

0 favorites 0 likes
#edge-computing

When Labels Are Scarce: An Oscillatory State Space Model for Vibration Diagnosis

arXiv cs.LG ↗ · 6d ago Cached

DualRes is a compact oscillatory state-space model for machine fault diagnosis from vibration data, achieving state-of-the-art performance with limited labels and reduced computational requirements for edge deployment.

0 favorites 0 likes
#edge-computing

ZO-COSMO: Index-Free One-Hop Mixing for Decentralized Zeroth-Order Optimization

arXiv cs.LG ↗ · 6d ago Cached

The paper introduces ZO-COSMO, an index-free method for decentralized zeroth-order optimization that improves accuracy by avoiding sparse index transmission through one-hop mixing. Experiments demonstrate gains over baseline methods on synthetic agents and Qwen LoRA workers.

0 favorites 0 likes
#edge-computing

@yoheinakajima: glance-vlm speedlab is now open source! read: https://glance.yohei.me/speed/ try: https://github.com/yoheinakajima/glan…

X AI KOLs Timeline ↗ · 2026-09-23 Cached

The article presents an open-source study on optimizing latency for local vision-language models through benchmarking and techniques like native batching and MLX quantization, achieving significant speedups while maintaining decision accuracy on Apple hardware.

0 favorites 0 likes
#edge-computing

@MSFTResearch: Robots are getting smarter, but how can their hardware match that growth? New Microsoft Research findings show that mov…

X AI KOLs Following ↗ · 2026-09-23 Cached

Microsoft Research findings demonstrate that offloading AI inference from robots to edge or cloud systems enhances task success rates, efficiency, and battery life for physical AI applications.

0 favorites 0 likes
#edge-computing

An Affordable AI-Integrated Smart Cane for Multimodal Mobility Assistance of Visually Impaired Users

arXiv cs.AI ↗ · 2026-09-23 Cached

This paper presents an affordable, offline AI-integrated smart cane designed to assist visually impaired users with multimodal mobility assistance, using edge computing on a Raspberry Pi Zero 2W for real-time obstacle detection and feedback.

0 favorites 0 likes
#edge-computing

TierKV: Long-Context On-Device LLMs via Predictive Multi-Tier KV Caching

arXiv cs.LG ↗ · 2026-09-21 Cached

TierKV proposes a predictive multi-tier KV caching framework to optimize memory usage and throughput for long-context LLMs on mobile devices, achieving significant performance improvements with minimal accuracy degradation.

0 favorites 0 likes
#edge-computing

Radio-Frequency Convolutional Neural Networks

arXiv cs.LG ↗ · 2026-09-18 Cached

Radio-frequency convolutional neural networks (RF-CNNs) repurpose existing wireless communication hardware for efficient AI inference on edge devices, demonstrating deep CNN performance with significant energy savings.

0 favorites 0 likes
#edge-computing

Small AI models let drones autonomously identify and attack battlefield targets

Ars Technica ↗ · 2026-09-17 Cached

A NATO-backed startup, Scaleout Systems, is using small AI models and federated learning to enable drones to autonomously identify and attack targets in battlefield settings, leveraging edge computing for real-time updates.

0 favorites 0 likes
#edge-computing

Where Should Agents Live? Energy-Memory Characterization of Agentic AI for the Edge-Cloud Continuum

arXiv cs.AI ↗ · 2026-09-17 Cached

This paper introduces agentic-eCAL, a metric for evaluating energy costs in multi-agent AI workflows across edge-cloud networks, demonstrating that transmission costs are minimal but inference costs are significant.

0 favorites 0 likes
#edge-computing

If you layer AI over your existing security cameras, how do you avoid destroying your upload bandwidth?

Reddit r/AI_Agents ↗ · 2026-09-16

The user explores adding AI object detection to older security cameras without straining network bandwidth, considering edge processing with on-premises appliances to analyze video locally and only send alerts or short clips, while seeking advice on hybrid architecture and API integration.

0 favorites 0 likes
Next →
← Back to home

Submit Feedback