edge-ai

Tag

Cards List
#edge-ai

Running a 13M ASR conformer on a microcontroller

Reddit r/LocalLLaMA · 2026-07-20

Details a method to run a 13 million parameter ASR Conformer model directly on a microcontroller, highlighting advances in edge AI deployment.

0 favorites 0 likes
#edge-ai

Introducing Cosmos 3 Edge

Hugging Face Blog · 2026-07-20 Cached

NVIDIA released Cosmos 3 Edge, a 4-billion-parameter open world model for edge devices that helps robots and vision AI agents understand surroundings, reason in real time, and generate actions. It achieves best-in-class throughput and accuracy among similar-sized models.

0 favorites 0 likes
#edge-ai

NVIDIA Introduces New Jetson Thor Computers to Advance Mainstream Robotics and Edge AI

NVIDIA Blog · 2026-07-15 Cached

NVIDIA announces the Jetson T3000 and T2000 modules based on the Thor architecture, delivering powerful AI compute for mainstream robotics and edge AI applications, with adoption by leading robotics companies.

0 favorites 0 likes
#edge-ai

@NVIDIAAI: NVIDIA DeepStream 9.1 is here, with 13 agentic skills for building video analytics pipelines. Instead of manually build…

X AI KOLs Timeline · 2026-07-15 Cached

NVIDIA released DeepStream 9.1, an update to its video analytics toolkit, featuring 13 agentic skills that allow users to describe pipelines in natural language, new multi-view 3D tracking, automatic camera calibration, and support for Jetson Orin and Thor. The update is open source on GitHub.

0 favorites 0 likes
#edge-ai

PrismML Bonsai 27B is surprisingly usable on the Jetson Orin Nano 8GB

Reddit r/LocalLLaMA · 2026-07-14

PrismML's Bonsai 27B model runs on the Jetson Orin Nano 8GB with 4.31 tokens/s and 27 t/s prompt processing, using 6.2GB RAM and about 25W power. It indicates surprisingly usable edge AI performance.

0 favorites 0 likes
#edge-ai

Closed-Loop Control with Rule-Aligned Small Language Models and Multi-Agent Self-Correction

arXiv cs.AI · 2026-07-14 Cached

This paper explores using a compact Small Language Model (Qwen2.5-1.5B) retrained with GRPO and combined with a validator-guided correction loop for autonomous industrial control. The framework achieves high alignment accuracy and low latency, demonstrating practical viability for edge deployment.

0 favorites 0 likes
#edge-ai

Edge-Aware Thermal Infrared UAV Swarm Tracking

Hugging Face Daily Papers · 2026-07-14 Cached

This paper proposes an edge-aware online tracking pipeline for thermal infrared UAV swarm tracking, featuring the Adaptive Kinematic Kalman Filter (AKKF) that balances efficiency and robustness under challenging conditions.

0 favorites 0 likes
#edge-ai

EvoLP: Self-Evolving Latency Predictor for Model Compression in Real-Time Edge Systems

arXiv cs.LG · 2026-07-13 Cached

EvoLP is a self-evolving latency predictor for neural network models on edge devices, designed to guide model compression while satisfying strict latency constraints. It outperforms prior methods across multiple edge devices and model variants.

0 favorites 0 likes
#edge-ai

NVIDIA Physical AI Guide: Cosmos, Isaac, Jetson, Omniverse Explained

Reddit r/ArtificialInteligence · 2026-07-09 Cached

An in-depth guide explaining NVIDIA's physical AI infrastructure stack, including Cosmos, Isaac, Jetson, Omniverse, and edge AI, positioning the company as the gravitational center of physical AI with both cloud and edge capabilities.

0 favorites 0 likes
#edge-ai

@googledevs: Meet LiteRT.js: @Google’s new Edge AI runtime for the web! We've made it easier to convert from PyTorch to #WebAI using…

X AI KOLs Timeline · 2026-07-09 Cached

Google announces LiteRT.js, a high-performance JavaScript runtime for running AI models directly in the browser using WebAssembly and hardware acceleration, as an evolution from TensorFlow.js.

0 favorites 0 likes
#edge-ai

Federated Learning for Object Detection: Enabling Collaborative Drone Learning Without Centralizing Data

arXiv cs.LG · 2026-07-07 Cached

Applies federated learning to object detection for drone fleets, enabling collaborative training without centralizing aerial imagery, achieving performance close to centralized training while preserving privacy and reducing bandwidth.

0 favorites 0 likes
#edge-ai

Small AI Models Gain Traction In places with unreliable networks

Hacker News Top · 2026-07-06 Cached

Small AI models are proving valuable in regions with unreliable networks, enabling life-saving applications like counterfeit drug detection and disease identification in crops without needing constant internet connectivity.

0 favorites 0 likes
#edge-ai

SupraLabs/Supra-Router-51M

Hugging Face Models Trending · 2026-07-05 Cached

SupraLabs releases Supra-Router-51M, a 51.7M parameter micro-LLM for multi-task infrastructure routing, designed to decide whether to process prompts locally on edge or send them to cloud-hosted models. Fine-tuned on a small dataset, it uses multi-task sequence generation for robust routing.

0 favorites 0 likes
#edge-ai

@PyTorch: PyTorch Foundation supported the ExecuTorch Hackathon in San Francisco, where more than 100 participants across 20+ tea…

X AI KOLs Timeline · 2026-07-02 Cached

The PyTorch Foundation supported the ExecuTorch Hackathon in San Francisco, where over 100 participants built real-time on-device AI applications using PyTorch and ExecuTorch on Snapdragon-powered Samsung Galaxy S25 Ultra devices. Winning projects included SafeScreen AI, SixthSense, and Toddle AI, showcasing local execution benefits for responsiveness, privacy, and offline capability.

0 favorites 0 likes
#edge-ai

@ben_burtenshaw: Super excited to launch this hackathon to port AI models direct to bare silicon! Its is a community hackathon for peopl…

X AI KOLs Following · 2026-07-02 Cached

Ben Burtenshaw announces a community hackathon focused on porting AI models directly to bare silicon for edge applications like vision, speech, and robotics.

0 favorites 0 likes
#edge-ai

Into the Omniverse: Three Workflows for Improving Vision AI Agent Accuracy With Synthetic Data and Fine-Tuning

NVIDIA Blog · 2026-06-30 Cached

NVIDIA discusses three workflows using synthetic data and fine-tuning on its Omniverse and Metropolis platforms to improve vision AI agent accuracy for edge deployment.

0 favorites 0 likes
#edge-ai

Firefly Aerospace Operates NVIDIA Jetson in Lunar Orbit for the First Time

NVIDIA Blog · 2026-06-29 Cached

Firefly Aerospace is leveraging NVIDIA Jetson edge AI on its Blue Ghost Mission 2 to perform AI inference in lunar orbit, enabling near real-time processing of imaging data and reducing reliance on slow downlinks.

0 favorites 0 likes
#edge-ai

A language model that runs on 5$ chip. Comes with 12 AI applications. No cloud, no internet. Universal installer + Open source Github + Huggingface available. Test it yourself.

Reddit r/artificial · 2026-06-29

A language model capable of running on a $5 chip with 12 AI applications, fully offline and open source, available on GitHub and Hugging Face.

0 favorites 0 likes
#edge-ai

NASA testing local LLM inference for future space missions

Reddit r/LocalLLaMA · 2026-06-29 Cached

NASA is testing Red Hat's RamaLama open source tool to run local LLM and VLM inference for a medical AI assistant on deep space missions, enabling autonomous real-time diagnostics without Earth communication.

0 favorites 0 likes
#edge-ai

@loktar00: Took a while to get this all working right it still likes to forget it has access to Stackchan functionality, but this …

X AI KOLs Following · 2026-06-27 Cached

A developer demonstrates running the Hermes AI model locally on an M5 Stackchan robot, with plans to do the same on a Reachy Mini. This showcases local AI robotics integration.

0 favorites 0 likes
← Previous
Next →
← Back to home

Submit Feedback