on-device-ai

Tag

Cards List
#on-device-ai

SDXL running locally in the browser on WebGPU, open-source

Reddit r/LocalLLaMA · 2026-06-24

Stable Diffusion XL (SDXL) can now run locally in the browser using WebGPU, enabling high-quality AI image generation directly on-device with open-source code.

0 favorites 0 likes
#on-device-ai

I built a workout app that gets your friends and family actually working out with you.

Reddit r/AI_Agents · 2026-06-22

RepSquad is a workout app that uses on-device AI to count reps and score form in real time, enabling social challenges with friends and family for motivation.

0 favorites 0 likes
#on-device-ai

Pixi’s new iOS app turns text messages into interactive AR experiences

TechCrunch AI · 2026-06-18 Cached

Pixi launches an iOS messaging app that uses on-device AI to send interactive AR characters that respond to real-world surroundings and user emotions, aiming to make conversations more playful and present.

0 favorites 0 likes
#on-device-ai

I made this android app which runs ai models locally

Reddit r/artificial · 2026-06-15

A developer created an Android app that runs AI models locally, supporting GGUF and LiteRT formats with multiple ways to add models.

0 favorites 0 likes
#on-device-ai

Energy-Efficient On-Device RAG on a Mobile NPU: System Design and Benchmark on Snapdragon X Elite

arXiv cs.CL · 2026-06-11 Cached

This paper presents the first end-to-end RAG pipeline running entirely on a mobile NPU (Qualcomm Hexagon on Snapdragon X Elite), achieving up to 18x faster LLM prefilling and 4x lower energy vs. CPU, with no quality regression.

0 favorites 0 likes
#on-device-ai

Semantic distance as routing layer: an on-device, serverless alternative to the central-index model

Reddit r/LocalLLaMA · 2026-06-09

Proposes a decentralized information discovery system using on-device embedding models and peer-to-peer gossip, eliminating the need for central indexes like search engines.

0 favorites 0 likes
#on-device-ai

Apple reveals new AI architecture built around Google Gemini models

Hacker News Top · 2026-06-08 Cached

Apple announced a major overhaul of its Apple Intelligence platform, revealing a new AI architecture built on foundation models co-developed with Google using Gemini technologies, enabling multimodal capabilities and privacy-preserving on-device and server processing via Private Cloud Compute.

0 favorites 0 likes
#on-device-ai

Bringing Gemma 4 12B to your Laptop: Unlocking Local, Agentic Workflows with Google AI Edge

Reddit r/LocalLLaMA · 2026-06-05 Cached

Google announces the availability of Gemma 4 12B on laptops via Google AI Edge, enabling local, agentic, and multimodal workflows with tools like AI Edge Gallery and Eloquent.

0 favorites 0 likes
#on-device-ai

Are We Underestimating Small Edge AI Models?[D]

Reddit r/MachineLearning · 2026-06-05

A developer argues that the edge AI community overlooks small, specialized models that can run locally on devices like smartphones, using a self-built offline Morse code recognition feature as an example. The project uses a sub-5 MB AI model with TensorFlow/Keras and LiteRT, and the entire pipeline from data generation to mobile integration was custom-built.

0 favorites 0 likes
#on-device-ai

Meta's ships facial recognition on smart glasses

Hacker News Top · 2026-06-04 Cached

A security researcher discovered that Meta's Stella companion app for smart glasses (v273.0.0.21) contains a fully assembled, functional facial recognition pipeline—including three on-device models, a biometric embedding database, and a notification system—that is dormant on stock accounts but operable when invoked directly. The pipeline can detect faces, generate 2048-dimension embeddings, and fire 'Person Recognized' notifications, raising significant privacy concerns even though Meta has not been observed activating it for regular users.

0 favorites 0 likes
#on-device-ai

Run (your largest) local models from your iPhone

Reddit r/LocalLLaMA · 2026-06-04

A tool or app that enables users to run large local AI models directly from their iPhone, bringing on-device LLM inference to iOS.

0 favorites 0 likes
#on-device-ai

Unsloth on Apple Silicon- Pre-announcement announcement

Reddit r/LocalLLaMA · 2026-06-04

Unsloth, a popular LLM fine-tuning library, announces upcoming support for Apple Silicon devices, expanding its optimization capabilities beyond NVIDIA GPUs.

0 favorites 0 likes
#on-device-ai

The AI war is moving from models to machines and I don’t think enough people are talking about it

Reddit r/artificial · 2026-06-04

A commentary arguing that the AI competition is shifting from model quality to hardware placement and infrastructure, highlighting Microsoft's Project Solara, NVIDIA's RTX Spark, and ByteDance's custom CPU efforts as signs that agentic workloads are driving new silicon and deployment strategies.

0 favorites 0 likes
#on-device-ai

Google’s Gemma 4 12B just dropped - here’s how to run it locally on your Mac

Reddit r/artificial · 2026-06-04

Google released Gemma 4 12B, an Apache 2.0 open-source multimodal model supporting text, vision, and audio with a 256K context window. The article provides a guide for running it locally on Macs using Ollama, LM Studio, or llama.cpp.

0 favorites 0 likes
#on-device-ai

The Data Center Moves to Your Machine (4 minute read)

TLDR AI · 2026-06-03

Perplexity unveiled a hybrid local-cloud inference system at Computex 2026 that intelligently routes queries between on-device and cloud models, building on its earlier Personal Computer agent.

0 favorites 0 likes
#on-device-ai

Next

Reddit r/ArtificialInteligence · 2026-06-01

The article discusses the shift of AI from data centers to laptops and desktops, driven by Nvidia's new Arm-based processors and Microsoft's AI integration, leading to a major enterprise hardware upgrade cycle.

0 favorites 0 likes
#on-device-ai

@cjzafir: Fine-tune your first AI model today. Run GPT4o level model and run on your phone or laptop. @OpenBMB released 15M sampl…

X AI KOLs Following · 2026-05-29 Cached

OpenBMB released UltraData-SFT-2605, a 15M-sample high-quality SFT dataset for fine-tuning AI models like MiniCPM5-1B to run on phones or laptops.

1 favorites 1 likes
#on-device-ai

Intent to Prototype: Embedding API

Lobsters Hottest · 2026-05-26 Cached

The Chromium team proposes a new Embedding API for the web platform that allows developers to generate vector embeddings on-device using Chrome's AI infrastructure, enabling privacy-preserving semantic search, retrieval-augmented generation, and content clustering while reducing latency and cost.

0 favorites 0 likes
#on-device-ai

Run Chrome’s tiny Gemma4 (aka Gemini Nano) directly on PC without GPU

Reddit r/LocalLLaMA · 2026-05-23

A developer created a Chrome extension called Dobby that runs Google's Gemma4 (Gemini Nano) locally on PC without needing a GPU, requiring only Chrome and 16GB RAM. The extension provides a simple interface to interact with the model for tasks like spell checking or summarizing.

0 favorites 0 likes
#on-device-ai

@sharbel: The 10 fastest growing GitHub repos this week: 1. codegraph (+14.1K stars) Pre-indexed code knowledge graph for Claude …

X AI KOLs Timeline · 2026-05-23 Cached

A weekly roundup of the 10 fastest-growing GitHub repositories, highlighting AI infrastructure tools like code knowledge graphs, agent memory, on-device intelligence, and more.

0 favorites 0 likes
← Previous
Next →
← Back to home

Submit Feedback