mobile-ai

Tag

Cards List
#mobile-ai

HybridInfer: Thermal-Aware Reinforcement-Learning Tier Routing for On-Device, Edge, and Cloud LLM Inference

arXiv cs.LG ↗ · yesterday Cached

HybridInfer introduces a thermal-aware reinforcement-learning router for multi-tier LLM inference, achieving higher quality and cost efficiency than heuristics on real Android devices.

0 favorites 0 likes
#mobile-ai

Qwen Intelligence Launches Three Mobile AI Agents (1 minute read)

TLDR AI ↗ · 6d ago Cached

Qwen Intelligence launches three state-of-the-art mobile AI agents: Mobile Planner Agent for task decomposition, Mobile-Use Agent with high end-to-end success rates, and Mobile Creative Agent for rapid content generation, alongside an open benchmark suite for evaluation.

0 favorites 0 likes
#mobile-ai

TierKV: Long-Context On-Device LLMs via Predictive Multi-Tier KV Caching

arXiv cs.LG ↗ · 2026-09-21 Cached

TierKV proposes a predictive multi-tier KV caching framework to optimize memory usage and throughput for long-context LLMs on mobile devices, achieving significant performance improvements with minimal accuracy degradation.

0 favorites 0 likes
#mobile-ai

Edge0/Edge0-35B-A3B-preview

Hugging Face Models Trending ↗ · 2026-09-08 Cached

Edge0-35B-A3B-preview is a sparse MoE model that enables efficient AI inference on mobile devices by using streaming expert offloading and quantization, achieving 15 tok/s with under 3 GiB of memory.

0 favorites 0 likes
#mobile-ai

SnapBench: Benchmarking Snap-and-Ask Multimodal Retrieval for Mobile Interactions

Hugging Face Daily Papers ↗ · 2026-08-30 Cached

SnapBench introduces a paired corruption benchmark for robust snap-and-ask multimodal retrieval on mobile, revealing image noise as a key degradation factor and proposing an adaptive fusion method for modality reliability.

0 favorites 0 likes
#mobile-ai

@MaximeRivest: what if: apple qualcomm google are the real ai companies of the future because we will have fable level ai in mobile? I…

X AI KOLs Following ↗ · 2026-08-28 Cached

A tweet speculates that Apple, Qualcomm, and Google could become leading AI companies by focusing on portable, edge AI devices, referencing Google Gemma running on NVIDIA Jetson.

0 favorites 0 likes
#mobile-ai

Running an Agent on a Ubuntu Touch smartphone feels like having a Black Mirror character in a box with you at all times.

Reddit r/AI_Agents ↗ · 2026-08-15

The article describes the experience of running an AI agent on an Ubuntu Touch smartphone, which is always on with access to sensors, allowing continuous conversation via Telegram.

0 favorites 0 likes
#mobile-ai

@GoogleDeepMind: We built SL2T with the Deaf community - guided by Deaf Googlers and our AI Sign Language Advisory Committee. Bringing A…

X AI KOLs ↗ · 2026-08-12 Cached

Google DeepMind announced SL2T, a system developed with the Deaf community to bring ASL input to phones, with plans to expand to more sign languages and applications.

0 favorites 0 likes
#mobile-ai

I thought I’d done something extraordinary by running massive models on standard smartphones but

Reddit r/AI_Agents ↗ · 2026-07-30

Creator of bigedgeonmoe open-source codebase enables running massive MoE models (up to 120B parameters) on mobile devices and consumer PCs, achieving 6 tokens/s for Qwen 35B on a mid-range phone.

0 favorites 0 likes
#mobile-ai

Keyword Matters: Unveiling the Energy Sensitivity of On-Device LLM Prompting

arXiv cs.AI ↗ · 2026-07-28 Cached

This paper empirically studies how prompt wording affects energy consumption for on-device LLMs, showing that keyword choices can significantly impact decoding length and total energy, suggesting prompt engineering as a lightweight energy optimization lever.

0 favorites 0 likes
#mobile-ai

LightMem-Ego: Your AI Memory for Everyday Life

Hugging Face Daily Papers ↗ · 2026-07-13 Cached

LightMem-Ego is a lightweight streaming multimodal memory system for everyday life assistance that continuously captures egocentric visual and audio streams, organizes them into hierarchical memory, and retrieves grounded answers to user queries about past experiences, deployable on smartphones and AI glasses.

0 favorites 0 likes
#mobile-ai

Compressed Version of Qwen-3.6-27B coming from PrismML - Khosla-Backed Startup Claims Breakthrough With Largest-Ever AI Model on an iPhone

Reddit r/LocalLLaMA ↗ · 2026-07-13

PrismML, a Khosla-backed startup, releases a compressed version of Qwen-3.6-27B, claiming it's the largest AI model ever to run on an iPhone.

0 favorites 0 likes
#mobile-ai

@PyTorch: PyTorch Foundation supported the ExecuTorch Hackathon in San Francisco, where more than 100 participants across 20+ tea…

X AI KOLs Timeline ↗ · 2026-07-02 Cached

The PyTorch Foundation supported the ExecuTorch Hackathon in San Francisco, where over 100 participants built real-time on-device AI applications using PyTorch and ExecuTorch on Snapdragon-powered Samsung Galaxy S25 Ultra devices. Winning projects included SafeScreen AI, SixthSense, and Toddle AI, showcasing local execution benefits for responsiveness, privacy, and offline capability.

0 favorites 0 likes
#mobile-ai

@heyshrutimishra: The keyboard hasn't changed in 20 years. Apple gave us touch typing in 2007. Phones grew from 3.5 inches to 6.7 inches.…

X AI KOLs Timeline ↗ · 2026-06-30 Cached

Announces Acti, an agentic keyboard that turns the passive keyboard into an active interface capable of executing actions directly within apps, such as pulling Notion docs or generating mini apps.

0 favorites 0 likes
#mobile-ai

Gemma 12b less than 10 watts 6.5pp 1.3tg

Reddit r/LocalLLaMA ↗ · 2026-06-14

Running Gemma 12B model on a Google Pixel 10 Pro using llama.cpp achieves 6.5 tokens per second prompt processing and 1.3 tokens per second generation with under 10 watts power consumption, demonstrating efficient on-device AI inference.

0 favorites 0 likes
#mobile-ai

Energy-Efficient On-Device RAG on a Mobile NPU: System Design and Benchmark on Snapdragon X Elite

arXiv cs.CL ↗ · 2026-06-11 Cached

This paper presents the first end-to-end RAG pipeline running entirely on a mobile NPU (Qualcomm Hexagon on Snapdragon X Elite), achieving up to 18x faster LLM prefilling and 4x lower energy vs. CPU, with no quality regression.

0 favorites 0 likes
#mobile-ai

Local iPhone AI image generation is getting practical - only 3 seconds per image

Reddit r/ArtificialInteligence ↗ · 2026-06-03

Benchmark shows local Stable Diffusion 1.5 on iPhone can generate 512x512 images in as little as 3.1 seconds using optimized models like Realistic Vision V5.1 Hyper, making on-device AI image generation practical.

0 favorites 0 likes
#mobile-ai

Ready or Not, the AI Phones Are Coming

Reddit r/artificial ↗ · 2026-06-01

This article discusses the imminent arrival of AI-powered smartphones and the implications for consumers and the tech industry.

0 favorites 0 likes
#mobile-ai

The question with Gemini on Android is not just privacy. It is the action boundary.

Reddit r/AI_Agents ↗ · 2026-05-24

This article argues that the real issue with integrating Gemini deeper into Android isn't just privacy, but the action boundary—what the AI can read, suggest, draft, change, send, buy, or delete—and proposes a tiered consent model for different levels of AI agency.

0 favorites 0 likes
#mobile-ai

Vibe coding is coming to your phone

The Verge ↗ · 2026-05-20 Cached

Google and Apple are bringing AI-powered 'vibe coding' to mobile, allowing users to create custom Android apps, widgets, and shortcuts via natural language prompts, as demonstrated at Google I/O 2026 and reported for iOS.

0 favorites 0 likes
Next →
← Back to home

Submit Feedback