local-ai

Tag

Cards List
#local-ai

I built a local realtime voice stack for Ollama: Parakeet STT → Qwen 2.5 7B → Qwen3-TTS

Reddit r/LocalLLaMA · 5h ago

The author built a local realtime voice stack using Parakeet STT, Qwen 2.5 7B, and Qwen3-TTS, integrated with Ollama.

0 favorites 0 likes
#local-ai

@TeksEdge: What is the cheapest sane way to get 128GB+ of memory for Local AI in 2026? A Reddit user did the math, and the choices…

X AI KOLs Timeline · 22h ago Cached

A Reddit user compares the cheapest hardware options for achieving 128GB+ memory for local AI in 2026, covering used GPUs, unified memory systems, and cloud alternatives.

0 favorites 0 likes
#local-ai

@TheAhmadOsman: All panels and presentations from our Local AI Summit at AIE's World's Fair 2026 is now available to watch online Watch…

X AI KOLs Timeline · yesterday Cached

All panels and presentations from the Local AI Summit at AIE's World's Fair 2026 are now available to watch online, showcasing talks on making local AI the default.

0 favorites 0 likes
#local-ai

@Nativ_AI: Run Qwen 3.5 9B locally with fully customizable system prompts. Make it sing. Make it rhyme. Make it brainstorm, write,…

X AI KOLs Timeline · yesterday Cached

Nativ is an open-source macOS app that runs Qwen 3.5 9B and other open models locally on Apple Silicon, offering customizable system prompts, telemetry, and integrations with coding agents, with no cloud or subscription required.

0 favorites 0 likes
#local-ai

@TheAhmadOsman: Last week @MikeBradleyAI went on @AndrewWarner's show to demo how ODS is the easiest way to get started with Local AI W…

X AI KOLs Following · 2d ago Cached

A tweet promoting OsmanticAI's ODS as the easiest way to get started with local AI, featuring a demo by Mike Bradley on Andrew Warner's show.

0 favorites 0 likes
#local-ai

Unsloth's Gemma 4 mmproj silently broke vision & audio on newer llama.cpp builds — anyone else hit this?

Reddit r/LocalLLaMA · 2d ago

A developer reports that Gemma 4 multimodal features broke in newer llama.cpp builds when using Unsloth's GGUF models, due to an incompatible mmproj file. Switching to ggml-org's official models fixed the issue, highlighting a recurring compatibility concern between third-party quantizers and llama.cpp updates.

0 favorites 0 likes
#local-ai

@yiyirats: When relying on third-party voice services, data privacy and stability are always factors to consider. Voicebox is a locally run AI voice studio; all processing is done locally, data is not uploaded to the cloud, and no account registration is required. The features are quite comprehensive: supports Qwen3-TTS, LuxTTS, Chatterbo…

X AI KOLs Timeline · 2d ago Cached

Voicebox is a locally run open-source AI voice studio that supports 7 TTS engines, 23 languages, and voice cloning. All processing is done locally to protect privacy. The project has received 33.8k stars on GitHub.

0 favorites 0 likes
#local-ai

@DanKornas: Cloud-based AI services can expose sensitive documents, communications, and creative work to privacy risks when they ar…

X AI KOLs Timeline · 2d ago Cached

NativeMind is a private, open-source browser extension that runs local AI models via Ollama or WebLLM, enabling offline AI features for privacy-conscious users.

0 favorites 0 likes
#local-ai

Inkling-Small 276B-A12B at ~2.9 tok/s on <10gb memory

Reddit r/LocalLLaMA · 3d ago

Mference, a Swift + Metal inference engine, now supports Inkling-Small 276B-A12B, running it at ~2.9 tok/s on under 10GB memory, enabling large MoE models on consumer Apple hardware.

0 favorites 0 likes
#local-ai

MacPaw taps Liquid AI to offer on-device inference to devs building for its app store

TechCrunch AI · 3d ago Cached

MacPaw partners with Liquid AI to bring on-device AI inference and local memory to its products and app store, planning to offer the tech stack to developers and introduce credit-based AI pricing.

0 favorites 0 likes
#local-ai

@Skaly__Bull: Traditional AI stack is walking dead They just don't know it yet $10K enterprise servers, data-center GPUs, racks and c…

X AI KOLs Timeline · 3d ago Cached

The author argues that the traditional enterprise AI stack is obsolete, claiming a $599 Mac mini running Ollama can handle 80% of AI workloads locally for a fraction of the cost of renting cloud GPUs.

0 favorites 0 likes
#local-ai

@ddalcu: What a crazy last 4 days for local AI... insane... https://github.com/ddalcu/mlx-serve/releases/tag/v26.8.2… @liquidai …

X AI KOLs Timeline · 4d ago Cached

A developer updates MLX-Serve, a fast local inference server for Apple Silicon, to support recent models like LiquidAI 2.6B, MiniMax H3 video generation, and DeepSeek V4 Flash, with AntLing 3.0-flash coming soon.

0 favorites 0 likes
#local-ai

A 2.6B model with tool calling and 128K context now runs at 30 tok/s on a phone

Reddit r/LocalLLaMA · 4d ago

Liquid AI released LFM2.5-2.6B, a 2.69B parameter model with 128K context and tool calling, optimized for multi-step agent workflows and capable of running at 30 tok/s on a phone with a 1.67GB Q4_K_M GGUF, though coding and knowledge-heavy tasks remain weak compared to larger models.

0 favorites 0 likes
#local-ai

@DanKornas: Local AI models are often disconnected from the webpages you need to read, question, and work with. Page Assist is an o…

X AI KOLs Timeline · 5d ago Cached

Page Assist is an open-source browser extension that adds a sidebar and web UI for chatting with local AI models on any webpage, supporting Ollama, Gemini Nano, and OpenAI-compatible endpoints.

0 favorites 0 likes
#local-ai

Show HN: Run an 80B Qwen in 4.3 GB of RAM on a Mac, and a 35B on an iPhone

Hacker News Top · 5d ago

Demonstrates running an 80B Qwen model in just 4.3 GB of RAM on a Mac and a 35B model on an iPhone, showcasing extreme memory optimization for local LLM inference.

0 favorites 0 likes
#local-ai

I CANNOT believe I've got DeepSeek-V4-Flash-0731, a frontier model, running on my home PC. Insane!

Reddit r/LocalLLaMA · 5d ago

A user expresses astonishment at running DeepSeek-V4-Flash-0731, a frontier model, on a mid-range Windows PC with 24GB VRAM via quantization, highlighting rapid progress in local AI.

0 favorites 0 likes
#local-ai

Show HN: Nightcrawler – A local AI pentesting agent running on a smartphone

Hacker News Top · 5d ago Cached

Nightcrawler is an autonomous penetration testing agent that runs entirely on a smartphone, using a local 1.2B AI model on the phone's GPU to discover hosts, map services, and generate pentest reports without cloud connectivity.

0 favorites 0 likes
#local-ai

SpeakoFlow

Product Hunt · 5d ago

SpeakoFlow is an open-source local voice assistant for desktop, allowing users to interact with their computer via voice without cloud dependency.

0 favorites 0 likes
#local-ai

@dee_hw: Qwen3.8-27B is coming. We open source our 2x RTX 5090 build, so you can host it locally: on-prem, private, no rate limi…

X AI KOLs Timeline · 5d ago Cached

Autonomous AI open-sources a Personal AI Computer build powered by 2x or 4x NVIDIA RTX 5090s, enabling fully local, private AI hosting without API costs or rate limits.

0 favorites 0 likes
#local-ai

Your AI has amnesia. Here's what I did about it

Reddit r/ArtificialInteligence · 5d ago

The author introduces Clippy Vision, a 100% local open-source desktop assistant that tracks your screen context so AI tools can help without copy-pasting or cloud uploads. A Windows executable is available and source code is shared.

0 favorites 0 likes
Next →
← Back to home

Submit Feedback