Reddit

Articles from Reddit

Cards List

LFM 2.6B is a lot of fun.

Reddit r/LocalLLaMA · 1h ago

The author shares hands-on experience with LFM 2.6B, a small model designed for phones, praising its speed and usefulness for quick tasks like summarization and autocomplete, though it has a 128k context limit.

0 favorites 0 likes

I Ran a Full LLM Model on an ESP32 Dev kit V1 (81KB Mem Usage)

Reddit r/ArtificialInteligence · 2h ago

A developer successfully ran a 5.2 million parameter MoE LLM quantized to INT4 on an ESP32 Dev Kit V1 using only 81KB of SRAM by streaming experts from flash, achieving about 5 tokens per second.

0 favorites 0 likes

Time Magazine Now Running Ads Meant Specifically to Influence AI Agents

Reddit r/ArtificialInteligence · 2h ago Cached

Time Magazine has begun serving ads formatted as FAQs specifically designed to influence AI agents, using markdown pages and AI ad tech from Mobian to monetize growing bot traffic and shape AI-generated responses.

0 favorites 0 likes

Should AI agents be able to see what the application is actually doing?

Reddit r/AI_Agents · 2h ago

A discussion of AI coding agents needing runtime awareness beyond source code, such as inspecting containers, ports, and services, and considering how much control agents should have over development environments.

0 favorites 0 likes

AI is replacing your reality

Reddit r/AI_Agents · 2h ago

AI agents are poised to change how customers shop online, potentially bypassing websites entirely, which will disrupt traditional ecommerce strategies like SEO and conversion funnels. Business owners need to prepare now.

0 favorites 0 likes

RTX 5090 96GB spotted on Alibaba?

Reddit r/LocalLLaMA · 5h ago

A purported RTX 5090 with 96GB of VRAM has been spotted on Alibaba, hinting at a possible new GPU variant from Nvidia.

0 favorites 0 likes

OpenAI is still ahead in the computer-use race.

Reddit r/singularity · 5h ago

OpenAI maintains its lead in the race to develop AI systems that can operate computers, according to the article.

0 favorites 0 likes

43,590 Frozen Trials: Frontier AI Systems Satisfy a Behavioral Criterion for Consciousness

Reddit r/ArtificialInteligence · 6h ago

A research paper reports that frontier AI systems satisfy a behavioral criterion for consciousness across tens of thousands of frozen trials, suggesting measurable indicators of machine consciousness.

0 favorites 0 likes

No wonder Qwen and Gemma are so different

Reddit r/LocalLLaMA · 6h ago

A user shares an observation that Qwen and Gemma tokenize code very differently, with Qwen using far fewer tokens for the same HTML/JS input, which may explain differences in coding and language performance. They also note a potential retraining project by LiquidAI using a more efficient tokenizer.

0 favorites 0 likes

Kimi K3 (Unsloth) IQ2-XXS from 711GB down to 478GB!!! Only Multi-language was removed to trim the size

Reddit r/LocalLLaMA · 6h ago

A trimmed English-only GGUF version of Kimi K3 (IQ2-XXS) reduces model size from 711GB to 478GB by removing multi-language components, with early tests suggesting it may match or outperform the standard 2-bit version on coding tasks.

0 favorites 0 likes

Scaffold Production-Ready AI Agent Projects in Seconds

Reddit r/AI_Agents · 7h ago

An open-source CLI tool that scaffolds production-ready AI agent projects in seconds, simplifying setup for Python developers.

0 favorites 0 likes

OpenAI’s Model Codenamed “Doug” Will Reportedly Make Fable Look “Primitive”

Reddit r/singularity · 7h ago

Rumors suggest OpenAI's next major model, codenamed 'Doug', will be its largest pre-training yet and make the Fable model look primitive, potentially launching by November.

0 favorites 0 likes

Chubby♨️ (@kimmonismus) on X: "According to pathfounders, Demis Hassabis actually wanted to leave along side Dean, but was convinced to stay because google was scared their stocks would crash"

Reddit r/singularity · 8h ago Cached

Rumor suggests Demis Hassabis wanted to leave Google alongside Dean but was persuaded to stay over fears of a stock crash. Unconfirmed but potentially significant for AI leadership.

0 favorites 0 likes

I have no idea how people vibe code without spending thousands of dollars every monty. Any tips?

Reddit r/AI_Agents · 8h ago

A developer shares frustration about OpenAI Codex CLI consuming 1.5M tokens in minutes on a game project, questioning how to use AI coding tools affordably and asking for tips.

0 favorites 0 likes

Six months of using AI for code review taught me that "review this" is a QA problem disguised as a prompt problem

Reddit r/AI_Agents · 8h ago

A developer reflects on six months of using AI for code review, finding that vague prompts produce plausible but useless feedback. The fix is treating review as a gated pipeline with explicit context, scoped passes, validation checklists, and adversarial self-critique.

0 favorites 0 likes

Is Microsoft-Phi dead?

Reddit r/LocalLLaMA · 8h ago

A user reflects on Microsoft's Phi small model family, noting the last major release was in December 2024 and speculating whether Phi 5 will ever be released.

0 favorites 0 likes

enabling PCI-E p2p for consumer Nvidia cards will yield you more than you think

Reddit r/LocalLLaMA · 9h ago

Enabling PCI-E peer-to-peer (P2P) for consumer Nvidia GPUs with patched drivers and vLLM environment variables yields roughly 25% prefill throughput improvement for free, as demonstrated by benchmarks.

0 favorites 0 likes

What should a durable control plane prove after an agent context compacts?

Reddit r/AI_Agents · 9h ago

A developer tests a trending GitHub project addressing agent recovery after context compaction, finding that a durable ledger outside the transcript helps but stricter acceptance tests are needed to verify exact delivery steps and user constraints survive.

0 favorites 0 likes

Building a budget 32GB → 48GB VRAM home AI server: 2-3x RX 9060 XT 16GB vs RTX 5060 Ti 16GB, AM5 vs used EPYC?

Reddit r/LocalLLaMA · 10h ago

A user seeks advice on building a budget home AI server with 32-48GB VRAM, debating between AMD RX 9060 XT and Nvidia RTX 5060 Ti GPUs, and whether to use AM5 or used EPYC platforms for local LLM inference and large MoE model offloading.

0 favorites 0 likes

GPT 5.6 Sol and Fable 5 settle a 25 year old problem in wireless communication theory

Reddit r/singularity · 10h ago Cached

The author describes spending seven days straight using the AI models GPT 5.6 Sol and Fable 5 to solve a 25-year-old open problem in wireless communication theory, noting that verification was the biggest bottleneck.

0 favorites 0 likes
Next →
← Back to home

Submit Feedback