Models

Cards List

Show HN: Needle2: 14MB agentic LLM for phones, wearables, smart home and robots

Hacker News Top · 11h ago Cached

Cactus Compute releases Needle 2, a 45M-parameter agentic LLM compressed to a 14MB binary for phones, wearables, smart home and robots, achieving 500+ tokens/sec on a Raspberry Pi 5 and running in 28MB RAM.

0 favorites 0 likes

@OpenAI: Advanced capabilities require strong safeguards. That’s why access is limited to approved defenders, with additional co…

X AI KOLs · 11h ago Cached

OpenAI is expanding Daybreak with two access tiers (Blue and Red) and introducing GPT-5.6-Cyber, a purpose-trained cybersecurity model that significantly reduces refusals for authorized defensive security work.

0 favorites 0 likes

@OpenAI: We've used GPT-5.6-Cyber extensively in real-world vulnerability research, including work that uncovered previously unk…

X AI KOLs · 11h ago Cached

OpenAI highlights the use of GPT-5.6-Cyber in real-world vulnerability research, including discovering previously unknown bugs in open-source software like Chrome's V8 engine.

0 favorites 0 likes

Needle 2: 14MB agentic LLM for phones, wearables, smart home and robots.

Reddit r/LocalLLaMA · 11h ago

Cactus releases Needle 2, a 14MB agentic LLM for phones, wearables, smart home devices, and robots, achieving fast inference on low-end hardware and supporting structured extraction and fine-tuning.

0 favorites 0 likes

GPT-5.6 Sol hits the ZeroBench human baseline at pass@5 without tools

Reddit r/singularity · 12h ago

GPT-5.6 Sol reportedly hits the ZeroBench human baseline at pass@5 without tools, meaning at least one of five attempts succeeds on the benchmark.

0 favorites 0 likes

Build Low-Latency Multilingual Voice Agents: Open Weights & Full Deployment Control with NVIDIA Magpie TTS

Hugging Face Blog · 12h ago Cached

NVIDIA announces Magpie Multilingual TTS, an open-weights text-to-speech model supporting 12 languages with low-latency deployment via NVIDIA NIM for building production voice agents.

0 favorites 0 likes

Meta’s new Glimmer AI model offers a hint at Zuckerberg’s personal intelligence vision

TechCrunch AI · 12h ago Cached

Meta released Muse Glimmer, an open-weight 30B-parameter AI model designed to run capable personal agents locally on consumer hardware, offering a concrete glimpse into Mark Zuckerberg's vision of distributed personal superintelligence.

0 favorites 0 likes

Muse Glimmer ACTUALLY fits on a single RTX 3090

Reddit r/LocalLLaMA · 14h ago

User reports that Muse Glimmer, a 30B model, fits on a single RTX 3090 with full 256k context using Q4_K_XL quantization and DFlash, achieving 64-124 tok/s and perfect long-context retrieval, unlike comparable models.

0 favorites 0 likes

@PyTorch: Today @AIatMeta introduced Muse Glimmer, an open-weight, 30-billion-parameter model distilled from Meta’s Muse Spark fo…

X AI KOLs Following · 14h ago Cached

Meta introduced Muse Glimmer, an open-weight 30B-parameter model distilled from Muse Spark for on-device agentic workflows, with ExecuTorch now supporting running it on NVIDIA GPUs and Apple silicon.

0 favorites 0 likes

@TheAhmadOsman: Will Qwen 3.8 27B make a comeback against Muse Glimmer 30B? In all cases, extremely happy about this release

X AI KOLs Following · 16h ago Cached

A user expresses excitement about a new AI model release and speculates whether Qwen 3.8 27B can compete with Muse Glimmer 30B.

0 favorites 0 likes

@heyshrutimishra: Hy3 hit #1 on OpenRouter in its first week. 68x the API calls of its predecessor. 295B total parameters, 21B active. On…

X AI KOLs Following · 17h ago Cached

Hy3 from Tencent Hunyuan hit #1 on OpenRouter in its first week, with 295B total parameters and 21B active, and is being offered free through WorkBuddy until August 31, 2026.

0 favorites 0 likes

Meta Open-Sources Muse Glimmer 30B Agent Model as Zuckerberg Pushes Personal-Superintelligence Vision

Reddit r/ArtificialInteligence · 17h ago

Meta open-sourced Muse Glimmer, a 30B-parameter multimodal agentic model under Apache 2.0, optimized for local tool use and coding with 4-bit quantization fitting under 20GB for consumer GPUs. Zuckerberg also promised open-weight Muse Spark 1.2 and a $1B community fund for data-center regions.

0 favorites 0 likes

Meta's new open-weight model targets local agentic AI

Hacker News Top · 17h ago Cached

Meta announces a new open-weight model aimed at local agentic AI, with Mark Zuckerberg sharing his essay on Meta's AI philosophy.

0 favorites 0 likes

unsloth/Muse-Glimmer-30B-GGUF · Hugging Face

Reddit r/LocalLLaMA · 18h ago Cached

Unsloth releases a GGUF-quantized version of Meta's Muse Glimmer 30B model, designed for local agentic tasks with multimodal input, tool use, and multi-step reasoning.

0 favorites 0 likes

Meta will open source their Muse Spark 1.2 and Muse Glimmer 30B

Reddit r/artificial · 18h ago

Meta announced it will open source its Muse Spark 1.2 and Muse Glimmer 30B models, described as the biggest open weights since Llama 4 and 3.

0 favorites 0 likes

Meta will soon release the weights for Muse Spark 1.2, their latest foundation model.

Reddit r/singularity · 18h ago

Meta is preparing to release the weights for Muse Spark 1.2, its latest foundation model, which will be available for broad use.

0 favorites 0 likes

@TheAhmadOsman: New opensource model from Meta Muse Glimmer 30B Apache-2.0

X AI KOLs Timeline · 18h ago Cached

Meta announces a new open-source model, Muse Glimmer 30B, released under the Apache-2.0 license.

0 favorites 0 likes

Introducing Muse Glimmer: an open-weight model optimized for always-on local agent workflows

Reddit r/LocalLLaMA · 18h ago

Meta releases Muse Glimmer, a 30B open-weight multimodal model optimized for local agent workflows, with permissive Apache 2.0 licensing, 4-bit quantization support, speculative decoding, and broad ecosystem integrations.

0 favorites 0 likes

Meta releases new on-device optimized open source model

Reddit r/singularity · 18h ago

Meta announces a new open-source model optimized for on-device deployment, aiming to bring efficient AI inference to edge devices.

0 favorites 0 likes

Meta Muse Glimmer – open weights 30B local coding model

Hacker News Top · 18h ago Cached

Meta introduces Muse Glimmer, a permissively licensed 30B-parameter model optimized for local agent workflows, coding, and tool use, with weights released on Hugging Face.

0 favorites 0 likes
← Previous
Next →
← Back to home

Submit Feedback