Models

Cards List

Updated benchmark: Deepseek V4 Flash on SlopCodeBench (local)

Reddit r/LocalLLaMA · 2h ago

A user shares updated benchmark results for DeepSeek V4 Flash on SlopCodeBench using local quants (antirez imatrix quant) with the pi harness, showing improved performance over previous runs but still slower than the hosted API.

0 favorites 0 likes

LFM 2.6B is a lot of fun.

Reddit r/LocalLLaMA · 4h ago

The author shares hands-on experience with LFM 2.6B, a small model designed for phones, praising its speed and usefulness for quick tasks like summarization and autocomplete, though it has a 128k context limit.

0 favorites 0 likes

I Ran a Full LLM Model on an ESP32 Dev kit V1 (81KB Mem Usage)

Reddit r/ArtificialInteligence · 4h ago

A developer successfully ran a 5.2 million parameter MoE LLM quantized to INT4 on an ESP32 Dev Kit V1 using only 81KB of SRAM by streaming experts from flash, achieving about 5 tokens per second.

0 favorites 0 likes

@rauchg: Grok Imagine Image 2.0 on Vercel AI Gateway Excellent model, #2 already on http://Arena.ai

X AI KOLs Timeline · 6h ago Cached

Guillermo Rauch highlights Grok Imagine Image 2.0, now available on Vercel AI Gateway and ranking #2 on Arena.ai's leaderboard. Vercel offers access via AI CLI and a live playground.

0 favorites 0 likes

Kimi K3 (Unsloth) IQ2-XXS from 711GB down to 478GB!!! Only Multi-language was removed to trim the size

Reddit r/LocalLLaMA · 9h ago

A trimmed English-only GGUF version of Kimi K3 (IQ2-XXS) reduces model size from 711GB to 478GB by removing multi-language components, with early tests suggesting it may match or outperform the standard 2-bit version on coding tasks.

0 favorites 0 likes

@interjc: @grok draw a picture of LeBron James leading the US men's soccer team to win the World Cup and receiving the trophy from Trump

X AI KOLs Following · 20h ago Cached

Grok announces Imagine Image 2.0, a next-generation image model with precision editing, crisp text rendering, improved factuality, and real-world usefulness.

0 favorites 0 likes

DeepMind’s hurricane breakthrough has surprised weather scientists

Ars Technica · 22h ago Cached

DeepMind's hurricane AI model gives forecasters an extra day of warning and is being open-sourced as WeatherNext models, though researchers don't fully understand how it works.

0 favorites 0 likes

@jefffhj: Try Imagine Image 2.0!

X AI KOLs Following · yesterday Cached

Grok announces Imagine Image 2.0, a next-generation image model with precision editing, crisp text rendering, and improved factuality for real-world use.

0 favorites 0 likes

@rohanpaul_ai: Grok Imagine Image 2.0 (Low) from @SpaceXAI jumped 12 places to #2, beating its own older quality model. The previous i…

X AI KOLs Following · yesterday Cached

Grok Imagine Image 2.0 (Low) from xAI jumped to #2 in the Text-to-Image Arena, beating its own older quality model and showing significant improvement.

0 favorites 0 likes

Anyone else amped up over Qwen 3.8?

Reddit r/LocalLLaMA · yesterday

The author shares excitement for the upcoming Qwen 3.8 model, highlighting their experience with Qwen 3.6 27B for local LLM use, and discusses the potential of self-hosted AI to replace subscription-based frontier models.

0 favorites 0 likes

@rohanpaul_ai: Meta's Muse Spark 1.2 reaches #4 while moving Text Arena's cost-quality Pareto frontier upward. you surrender 9 points …

X AI KOLs Following · yesterday Cached

Meta's Muse Spark 1.2 reaches #4 in the Text Arena, moving the cost-quality Pareto frontier upward with a ~91% price cut while sacrificing only 9 points relative to the top model.

0 favorites 0 likes

@ayushrajnc1z: This feels less like another Al video update and more like the start of a proper production workflow. Thirty-second vid…

X AI KOLs Timeline · yesterday Cached

Seedance 2.5 introduces 30-second video generation, multimodal references, multilingual creation, and targeted editing, with API access coming soon on BytePlus for developers and enterprises.

0 favorites 0 likes

@TheAhmadOsman: Some numbers from running DeepSeek V4 Flash 0731 on a DGX Station

X AI KOLs Timeline · yesterday Cached

Ahmad Osman shares performance numbers from running DeepSeek V4 Flash 0731 on an NVIDIA DGX Station.

0 favorites 0 likes

The best AI Model in Africa and the middle east

Reddit r/artificial · yesterday

TokenAI, an Egyptian startup, announces Early Access for Horus Cyper Nano 1.0 BETA, a specialized cybersecurity model for offensive security and red teaming research, with open weights planned for September 2026.

0 favorites 0 likes

@rohanpaul_ai: Alibaba released Qwen3.8-Max, a 2.4 trillion parameter model that activates only about 95 bn parameters per token. A th…

X AI KOLs Timeline · yesterday Cached

Alibaba released Qwen3.8-Max, a 2.4 trillion-parameter sparse MoE model with 95B active parameters per token, 1M token context, and strong agentic and benchmark results, including autonomously coding for days, circuit design, and outperforming rivals on Terminal Bench and PaperBench.

0 favorites 0 likes

@gabriel1: so exciting, can't wait to try this beast

X AI KOLs Following · yesterday Cached

Sam Altman announces that 'astra' is a powerful AI model being prepared for general availability, with additional safety time needed due to its cyber capabilities.

0 favorites 0 likes

@sama: astra is a powerful model and we are working to make it generally available. we do not think it is a good strategy to k…

X AI KOLs Timeline · yesterday Cached

Sam Altman announces that the Astra model is powerful and OpenAI is working to make it generally available, while taking extra time to ensure safety given its cyber capabilities.

0 favorites 0 likes

@gdb: GPT-5.6 Sol for cybersafety:

X AI KOLs Following · yesterday Cached

Greg Brockman shares a testimonial from Grigori Karapetyan praising GPT-5.6 Sol (Codex) for enabling speedy cybersafety investigations and responses.

0 favorites 0 likes

Moonlight & Mayhem (Raccoon Heist by Codex + GPT-5.6 Sol Ultra)

Simon Willison's Blog · yesterday Cached

Simon Willison tests GPT-5.6 Sol Ultra via Codex Desktop by asking it to recreate a 'Raccoon Heist' game from a four-year-old prompt, resulting in a much better game than Claude Fable 5's version, though it had a bug with oversized eyeballs.

0 favorites 0 likes

@gdb: Evaluations of our next major model, Astra, indicate significant capability advancements in agentic coding and cybersec…

X AI KOLs Following · yesterday Cached

OpenAI's next major model, Astra, shows significant capability gains in agentic coding and cybersecurity, and is being treated as a 'critical' model for cybersecurity under the Preparedness Framework, with additional controls planned.

0 favorites 0 likes
Next →
← Back to home

Submit Feedback