Tag
MiMo v2.5 is praised for its impressive token generation speed in OpenCode, suggesting it's an underrated model update.
Elon Musk states that Grok is making progress.
OpenAI has made GPT-Realtime-2.1-mini available via its API, providing a cost-efficient option for real-time applications.
Ahmad Osman highlights improvements in Tencent Hy3 over its preview version and compares it to GLM 5.2, which is twice its size, while sharing a prediction about running similar intelligence on an RTX 5090.
Openai is reportedly preparing GPT-5.6 for release next week, with three tiers (Sol, Terra, Luna) and a new reasoning-effort control slider in the Codex app, potentially challenging Anthropic's Fable 5.
Meta is reportedly preparing an update to its flagship AI model Muse Spark.
Third-party tests show that after returning, the Fable 5 model's performance dropped significantly on the BridgeBench benchmark. Debugging score fell from 86.2 to 25.9, Refactoring from 73.6 to 38.4, Hallucination from 75.9 to 61.7. It is speculated that the new safety guardrails caused many tasks to be handed off to Opus 4.8.
Universal-3.5 Pro improves native code switching, diarization, and adds more languages, enhancing speech recognition capabilities.
Anthropic is redeploying Claude Fable 5 globally with updated classifiers to block cybersecurity tasks, though routine coding may trigger fallback to Opus.
Claude Sonnet 5 has appeared in the latest Claude Code CLI build, suggesting a new model release or update from Anthropic.
OpenAI announced GPT-5.5 Instant as its most used model, but critics argue the usage numbers are inflated because free users are forced to use it.
Sam Altman announced an update to the 5.5 instant model used in ChatGPT, saying he likes its vibes.
OpenAI has released an updated version of GPT-5.5 Instant that improves its ability to understand intent, handle complex constraints, and provide better recommendations.
GLM 5.2 appears to be a strong model update, but its launch is controversially conflating two different benchmark metric sets.
OpenAI announces significant improvements in health-related responses within ChatGPT using GPT-5.5 Instant, achieving accuracy comparable to frontier models and reducing factuality issues by 71% through physician-led evaluations.
GLM5.2 reintroduces a critic component for fine-grained variance reduction, suggesting that group-based methods are ineffective for long horizons. The author believes OpenAI and Anthropic already use value models.
User praises GLM 5.2 for being reliable and smart, but points out that lack of compute power leads to instability.
A user observes that the Kimi K2.6 model's chain-of-thought has become shorter and more concise, improving coding performance in Kimi Code, and expresses hope for continued open-source competition with upcoming GLM 5.2 and Fable 5.
A user shares their initial experience with Claude Fable 5, noting it feels incrementally smarter than previous versions but not revolutionary, and asks the community for their thoughts.
Google's Gemma 4 31B IT model now has a chat template fix that preserves thinking and improves null handling, reasoning preservation, and input validation.