Tag
Laguna s2.1 launched with impressive benchmarks but suffered from looping and tool usage issues; recent updates may have fixed them, and users are asking if it is now stable.
Muse Spark 1.1 has improved its Text Arena ranking, now close to Fable 5.
xAI announces upcoming releases of Grok 4.6 in two weeks and Grok 4.7 in four weeks.
Google has deprecated temperature, top_p, and top_k parameters starting with Gemini 3.6 Flash and 3.5 Flash-Lite models, as detailed in their developer guide.
Anthropic updated Claude's voice mode to support Opus, Sonnet, and Haiku models, and integrated with apps like Gmail, Slack, Canva, and Notion. Multilingual support in beta.
Kimi-K3 is not yet outperforming Fable, but it is showing significant progress and closing the gap.
Claude Fable 5 will be included in all Max and Team Premium plans starting July 20, with 50% of limits. Pro and Team Standard users get a one-time $100 credit and access via usage credits.
Alibaba's Qwen AI model is being updated, as hinted by a new announcement from the official Qwen account.
This thread presents research on whether new facts can be added to an LLM's weights without breaking the model, and finds that it breaks unexpectedly, making compressed KV caches and in-context learning more promising for continual learning.
After testing Kimi 3, the user found its performance far exceeds the previous generation Kimi 2.7, calling it a 'crushing' victory, and mentioned that a demo is already available on the official website.
A developer demonstrates adding vision capabilities to the GLM language model, showcasing a significant multimodal extension.
Google updated Gemma 4's chat templates with major fixes to tool calling, reduced laziness, enabled Flash Attention 4 on Hopper GPUs, and released an interactive vision guide. The updates are available on Hugging Face.
Google Gemma is rolling out significant improvements to Gemma 4, driven by community feedback and contributions, as detailed in a thread.
The tweet praises SAM3.1 for improved segmentation with faint borders and better tail coverage, leading to more accurate length measurements.
Kimi K3 model is nearing completion, suggesting an upcoming release.
Sam Altman announces that GPT-5.6 sol is half the price and roughly twice as token efficient as fable for many tasks, with plans to deliver at one-quarter of the price.
Devin Fusion now integrates Fable 5, which surprisingly offers lower cost per task than Opus 4.8 due to efficiency improvements in delegation and reasoning.
Sam Altman announces that GPT 5.6 Sol will remain available in ChatGPT subscriptions (Go, Plus, Pro) until a better model is shipped, providing clarity on the model's availability.
OpenAI announces GPT-5.6, a major step forward for health intelligence with stronger performance and 25x lower cost for GPT-5.6 Luna compared to GPT-5.5, making advanced models more accessible globally.
GPT-5.6 achieves a breakthrough by solving a previously unsolved problem, marking a significant advancement in AI capabilities.