Tag
The Ornith 1.5 35B A3B model's MTP tensors appear to be uninitialized, causing poor speculative decoding performance, and grafting the trained head from Qwen3.6-35B-A3B improves speed by 29%.
The author shares a hunch that Qwen3.8-27B has pruned general knowledge to improve coding and agentic skills, based on reduced knowledge of a specific German town compared to earlier Qwen models.
SpaceXAI announces Grok 4.6, claiming it delivers frontier intelligence and is a significant improvement over Grok 4.5 at the same price.
xAI's Grok is making rapid progress on writing quality and design taste, with Grok 4.6 expected soon.
Greg Brockman announces ChatGPT updates, including rich formatting in the web composer and a new GPT-5.6 model for paid users.
A user shares a head-to-head comparison of Seedance 2.5 versus Seedance 2.0 across four cinematic genres, noting significant differences in output quality.
Anthropic is updating Claude Fable 5's biology safeguards to reduce false positives, cutting biology-related fallbacks by about 85% while still restricting dual-use capabilities like virology and molecular design.
A circulating update claims OpenAI's GPT Astra is its largest pretraining run since GPT-4.5, with internal codename Mewfour, already used by employees, and possibly releasing as soon as next week.
OpenAI announces GPT-5.6 Luna with an intelligence upgrade, and now Free and Go users can use the 'Think' button for more reasoning on harder questions.
OpenAI removes text chat limits for free ChatGPT users, introducing the GPT-5.6 Luna model as the default and a new Think button, while Plus/Pro users get an upgraded GPT-5.6 Sol with a thinking slider.
Kyutai Labs announces that their audio-to-MIDI model MuScriptor now detects tempo, allowing direct drag-and-drop of MIDI into a DAW without manual tempo matching.
Higgsfield announces unlimited Seedance usage for 11 days, with 4 days of Seedance 2.0 in 4K and 7 days of any other Seedance model, ahead of the upcoming Seedance 2.5 release.
DeepSeek V4 Flash High ranks 7th on the overall leaderboard in the frontend code arena with a score of 1586, 3rd among open-source models, a significant improvement over the preview version.
DeepSeek released V4-Flash-0731, an AI model that surpasses Fable-5, Sol, and Kimi-K3 on a chess benchmark.
Nova versão do DS v4 flash 0731 promete grande melhoria por preço baixo, com desempenho superior ao GLM 5.1 em testes caseiros, embora haja desconfiança sobre os benchmarks.
DeepSeek quietly updated its changelog with a V4-Flash upgrade, boosting its Terminal-Bench score to 82.7, a +25.8 leap from the April preview. It is currently API-only, with open weights coming soon.
DeepSeek 宣布更新了 DeepSeek-V4-Flash,并预告 DeepSeek-V4-Pro 的正式发布将很快到来。
Sam Altman announces major price cuts for GPT-5.6 models: an 80% drop for Luna, a 20% drop for Terra, and a new Fast mode for Sol in the API.
Pangram Labs announces Pangram 4, its most powerful AI detector yet, with significantly reduced false positive and false negative rates, robust detection across frontier models, and new image scan features.
Kimi Code releases Kimi K3-256k, a 256k-context version of its flagship K3 coding model, offering reduced quota consumption while maintaining similar performance for most tasks.