model-update

Tag

Cards List
#model-update

If you are wondering why Ornith 1.5 35B A3B with MTP is so slow, this is why

Reddit r/LocalLLaMA ↗ · 2026-08-20 Cached

The Ornith 1.5 35B A3B model's MTP tensors appear to be uninitialized, causing poor speculative decoding performance, and grafting the trained head from Qwen3.6-35B-A3B improves speed by 29%.

0 favorites 0 likes
#model-update

A hunch: Qwen3.8-27B's general knowledge got pruned (good, if true)

Reddit r/LocalLLaMA ↗ · 2026-08-14

The author shares a hunch that Qwen3.8-27B has pruned general knowledge to improve coding and agentic skills, based on reduced knowledge of a specific German town compared to earlier Qwen models.

0 favorites 0 likes
#model-update

@charliermarsh: The chart crimes will continue until morale improves

X AI KOLs Following ↗ · 2026-08-13 Cached

SpaceXAI announces Grok 4.6, claiming it delivers frontier intelligence and is a significant improvement over Grok 4.5 at the same price.

0 favorites 0 likes
#model-update

@leerob: We’re making great progress on improving the writing quality and design taste of Grok. Still a lot more to do, but I’m …

X AI KOLs Timeline ↗ · 2026-08-08 Cached

xAI's Grok is making rapid progress on writing quality and design taste, with Grok 4.6 expected soon.

0 favorites 0 likes
#model-update

@gdb: ChatGPT team is shipping:

X AI KOLs Following ↗ · 2026-08-08 Cached

Greg Brockman announces ChatGPT updates, including rich formatting in the web composer and a new GPT-5.6 model for paid users.

0 favorites 0 likes
#model-update

@VraserX: Seedance 2.5 is finally here. I put it head-to-head against Seedance 2.0 using the exact same prompts across four compl…

X AI KOLs Following ↗ · 2026-08-07 Cached

A user shares a head-to-head comparison of Seedance 2.5 versus Seedance 2.0 across four cinematic genres, noting significant differences in output quality.

0 favorites 0 likes
#model-update

@claudeai: We’re updating Claude Fable 5’s biology safeguards to reduce false positives. In our testing, this update reduced biolo…

X AI KOLs Timeline ↗ · 2026-08-07 Cached

Anthropic is updating Claude Fable 5's biology safeguards to reduce false positives, cutting biology-related fallbacks by about 85% while still restricting dual-use capabilities like virology and molecular design.

0 favorites 0 likes
#model-update

@0x0SojalSec: GPT Astra update that’s been circulating: - Largest OpenAI pretraining run since GPT-4.5 - Internal codename: Mewfour -…

X AI KOLs Timeline ↗ · 2026-08-06 Cached

A circulating update claims OpenAI's GPT Astra is its largest pretraining run since GPT-4.5, with internal codename Mewfour, already used by employees, and possibly releasing as soon as next week.

0 favorites 0 likes
#model-update

@OpenAI: In addition to the upgrade in intelligence with GPT-5.6 Luna, Free and Go users can now use the “Think” button for more…

X AI KOLs ↗ · 2026-08-06

OpenAI announces GPT-5.6 Luna with an intelligence upgrade, and now Free and Go users can use the 'Think' button for more reasoning on harder questions.

0 favorites 0 likes
#model-update

ChatGPT brings unlimited text chats to free users

TechCrunch AI ↗ · 2026-08-06 Cached

OpenAI removes text chat limits for free ChatGPT users, introducing the GPT-5.6 Luna model as the default and a new Think button, while Plus/Pro users get an upgraded GPT-5.6 Sol with a thinking slider.

0 favorites 0 likes
#model-update

@kyutai_labs: Our audio-to-MIDI model, MuScriptor, now also detects tempo! You can directly drag-and-drop the MIDI into a DAW and it …

X AI KOLs Timeline ↗ · 2026-08-06 Cached

Kyutai Labs announces that their audio-to-MIDI model MuScriptor now detects tempo, allowing direct drag-and-drop of MIDI into a DAW without manual tempo matching.

0 favorites 0 likes
#model-update

@higgsfield: Unlimited Seedance for 11 days. Seedance 2.5 is coming soon to Higgsfield. Start with 4 days of Seedance 2.0 in 4K, the…

X AI KOLs Following ↗ · 2026-08-03 Cached

Higgsfield announces unlimited Seedance usage for 11 days, with 4 days of Seedance 2.0 in 4K and 7 days of any other Seedance model, ahead of the upcoming Seedance 2.5 release.

0 favorites 0 likes
#model-update

@FinanceYF5: DeepSeek V4 Flash High reached 7th on the overall leaderboard in the frontend code arena with a score of 1586, ranking 3rd among open-source models. Consumer Products 4th, Reference Design, Data Analysis, and Games all 6th. Up 154 points from Flash High Preview, even higher than V4 Pro P…

X AI KOLs Following ↗ · 2026-08-03 Cached

DeepSeek V4 Flash High ranks 7th on the overall leaderboard in the frontend code arena with a score of 1586, 3rd among open-source models, a significant improvement over the preview version.

0 favorites 0 likes
#model-update

DeepSeek-V4-Flash-0731: surpasses Fable-5, Sol & Kimi-K3 on Chess Benchmark

Reddit r/LocalLLaMA ↗ · 2026-08-02

DeepSeek released V4-Flash-0731, an AI model that surpasses Fable-5, Sol, and Kimi-K3 on a chess benchmark.

0 favorites 0 likes
#model-update

Nova versão do DS v4 flash 0731

Reddit r/openclaw ↗ · 2026-08-01

Nova versão do DS v4 flash 0731 promete grande melhoria por preço baixo, com desempenho superior ao GLM 5.1 em testes caseiros, embora haja desconfiança sobre os benchmarks.

0 favorites 0 likes
#model-update

@cline: DeepSeek silently updated their changelog with a new V4-Flash upgrade 1 hour ago. Their new Terminal-Bench score is 82.…

X AI KOLs Following ↗ · 2026-07-31 Cached

DeepSeek quietly updated its changelog with a V4-Flash upgrade, boosting its Terminal-Bench score to 82.7, a +25.8 leap from the April preview. It is currently API-only, with open weights coming soon.

0 favorites 0 likes
#model-update

DeepSeek-V4-Flash has been updated, "The official release of DeepSeek-V4-Pro will follow soon"

Reddit r/LocalLLaMA ↗ · 2026-07-31

DeepSeek 宣布更新了 DeepSeek-V4-Flash,并预告 DeepSeek-V4-Pro 的正式发布将很快到来。

0 favorites 0 likes
#model-update

@vasuman: This is kinda like Citadel vs Leopold Aschenbrenner but for vibe coders

X AI KOLs Following ↗ · 2026-07-30 Cached

Sam Altman announces major price cuts for GPT-5.6 models: an 80% drop for Luna, a 20% drop for Terra, and a new Fast mode for Sol in the API.

0 favorites 0 likes
#model-update

Introducing Pangram 4 (2 minute read)

TLDR AI ↗ · 2026-07-30 Cached

Pangram Labs announces Pangram 4, its most powerful AI detector yet, with significantly reduced false positive and false negative rates, robust detection across frontier models, and new image scan features.

0 favorites 0 likes
#model-update

Kimi K3-256k

Hacker News Top ↗ · 2026-07-29 Cached

Kimi Code releases Kimi K3-256k, a 256k-context version of its flagship K3 coding model, offering reduced quota consumption while maintaining similar performance for most tasks.

0 favorites 0 likes
← Previous
Next →
← Back to home

Submit Feedback