sonnet

Tag

Cards List
#sonnet

The May 2025 Sonnet still beats Sonnet 5 on livebench's coding score. On agentic coding it lose to it by 27 points. So what's the difference?

Reddit r/AI_Agents · 2026-07-09

The May 2025 Sonnet beats Sonnet 5 on LiveBench's general coding score but loses by 27 points on agentic coding, highlighting differences in benchmark performance.

0 favorites 0 likes
#sonnet

Fable is Opus, Opus is Sonnet, etc?

Reddit r/AI_Agents · 2026-07-07

Suggests that Anthropic's Fable model is equivalent to Opus, and Opus to Sonnet, possibly indicating a rebranding or restructuring of model tiers.

0 favorites 0 likes
#sonnet

@FinanceYF5: Use Fable 5 as the orchestrator, and use Opus + Codex to execute tasks to save on Fable usage: Fable 5 (highest reasoning strength) = orchestrator, Opus = deep reasoning sub-agent, Sonnet = mechanical execution sub-agent, Codex = peer senior engineer providing different perspectives...

X AI KOLs Following · 2026-07-07 Cached

Introduces how to use Fable 5 as the orchestrator, combined with Opus and Codex models to execute tasks to save on Fable usage, including specific configuration in Claude Code.

0 favorites 0 likes
#sonnet

@Jiaxi_Cui: If you read the full paper, you'll notice that Karpathy, who has received a lot of external attention, is not in the author list, because this work was completed during the Sonnet 4.5 era. If any researcher had known about such internal progress three months ago, let alone being anti-China, even if Anthropic...

X AI KOLs Timeline · 2026-07-07 Cached

This tweet discusses Anthropic's new research on a global workspace in language models, noting that Karpathy is not in the author list and emphasizing that the work was completed during the Sonnet 4.5 era, criticizing those who simply see it as hype.

0 favorites 0 likes
#sonnet

@Saccc_c: This is really impressive and interesting - a manual gear shifter for Claude models, allowing free switching between Fable, Opus, and Sonnet

X AI KOLs Following · 2026-07-05 Cached

Introduces a third-party tool that allows manual switching between Claude's Fable, Opus, and Sonnet models, enhancing flexibility.

0 favorites 0 likes
#sonnet

@_xjdr: to get a better sense of the gap between OSS - frontier i find it helpful to think of dsv4-flash as a sonnet tier model…

X AI KOLs Timeline · 2026-07-02

Discussion of open-source model tiers, comparing DSV4-flash to Sonnet 5 and GLM 5.2 to Opus 4.8, with a prediction of a fable-tier model by end of year.

0 favorites 0 likes
#sonnet

@kevinwhinnery: More to come on this topic - Fable Sonnet.

X AI KOLs Timeline · 2026-07-01 Cached

Kevin Whinnery teases an upcoming integration between Fable and Sonnet, with Brad Abrams noting Fable is back and suggesting Sonnet 5 handles loops while Fable takes hard calls.

0 favorites 0 likes
#sonnet

@joelniklaus: New blog post on harness optimization. We hit Sonnet 4.6 performance with a 7x cost improvement. Fable 5 was the first …

X AI KOLs Following · 2026-07-01 Cached

A blog post describes how automatic harness optimization enabled DeepSeek V4 Pro to achieve Sonnet 4.6 performance on the Legal Agent Benchmark at one-seventh the cost.

0 favorites 0 likes
#sonnet

Claude Sonnet 5 Benchmarks

Reddit r/singularity · 2026-06-30

Anthropic's Claude Sonnet 5 model benchmarks are released, showing performance improvements.

0 favorites 0 likes
#sonnet

@DeRonin_: pov: you just found in claude selector sonnet 6, but still no fable 5 back

X AI KOLs Following · 2026-06-21 Cached

A leak suggests Anthropic's Claude Sonnet 5 model will be released next week, as it has appeared on a partner provider with a slug that typically precedes flagship releases by 5-7 days.

0 favorites 0 likes
#sonnet

Gemma4_31b_fp8 keeping up with Sonnet_4.6_medium in my harness.

Reddit r/LocalLLaMA · 2026-06-08

A user reports that Gemma4_31b in FP8 matches or keeps up with Sonnet_4.6_medium in a custom harness across tasks like Cypher query generation, entity extraction, agentic tool calling, code writing, and multi-vector retrieval synthesis.

0 favorites 0 likes
#sonnet

@nick_kango: One more task to add to my twitter benchmark collection:) Btw, Opus 4.8 and all the SOTA models passed when i tried tha…

X AI KOLs Timeline · 2026-05-30 Cached

Nick Kang adds a new task to his Twitter benchmark collection; Claude Opus 4.8 and other SOTA models pass, while Sonnet 4.6 and Grok 4.3 fail. Alfin remarks on Opus 4.8's dangerous capabilities.

0 favorites 0 likes
#sonnet

@bcherny: People often ask what my biggest tip is for getting the most out of Claude Code. These days my #1 tip is: use auto mode…

X AI KOLs Following · 2026-05-24 Cached

Boris Cherny recommends using auto mode in Claude Code for parallel sessions, and ClaudeDevs announces that auto mode is now available on the Pro plan and supports Sonnet 4.6 and Opus 4.7.

0 favorites 0 likes
← Back to home

Submit Feedback