@rohanpaul_ai: Claude Sonnet 5.5 is out and it scores 70.6% on Terminal-Bench 4.0, up from Sonnet 5's 10.3%, at unchanged prices. Over…

X AI KOLs Following Models

Summary

Claude Sonnet 5.5 is released by Anthropic, offering major improvements in benchmark scores, cost efficiency, and speed compared to Sonnet 5, with a 70.6% score on Terminal-Bench 4.0.

Claude Sonnet 5.5 is out and it scores 70.6% on Terminal-Bench 4.0, up from Sonnet 5's 10.3%, at unchanged prices. Overall, 30% cost reduction per-task due to faster speeds and fewer tool calls. Keeps Sonnet 5's $2/$10 per million input/output tokens, half Opus 5.5's rates. Its savings instead come from doing less work per job, since Anthropic says fewer tokens and tool calls cut the total cost of a task by up to 30%. Output also arrives more than 30% faster than Sonnet 5's, making Sonnet 5.5 Anthropic's quickest Sonnet yet. Users can dial an effort setting, trading longer reasoning and more self-checking for a higher cost per task. The economics might count for more than the leaderboard. Sonnet 5.5 operating at Low or Medium effort is able to exceed Sonnet 5's top score at about one-tenth the cost per task. On FrontierCode, it says Sonnet 5.5 at High effort scores roughly 10 points above Sonnet 5 at the same setting while costing approximately one-fifteenth as much per task. Anthropic characterizes Sonnet 5.5 as ideal for comparatively well-defined routine work, such as software debugging, coding, document production, building presentations and spreadsheets, and designing or refining interfaces.
Original Article
View Cached Full Text

Cached at: 09/29/26, 05:42 AM

Claude Sonnet 5.5 is out and it scores 70.6% on Terminal-Bench 4.0, up from Sonnet 5’s 10.3%, at unchanged prices.

Overall, 30% cost reduction per-task due to faster speeds and fewer tool calls.

Keeps Sonnet 5’s 2/10 per million input/output tokens, half Opus 5.5’s rates.

Its savings instead come from doing less work per job, since Anthropic says fewer tokens and tool calls cut the total cost of a task by up to 30%.

Output also arrives more than 30% faster than Sonnet 5’s, making Sonnet 5.5 Anthropic’s quickest Sonnet yet.

Users can dial an effort setting, trading longer reasoning and more self-checking for a higher cost per task.

The economics might count for more than the leaderboard. Sonnet 5.5 operating at Low or Medium effort is able to exceed Sonnet 5’s top score at about one-tenth the cost per task. On FrontierCode, it says Sonnet 5.5 at High effort scores roughly 10 points above Sonnet 5 at the same setting while costing approximately one-fifteenth as much per task.

Anthropic characterizes Sonnet 5.5 as ideal for comparatively well-defined routine work, such as software debugging, coding, document production, building presentations and spreadsheets, and designing or refining interfaces.

Claude (@claudeai): Introducing Claude Sonnet 5.5, the second model in the Claude 5.5 family.

It’s a clear upgrade over Sonnet 5, runs more than 30% faster, and costs up to 30% less for most work.

Similar Articles

What's new in Claude Sonnet 5

Simon Willison's Blog

Anthropic released Claude Sonnet 5, a model with performance near Opus 4.8 at lower prices, but featuring a new tokenizer that increases token counts for English and code by ~30%, effectively raising costs.

Claude Sonnet 5 Benchmarks

Reddit r/singularity

Anthropic's Claude Sonnet 5 model benchmarks are released, showing performance improvements.