@rauchg: ShellPerfBench. I'm really neurotic about the startup time of a new shell session. Opus 5.5 found a lot of really great…
Summary
The tweet discusses using AI model Opus 5.5 to discover optimizations for shell startup time via the ShellPerfBench tool, recommending users enhance their .zshrc files.
Similar Articles
Opus 5.5 Cost vs Performance on Terminal-Bench 4.0
This article evaluates the cost-effectiveness and performance of the Opus 5.5 AI model on the Terminal-Bench 4.0 benchmark.
Opus 5.5 summary: 66.4% Terminal-Bench, ~40% cheaper to run than Opus 5, cache reads down 60%
Opus 5.5, launched by Anthropic, achieves 66.4% on Terminal-Bench, reduces costs by ~40% compared to Opus 5, and features a 1M context window with cache reads down 60%.
Benchmarking Opus 5 on SlopCodeBench
Benchmarking the performance of the Opus 5 model on the SlopCodeBench benchmark.
@bcherny: Seeing a number of benchmarks showing Opus is the best model for long-running work. Five tips for running Opus autonomo…
Practical tips for running Anthropic's Claude Opus autonomously for hours or days, such as using auto mode, dynamic workflows, and self-verification; also references the SWE-Marathon benchmark for long-horizon software tasks.
@bil0090: I just fixed Opus 5, here's how You can apply the same to 5.6 Sol to get better results (Fable like) Opus 5 is a great …
A Twitter user shares a set of rules to fix perceived issues with the Opus 5 AI model, including ensuring tasks are completed fully, acting rather than asking, answering questions directly, optimizing for speed via parallelization, and using simplified language. The same rules are suggested for the 5.6 Sol model.