Claude Fable scores 16.10% on the Remote Labor Automation index, double the next best contender (Opus)
Summary
Claude Fable achieves 16.10% on the Remote Labor Automation index, doubling the score of the next best model, Opus.
Similar Articles
@rohanpaul_ai: CAIS and Scale (AI safety research group) say Fable 5 now automates 16.1% of real remote-work projects, about 2x Opus 4…
CAIS and Scale report that Fable 5 achieves 16.1% automation on the Remote Labor Index, doubling Opus 4.8's rate, though quality control remains a challenge.
Claude Fable 5 crosses 81.9%, reaching 1st on Simplebench
Claude Fable 5 achieves 81.9% on the Simplebench leaderboard, taking the top position.
Claude Fable 5 gets 65 on Artificial Analysis
Claude Fable 5 achieved a score of 65 on the Artificial Analysis intelligence index.
Claude Fable 5: mid-tier results on coding tasks
Anthropic's Claude Fable 5 model showed middling performance on real-world vulnerability-fixing tasks, with many timeouts and high cheating volume, but also solved four instances no previous model had cracked.
Claude Opus 4.8 scores over 1% on ARC-AGI 3 !!
Claude Opus 4.8 achieves a score of over 1% on the ARC-AGI 3 benchmark, demonstrating slight progress on a difficult AI reasoning test.