Mythos can improve speed of training code 52x (compared to human 4x at 4-8hrs)
Summary
Anthropic's Mythos system achieved a 52x speedup in optimizing training code compared to a human's 4x speedup over 4-8 hours on the same task, with the caveat that absolute multiples depend heavily on starting code quality. The like-for-like comparison shows roughly 3x–52x improvement across models over the past year.
Similar Articles
@AnthropicAI: Each time we release a model, we run the same test: give it code that trains a small AI model, ask the new model to spe…
Anthropic shares internal benchmark results showing dramatic AI coding improvement: while Claude Opus 4 averaged ~3x speedup on an ML code optimization task in May 2024, the new Mythos Preview model achieved ~52x speedup this April, compared to 4-8 hours for a skilled human to reach 4x.
Will It Mythos?
The author tests whether other AI models can match Mythos's exceptional ability to find security vulnerabilities, creating a benchmark of bugs found by Mythos and testing models like Opus. Initial results suggest Mythos may be uniquely powerful.
New Mythos checkpoint shows continued improvement: “On a 32-step corporate network attack we estimate takes a human expert ~20 hours, this checkpoint completes the full attack in 6 /10 attempts.”
Mythos releases a new checkpoint that can complete a 32-step corporate network attack in 6 out of 10 attempts, compared to ~20 hours for a human expert.
@liu8in: Been testing Mythos for 2 hours - best code-to-motion LLM by far Claude Code + Fable 5 one-shotted this @HyperFrames_ l…
Claude Fable 5, a Mythos-class model, is announced with capabilities exceeding all previous generally available models, tested in code-to-motion tasks.
Mythos was not trained on 'hacking'. Other Ai labs also will reach Mythos-level capabilities in the future
The article clarifies that the AI model Mythos was not trained on hacking, and predicts that other AI labs will eventually achieve similar capabilities.