@MaxForAI: A brutal fact: AI's mathematical abilities have already surpassed 99% of humans on this planet. @MenloVentures partner Deedy @deedydas, who invested in Anthropic and OpenRouter, conducted a test. He used the just-concluded 2026 International Mathematical Olympiad...

X AI KOLs Timeline News

Summary

AI models Fable, Sol, K3, and Axiom all achieved a perfect score of 42/42 in the 2026 International Mathematical Olympiad, solving the competition completely for the first time at low cost. Among them, Claude Fable 5 was the fastest, while GPT 5.6 Sol had the lowest cost.

A brutal fact: AI's mathematical abilities have already surpassed 99% of humans on this planet. @MenloVentures partner Deedy @deedydas, who invested in Anthropic and OpenRouter, conducted a test. He tested AI using the original problems from the just-concluded 2026 International Mathematical Olympiad (IMO). This is the world's hardest math competition for high school students, and was previously one of the most common benchmarks for evaluating AI models' mathematical abilities. He had Fable (high config), Sol (ultra-high config), K3 (max config), and Axiom participate in the competition, and they all achieved a perfect score of 42/42. As a reference, in the past 7 years of competition, a total of 4,347 human contestants participated, and only 30 achieved a perfect score of 42, a rate of 0.69%. — Claude Fable 5 solved all problems in just one attempt, and was the fastest. — GPT 5.6 Sol took one more attempt but had the lowest cost. — Kimi K3 succeeded, but took 4 more attempts and consumed a large number of tokens. Based on the number of attempts and tokens consumed, problems 3 and 6 were the hardest, followed by problem 2. Human students had 9 hours to solve these 6 problems, while Fable and Sol both completed them in under 4 hours. This is the first time the IMO has been completely solved. Last year, the best-performing models were the unreleased Gemini Deep Think and an experimental OpenAI model, both scoring 35/42. This year, three publicly available models, including one that will soon open-source its weights (K3), achieved this at a cost of only $10 to $50! AI capabilities have officially far surpassed the level of the International Mathematical Olympiad.
Original Article
View Cached Full Text

Cached at: 07/21/26, 04:46 PM

A brutal fact: AI’s math ability has already surpassed 99% of humans on this planet.

@MenloVentures partner Deedy @deedydas, who has invested in Anthropic and OpenRouter, ran a test.

He used the original problems from the just-concluded 2026 International Mathematical Olympiad (IMO) to test AI.

This is the world’s toughest high school math competition, and it has long been one of the most common benchmarks for evaluating AI models’ mathematical capability.

He put Fable (high), Sol (xhigh), K3 (max), and Axiom through the contest — all scored a perfect 42/42.

For reference, over the past 7 competitions, a total of 4,347 human participants have taken part, and only 30 achieved a perfect score of 42 — a rate of 0.69%.

  • Claude Fable 5 solved all problems in a single attempt and was the fastest.
  • GPT 5.6 Sol took one extra attempt but had the lowest cost.
  • Kimi K3 also managed it, but required 4 extra attempts and consumed a large number of tokens.

Based on the number of attempts and token consumption, problems 3 and 6 were the hardest, followed by problem 2.

Human students have 9 hours to solve these 6 problems; Fable and Sol both finished within 4 hours.

This is the first time the International Mathematical Olympiad has been completely solved.

Last year, the best-performing models were the unreleased Gemini Deep Think and an experimental OpenAI model — both scored 35/42.

This year, three publicly accessible models, including one (K3) about to open-source its weights, accomplished this at a cost of just $10 to $50!

AI’s capability has officially far surpassed the level of the International Mathematical Olympiad.

Deedy (@deedydas): The International Math Olympiad (IMO) 2026, the hardest math contest for high schoolers, just ended.

I ran Fable (high), Sol (xhigh), K3 (max) and Axiom against it and all got a perfect score of 42/42 (repo below if you want to check their solutions): — Claude Fable 5 was the

Similar Articles

@paperpaper886: Last week, I discussed the current state and future of AI4Math with a friend from the math department. He said that current AI is already powerful enough as an auxiliary tool, but there is still a long way to go for AI to achieve independent discovery.

X AI KOLs Timeline

Discussed the current state and future of AI in mathematics. Citing an example, ChatGPT 5.5 Pro autonomously solved the farthest pair problem in high-dimensional computational geometry, which had been stuck for years, demonstrating AI's potential in mathematical discovery.

@MaxForAI: Tian Yuandong @tydsh's startup team Recursive @Recursive_SI released a milestone: an automated AI research system. In this system, AI can complete the entire research loop of 'propose ideas → implement → run experiments → verify → select next experiment based on results'. Results show that with clear objectives...

X AI KOLs Timeline

The Recursive team released an automated AI research system that can autonomously complete the research loop, surpassing existing human community solutions on multiple benchmarks. For example, on NanoGPT Speedrun it compressed training time from 79.7 seconds to 77.5 seconds, and on SOL-ExecBench it improved the score to 0.754.

@FinanceYF5: This conclusion might surprise the market: In fact, OpenAI's growth rate has already surpassed Anthropic's. Ramp data shows that the quarter-over-quarter growth for Q3 so far is 82% and 76%, respectively. Why? GPT-5.6 Sol is becoming the choice of more developers; Fable…

X AI KOLs Timeline

OpenAI's growth rate exceeds Anthropic's, with Q3 quarter-over-quarter growth at 82% and 76% respectively, mainly because GPT-5.6 Sol is more popular among developers, while Fable 5's adoption is below expectations due to cost and data issues.