frontier-math

Tag

Cards List
#frontier-math

GPT-6.1 sol (max) scores 100% in frontier math 4

Reddit r/singularity ↗ · 7h ago

OpenAI 的 GPT-6.1 sol (max) 在 Epoch AI 的 FrontierMath Tier 4 v2 基准测试中取得 100% 的成绩,登上该基准的榜首。

0 favorites 0 likes
#frontier-math

FrontierMath’s First “Major Advance” Problem Has Been Solved

Reddit r/singularity ↗ · 2026-09-18

FrontierMath has successfully solved its first 'Major Advance' problem, marking a significant milestone in mathematical research and AI collaboration.

0 favorites 0 likes
#frontier-math

@gdb: congrats on the amazing result!

X AI KOLs Following ↗ · 2026-07-28 Cached

David Turturean solved a 40-year-old open problem in p-adic Galois theory using voice input, in collaboration with problem proposer David Roe, under EpochAI Research's FrontierMath initiative.

0 favorites 0 likes
#frontier-math

Claude Fable 5's FrontierMath scores

Reddit r/singularity ↗ · 2026-06-12

Epoch AI released a v2 update to the FrontierMath benchmark, correcting errors in 42% of problems and increasing scores across all models, though rankings remained largely unchanged; Tiers 1-4 are approaching saturation.

0 favorites 0 likes
#frontier-math

GPT-5.5 was used to flag fatal errors in FrontierMath problems

Reddit r/singularity ↗ · 2026-05-12

GPT-5.5 was used by Epoch to identify fatal errors in approximately one-third of the FrontierMath benchmark problems, demonstrating the model's capability to sanity-check evaluation standards.

0 favorites 0 likes
#frontier-math

[Google DeepMind] the AI co-mathematician also achieves state of the art results on hard problemsolving benchmarks, including scoring 48% on FrontierMath Tier 4, a new high score among all AI systems evaluated.

Reddit r/singularity ↗ · 2026-05-08

Google DeepMind's AI co-mathematician achieves state-of-the-art results on hard problem-solving benchmarks, scoring 48% on FrontierMath Tier 4, the highest among all AI systems evaluated.

0 favorites 0 likes
#frontier-math

AI Co-Mathematician: Accelerating Mathematicians with Agentic AI

Hugging Face Daily Papers ↗ · 2026-05-07 Cached

This paper introduces the AI Co-Mathematician, a workbench that uses agentic AI to support mathematicians in open-ended research tasks like ideation and theorem proving. Early tests show the system achieving state-of-the-art results on hard problem-solving benchmarks, including a 48% score on FrontierMath Tier 4.

0 favorites 0 likes
← Back to home

Submit Feedback