olympiad

Tag

Cards List
#olympiad

Solving Is Not Drawing: A Benchmark for Diagrammatic Reasoning in Olympiad Geometry

arXiv cs.AI ↗ · 2026-08-20 Cached

This paper introduces an open-source benchmark for evaluating diagrammatic reasoning in geometry, revealing that foundation models like GPT and Claude struggle to produce faithful diagrams despite their problem-solving abilities.

0 favorites 0 likes
#olympiad

Opus 5 received a perfect score on the IMO

Reddit r/singularity ↗ · 2026-07-24

Opus 5, an AI model, achieved a perfect score on the International Mathematical Olympiad, demonstrating advanced mathematical reasoning.

0 favorites 0 likes
#olympiad

@jasongyang365: The competition paradigm is mass-producing founders. 1. Individuals with math/programming Olympiad backgrounds make up a disproportionate share of today's tech founders—founders of Hyperliquid, Cognition, Scale, Perplexity, Pika, Cartesia all come from the same circle. 2. Core…

X AI KOLs Timeline ↗ · 2026-07-14 Cached

This thread explores how math/programming Olympiad backgrounds mass-produce tech founders, pointing out that the internalized systematic problem-solving ability, belief, and peer effect from the competition paradigm are the core engine, with quantitative finance as a transfer station, but also reminds that entrepreneurship requires skills beyond problem-solving.

0 favorites 0 likes
#olympiad

@ClementDelangue: Paper of the day! https://huggingface.co/papers/2605.13301…

X AI KOLs Following ↗ · 2026-05-15 Cached

A paper introduces a unified recipe (SU-01) that combines reverse-perplexity curriculum, two-stage reinforcement learning, and test-time scaling to achieve gold-medal-level performance on IMO and IPhO problems using a 30B-A3B backbone.

0 favorites 0 likes
#olympiad

Physics-R1: An Audited Olympiad Corpus and Recipe for Visual Physics Reasoning

arXiv cs.CL ↗ · 2026-05-15 Cached

This paper audits multimodal physics evaluation pipelines, revealing issues like train-eval contamination, translation drift, and MCQ saturation. It releases new datasets (PhysCorp-A, PhysR1Corp, PhysOlym-A) and a training recipe (Physics-R1) that significantly improves performance on held-out olympiad problems.

0 favorites 0 likes
#olympiad

Achieving Gold-Medal-Level Olympiad Reasoning via Simple and Unified Scaling

Hugging Face Daily Papers ↗ · 2026-05-13 Cached

A paper presenting SU-01, a 30B-A3B reasoning model that achieves gold-medal-level performance on IMO and IPhO problems via reverse-perplexity curriculum, two-stage reinforcement learning, and test-time scaling.

0 favorites 0 likes
#olympiad

MIT scientists build the world’s largest collection of Olympiad-level math problems, and open it to everyone

MIT News — Artificial Intelligence ↗ · 2026-04-24 Cached

MIT researchers, in collaboration with KAUST and HUMAIN, have released MathNet, the largest open-source dataset of Olympiad-level math problems, containing over 30,000 expert-authored problems from 47 countries.

0 favorites 0 likes
#olympiad

Solving (some) formal math olympiad problems

OpenAI Blog ↗ · 2022-02-02 Cached

OpenAI achieved a new state-of-the-art 41.2% on the miniF2F formal math olympiad benchmark using a technique called 'statement curriculum learning,' which iteratively trains a neural prover on proofs of increasing difficulty. The approach builds on iterative proof search and retraining over 8 iterations to significantly outperform the previous best of 29.3%.

0 favorites 0 likes
← Back to home

Submit Feedback