Is anyone else surprised Google DeepMind isn't leading these mathematical benchmarks?

Reddit r/singularity News

Summary

The author expresses surprise that Google DeepMind is not leading mathematical AI benchmarks despite its past foundational work in the area, while noting OpenAI's recent progress.

Seeing OpenAI's recent progress surprised me, but what surprised me even more is how little Google DeepMind appears in comparison. DeepMind has arguably done more foundational work in AI for mathematics than anyone else over the past few years, they got AlphaGeometry, AlphaProof, AlphaEvolve and a long history of research on reasoning and search. If you'd asked me a 1-2 years ago which lab would dominate difficult mathematical benchmarks, I probably would have said DeepMind. Is this simply because Gemini is optimized differently from OpenAI's models? Is Google DeepMind focusing on broader product capabilities rather than pushing frontier research? Or do you simply find these problems not so relevant? I'm sort of surprised that DeepMind, of all labs, doesn't seem to be leading in an area that has historically been one of its biggest strengths. More about these stats you can see here https://vibemathed.com/stats And here OpenAI's latest blog about ten problems they contributed to https://openai.com/index/ten-advances-in-mathematics/
Original Article

Similar Articles

Bench Maxing

Reddit r/ArtificialInteligence

Discusses slipping market share for OpenAI while Meta and Google gain, questioning whether high benchmark scores matter to average users and suggesting AI's true value is as a feature within existing product ecosystems.