Claude-shaped science: a correct calculation still needs a worthwhile question
Summary
A commentary on an Anthropic guest essay describing BootLoops, tools in which Claude-generated calculations were technically correct but only became scientifically meaningful after domain experts redirected the questions. The piece argues that arithmetic verification and relevance/novelty judgment require separate expert reviews before AI output is called a discovery, while noting the essay's author is an Anthropic visiting researcher rather than an independent benchmark.
Similar Articles
Oct 1, 2026ScienceClaude-shaped science
In this Anthropic guest post, Prof. Matthew Schwartz describes BootLoops, an open-source toolkit of scientific software and protocols built with Claude to do exact calculations in quantitative science. By deliberately targeting "Claude-shaped" problems, the tools surfaced cross-domain connections to ecology, population genetics and other fields, which domain experts then helped steer toward scientifically meaningful questions.
Anthropic just shipped claude science, basically claude code but for research
Anthropic launched Claude Science, a new tool for research similar to their existing Claude Code for coding.
Claude solves an important mathematical problem.
Claude reportedly solved an important mathematical problem, drawing attention from Scientific American. The achievement highlights advances in AI-driven mathematical reasoning.
@AnthropicAI: New on the Science Blog: Yes, Claude can do Nine Loops. Theoretical physicists predict how particles behave using formu…
Claude AI solved a nine-loop scattering amplitude problem in theoretical physics, setting a new record and demonstrating AI's potential in complex scientific computation.
Claude Says
A blog post criticizing the reflexive use of LLMs like Claude in software work, arguing that AI-driven solutions lead to overengineered systems, that deferring to an AI signals a lack of expertise, and that prompt-driven development erodes critical thinking.