
AI keeps cracking unsolved math problems, and mathematicians have mixed feelings
Quick Answer
AI models like OpenAI's Astra and Epoch AI's FrontierMath are revolutionizing mathematics by solving significant problems and generating proofs, but concerns about the potential loss of mathematical culture persist.
Quick Take
While AI excels in certain areas, major challenges like the Millennium Prize Problems remain unsolved, indicating a mixed future for human mathematicians.
Key Points
- Epoch AI's FrontierMath benchmark highlights AI's recent success in solving significant math problems.
- OpenAI's Astra model introduced ten solutions of varying difficulty for unsolved math questions.
- Concerns arise over AI's impact on mathematical culture and expertise development.
- AI struggles with major challenges like the Millennium Prize Problems, still unsolved.
- Mathematicians like Terence Tao see parallels between current AI advancements and historical foundational crises.
DeepSignal Analysis
What happened
AI models like OpenAI's Astra and Epoch AI's FrontierMath are making strides in solving mathematical problems, including some previously deemed significant. However, the field remains divided on the implications of AI's role, with some researchers embracing it as a tool while others express concern over the potential loss of mathematical culture and expertise.
Key evidence
- Epoch AI's FrontierMath benchmark has yet to see AI solve problems in the 'Major Advance' and 'Breakthrough' categories, indicating limitations in AI's capabilities.
- Mathematician Trefor Bazett notes that while AI has made progress, many major problems, including the Millennium Prize Problems, remain unsolved for both AI and humans.
- Abhishek Saha, a math professor, suggests that AI models are comparable to diligent PhD students, indicating a shift in how mathematicians might approach their work.
Why it matters
The integration of AI into mathematics could redefine the discipline, enhancing productivity but also raising concerns about the depth of human understanding and expertise. As AI continues to advance, the balance between leveraging technology and preserving mathematical culture will be crucial for the future of the field.
Source Excerpt
OpenAI's refutation of the Unit Distance Conjecture has sparked a wave of AI-assisted advances in mathematics. Fields Medal winner Timothy Gowers says GPT 5. 6 Pro solved two problems he had spent considerable time working on, each on its first attempt. He warns of the "possible destruction of mathematical culture" if mathematicians stop building the expertise needed to understand such results. Others simply see AI as a productivity tool.
Want this in your inbox every morning?
Daily brief at your local 8am — bilingual EN/中文, free.
More from The Decoder
See more →
An AI model programmed nonstop for 19 days on a single MirrorCode task that cost $2,600 to run
Epoch AI's MirrorCode benchmark reveals Claude Opus 4.7 as the leader with a 56% solve rate, reconstructing a 16,000-line toolkit in 14 hours. Despite this, all models tested struggle with the most complex tasks, highlighting limitations in current AI capabilities. The single task consumed $2,600 over 19 days, raising questions about cost-effectiveness in AI development.

