AI research
AI is now doing real mathematics
Original, verifiable maths — not just summaries. The brief.
The answer
In 2026 AI began producing original, checker-verified maths on long-open problems.
What happened
In May 2026, AI from OpenAI and Google DeepMind produced original results on long-open Erdős problems — questions easy to state, impossible to fake, with no memorised answer to copy. OpenAI's model disproved a geometry conjecture dating to 1946. DeepMind's AlphaProof Nexus solved nine open Erdős problems plus 44 conjectures from the integer-sequence catalogue, two of which had been unsolved for 56 years. Different companies, different methods, same week.
Why it's credible
DeepMind's results are verified step-by-step in the Lean proof assistant — the AI proposes, Lean certifies, and a flawed step is simply rejected and fed back. That guards against AI's worst habit: sounding right while being wrong. OpenAI's construction was checked by human mathematicians, including Timothy Gowers, with formal peer review still to come — credible, but a softer guarantee. Same word, 'solved'; two different levels of proof.
Hassabis moved quickly to temper expectations, saying the system is 'still not AGI' even as it points toward a more practical role for AI in verified mathematical research.
OpenAI described the result as the first time a prominent open problem, central to a subfield of mathematics, has been solved autonomously by AI.
What to watch
The durable result isn't any single proof — it's the pipeline. 'AI proposes, formal checker certifies' neutralises AI's worst failure mode wherever claims can be mechanically verified. Mathematics was the hardest proving ground precisely because the bar for correctness is absolute: Lean either accepts every step or it doesn't. If that generate-then-verify template spreads to software testing, computational biology, or other formalisable fields, the May 2026 maths results will look, in hindsight, like the prototype — the moment a reliable pattern was first demonstrated — rather than the prize itself.
Frequently asked questions
Did AI actually do new maths?
What is Lean and why does it matter?
Does this make AI generally intelligent?
Sources
- An OpenAI model has disproved a central conjecture in discrete geometry — OpenAI, 20 May 2026
- Solving open problems with AlphaProof Nexus (preprint, arXiv:2605.22763) — Google DeepMind / arXiv, 21 May 2026
- OpenAI's milestone math breakthrough played to AI's strengths — Understanding AI, 22 May 2026
- Google DeepMind's AlphaProof Nexus solves 9 Erdős problems and 44 conjectures — Crypto Briefing, 26 May 2026