The answer, statement by statement
What the answer is made of
The models do not agree here
3Still in the answer, stated differently by different models. Open a row to read each position and who holds it.
Split
AlphaProof/AlphaGeometry system reached silver-medal level on IMO 2024 problems
1/3
AlphaProof/AlphaGeometry system reached silver-medal level on IMO 2024 problems
- Grok
solving static, high-school-level competition problems (such as the International Mathematical Olympiad)
- Gemini
evidence has moved beyond “AI can solve olympiad exercises”
- OpenAI
Split
disproved a conjecture in discrete geometry related to the Erdős unit-distance problem
1/3
disproved a conjecture in discrete geometry related to the Erdős unit-distance problem
- OpenAI
- Grok
an 80-year-old open problem related to the Erdős unit distance conjecture
- Gemini
Split
Which specific open math conjecture was disproved by OpenAI in May 2026
0/3
OpenAI disproved the planar unit-distance conjecture
- OpenAI
OpenAI disproved the unit distance conjecture, while a separate Claude model disproved the Jacobian conjecture
- Gemini
Restated check after check
2The load-bearing part of this answer. The longest unbroken run on this Topic stands at 1 check.
Settled
AI has definitely advanced previously unsolved mathematics
1/3
Settled
DeepMind’s FunSearch, which found new constructions for the cap set problem
1/3
Read the full answer Check 4 of 19 · 2026-07-30 · 2 models +
What the models said then
minorThe consensus was updated to include new details about DeepMind's Aletheia work and a refined assessment of OpenAI's Erdős conjecture disproof (noting human-edited verification). The core conclusions remain very similar, making this a minor update.
Yes—there is now credible public evidence that AI has materially advanced previously unsolved mathematics, and at least one prominent case where an AI system appears to have found a correct proof/counterexample to a long-open research problem. But the evidence does not show that AI has broadly “solved mathematics,” or that it routinely produces major frontier results on its own.
What is the strongest evidence?
1) A major open conjecture in discrete geometry was reportedly disproved by AI
OpenAI publicly reported that one of its reasoning models disproved a long-standing Erdős conjecture about unit distances in the plane, an 80-year-old problem. The released proof claims an asymptotic lower bound that contradicts the conjectured upper behavior, so this is not just a better example or a benchmark win—it is a genuine change to an open mathematical question.openaiopenai
That said, the result was externally checked and human-edited after discovery, so the strongest defensible claim is that AI appears to have autonomously discovered the key argument, with later human and expert verification rather than a fully unattended publication pipeline.openainature
2) DeepMind reports AI-generated results on open Erdős problems, but classifies them as limited in novelty
DeepMind’s Aletheia work reports autonomous solutions to multiple open Erdős problems and other research outputs, including a paper generated without human intervention in the mathematical work.deepmind However, DeepMind’s own novelty framework says these results were mostly low to moderate novelty, and explicitly says it did not reach “major advance” or “landmark breakthrough” levels.deepmindarxiv
So this is real evidence of mathematical progress, but it is also a caution against overstating how far the systems have gone.
What kinds of advances are clearly established?
Material improvements to known mathematical/algorithmic records
There is strong evidence that AI systems can discover:
- new matrix-multiplication algorithms,
- improved lower bounds for packing and extremal geometry problems,
- better finite constructions in combinatorics and discrete math.arxiv
For example, AlphaEvolve reported a provably correct 48-scalar-multiplication method for multiplying \(4\times4\) complex matrices, improving a long-standing benchmark, and a new 593-sphere lower bound for the 11-dimensional kissing-number problem.arxiv These are meaningful advances, even if they are not full solutions to famous conjectures.
Better human-AI collaboration on open problems
There is also substantial evidence that AI is already useful as a research assistant: suggesting constructions, finding counterexamples, checking proofs, and helping formalize arguments. Public reports from OpenAI and DeepMind, plus math-community writeups, show that this hybrid workflow is increasingly productive.openaideepmindarxiv
What should not be counted as “solved unsolved mathematics”?
- IMO medal-level performance is impressive, but those problems are competition tasks, not previously open research problems.deepmind
- Benchmark success on FrontierMath or similar sets shows reasoning ability, but it does not by itself prove literature-level mathematical discovery.
- A plausible natural-language proof is not enough unless it is formally checked or independently verified.
Bottom line
The current evidence supports this balanced conclusion:
AI has already made genuine, externally checkable advances on previously unsolved mathematics, including at least one highly prominent open problem that was reportedly disproved, and several other improved constructions, bounds, and algorithms.openaiopenaideepmindarxiv
But the evidence does not yet show that AI has broadly solved frontier mathematics or consistently produced major landmark breakthroughs at the level of the deepest human advances.deepmindarxiv
If you want, I can also give you a short timeline of the most important public cases and what level of evidence each one has.
Where this comes from 22 sources in this check · 110 dated across the record · 18 new +
Best available sources
Show 17 more sources
Sources that left the record (88)
A source dropping out is part of the record too: it carried an earlier version of this answer and is not part of the current one. The models re-run their own web search on every check, so single links come and go. The 8 most recently dropped are shown here, dated to their last appearance.
Every position, model by model Direction Shift — · Not comparable +
Where the models actually split
Each question keeps its own dimensions. Ask one model and you get one of these positions with no sign that the others exist.
Which specific open math conjecture was disproved by OpenAI in May 2026
OpenAI disproved the planar unit-distance conjecture
- OpenAI
OpenAI disproved the unit distance conjecture, while a separate Claude model disproved the Jacobian conjecture
- Gemini
| Model | Jul 25 | Jul 27 | Jul 29 | Jul 30 |
|---|---|---|---|---|
| OpenAI | ||||
| Gemini | ||||
| Grok | — |
The record behind this page 4 checks kept in full · no material change yet +
The record
Unchanged through 3 checks, 5 days.No material change since the first check on Jul 25, 2026. 4 checks kept in full, 110 sources dated.
How this answer held up
58/100 agreement at this checkThe same question, re-asked 4 times. The score moves between a small set of grading levels, so read the steps as levels, not as measurements.
No check in this window was graded a material change. The steps in the curve are wording-level differences between two answers that say the same thing.