The answer, statement by statement
What the answer is made of
Restated check after check
2The load-bearing part of this answer. The longest unbroken run on this Topic stands at 1 check.
Settled
found new constructions for the cap-set problem
1/7
Settled
disproved a longstanding conjecture in discrete geometry
1/7
Entered the record recently
2First stated within the last 3 checks — the least tested part of this answer.
New
The specific nature and origin of other major open problem breakthroughs besides the Erdős conjecture.
0/7
Model B claims breakthroughs include the Jacobian conjecture via Levent Alpöge and Anthropic's Claude, and AlphaProof Nexus solving 9 out of 353 open Erdős problems.
- Gemini
Model A and Model C focus on DeepMind's Aletheia/Gemini work, FunSearch, and general theorem-proving benchmarks rather than Model B's specific novel claims.
- Grok
- OpenAI
New
or other theorem provers to generate machine-checkable proofs.
1/7
Dropped 2 statements the models stopped making · 14 one-off phrasings left out +
Stated in more than one earlier check, absent from this one. Nothing here was retracted — the models simply stopped saying it. 4 of the 11 checks listed no statements at all and count as gaps, not as absences.
Dropped
AI has materially advanced several previously unsolved mathematical problems
last Aug 01, 2026
Dropped
AlphaProof and AlphaGeometry 2 solved 4 of 6 problems at the 2024 IMO
last Aug 01, 2026
Read the full answer Check 11 of 19 · 2026-08-06 · 3 models +
What the models said then
stableThe new consensus reorganizes the sections and refines the framing of AI's mathematical capabilities, but the core conclusions regarding AI's progress on open problems, formal proofs, and limitations remain consistent.
As of mid-2026, the evidence is strong that AI has materially advanced previously unsolved mathematical problems, and in a few cases has produced genuinely new, checkable mathematical results. But it is not evidence that AI has broadly “solved mathematics,” or that it has resolved the biggest open problems of the field.
Bottom line
The clearest current evidence falls into three buckets:
- AI has disproved or materially advanced some long-open conjectures.
- AI has generated formal, machine-checkable proofs and new constructions.
- AI still relies heavily on human verification, problem selection, and interpretation.
Strongest evidence so far
1) A counterexample to a long-standing discrete-geometry conjecture
OpenAI reported that one of its reasoning models found a counterexample to the planar unit-distance conjecture, a problem associated with Erdős. The result was not just a numerical pattern: it produced an infinite family of point sets with more than the conjectured \(n^{1+o(1)}\) unit-distance behavior, and OpenAI said the proof was externally checked by mathematicians.openai A companion paper by leading mathematicians gives a human-verified version of the argument and explains the result in standard mathematical terms.openai
This is the best public example of AI-originated work that appears to have changed the status of a previously open mathematical question.
2) Semi-autonomous solutions to smaller research problems
DeepMind’s Aletheia/Gemini work reports large-scale exploration of open Erdős problems and related research tasks, including cases where the system found previously open problems already solved in the literature, as well as some apparent novel solutions.deepmindopenai One published example concerns an Erdős–Graham irrationality problem, where the authors describe the original problem as being autonomously solved by the AI system, with the final paper developed through human–AI interaction.deepmind
This is meaningful evidence that AI can now contribute to real research-level mathematics, but the results vary in autonomy and scope.
3) Formal proof-search systems are now resolving open problems at scale
A more technical line of work uses systems that combine language models, search, and proof assistants such as Lean. One reported framework solved a number of open Erdős and OEIS problems and produced machine-checked proofs.deepmind The important point here is not just that the model “sounds right,” but that the outputs can be formally verified by a proof checker.
That said, these are still often special-purpose, scaffolded systems, not unconstrained autonomous mathematicians.
4) AI has produced publishable-level mathematical research outputs
There are also AI-assisted or AI-generated papers in arithmetic geometry and related areas, including work reported by DeepMind on eigenweights in arithmetic Hirzebruch proportionality.deepmind These show that AI is starting to operate inside genuine research workflows, not only contest-style benchmarks.
What this does and does not mean
It does mean:
- AI can now find new mathematical constructions.
- AI can sometimes disprove conjectures or improve known bounds.
- AI can produce machine-checkable proofs in formal systems.
- AI is no longer limited to olympiad-style or textbook mathematics.deepmindopenaideepmind
It does not yet mean:
- AI has solved the hardest famous open problems, such as the Millennium Prize problems.
- AI can reliably do open-ended mathematical research without human guidance.
- AI has replaced mathematicians in framing problems, judging significance, or turning raw discoveries into accepted theory.
Best current summary
The current evidence supports this statement:
AI has begun to make genuine mathematical discoveries, including some that resolve or overturn previously open conjectures, but it has not yet demonstrated broad autonomous mastery of frontier mathematics.
If you want, I can also give you a short table of the strongest examples, ranked by how convincing the evidence is.
Where this comes from 33 sources in this check · 249 dated across the record · 12 new +
Best available sources
Show 28 more sources
Sources that left the record (216)
A source dropping out is part of the record too: it carried an earlier version of this answer and is not part of the current one. The models re-run their own web search on every check, so single links come and go. The 8 most recently dropped are shown here, dated to their last appearance.
Every position, model by model Direction Shift 0/100 · Stable +
Where the models actually split
Each question keeps its own dimensions. Ask one model and you get one of these positions with no sign that the others exist.
The specific nature and origin of other major open problem breakthroughs besides the Erdős conjecture.
Model B claims breakthroughs include the Jacobian conjecture via Levent Alpöge and Anthropic's Claude, and AlphaProof Nexus solving 9 out of 353 open Erdős problems.
- Gemini
Model A and Model C focus on DeepMind's Aletheia/Gemini work, FunSearch, and general theorem-proving benchmarks rather than Model B's specific novel claims.
- Grok
- OpenAI
| Model | Jul 25 | Jul 27 | Jul 29 | Jul 30 | Jul 31 | Aug 01 | Aug 02 | Aug 03 | Aug 04 | Aug 05 | Aug 06 |
|---|---|---|---|---|---|---|---|---|---|---|---|
| OpenAI | |||||||||||
| Gemini | |||||||||||
| Grok | — | — |
The record behind this page 11 checks kept in full · no material change yet +
The record
Unchanged through 10 checks, 12 days.No material change since the first check on Jul 25, 2026. 11 checks kept in full, 249 sources dated.
How this answer held up
64/100 agreement at this checkThe same question, re-asked 11 times. The score moves between a small set of grading levels, so read the steps as levels, not as measurements.
No check in this window was graded a material change. The steps in the curve are wording-level differences between two answers that say the same thing.