The answer, statement by statement
What the answer is made of
Restated check after check
1The load-bearing part of this answer. The longest unbroken run on this Topic stands at 1 check.
Settled
AI has materially advanced several previously unsolved mathematical problems
1/5
Entered the record recently
2First stated within the last 3 checks — the least tested part of this answer.
New
disproves Erdős’s planar unit-distance conjecture
1/5
New
AlphaProof and AlphaGeometry 2 solved 4 of 6 problems at the 2024 IMO
1/5
Read the full answer Check 6 of 19 · 2026-08-01 · 3 models +
What the models said then
minorThe consensus was restructured with minor updates to phrasing and evidence framing (such as clarifying the nature of the Erdős conjecture disproof and the status of IMO results), but the core conclusions and main takeaways remain substantially the same.
Short answer
The current evidence says AI has materially advanced several previously unsolved mathematical problems, and in a few cases appears to have genuinely solved or disproved open conjectures, but the strongest evidence is still narrow and problem-specific, not a general proof that AI can independently solve frontier mathematics. The best-supported cases are open-problem disproofs, new constructions, and formal proof search, rather than landmark problems like the Riemann hypothesis or \(P \ne NP\).
What is clearly evidenced
1) A major open conjecture in discrete geometry was disproved
OpenAI reported that one of its reasoning models produced a proof that disproves Erdős’s planar unit-distance conjecture by constructing an infinite family with more unit-distance pairs than the conjecture allowed.openai
A companion paper by human mathematicians states that the AI-generated argument was mathematically valid, while also refining and simplifying parts of it.openai
This is the strongest public example so far of AI contributing to a genuine advance on a longstanding open problem.
2) AI has reached research-level formal proof capability
DeepMind’s AlphaProof and AlphaGeometry 2 solved 4 of 6 problems at the 2024 IMO, reaching a silver-medal-equivalent score.deepmind
That is not an “open problem” result, but it is strong evidence that AI can handle long, rigorous mathematical reasoning in formal settings.
3) AI systems have found new results in open mathematical research
DeepMind’s FunSearch used an LLM-guided search to discover new constructions related to the cap set problem, a long-standing open question in extremal combinatorics.deepmind
This counts as material progress on an open problem, though it is best described as a new construction/better bound rather than a complete resolution.
4) More recent work shows broader research capability
Recent formal-proof-search and agentic systems report solving subsets of open problems in databases such as Erdős and OEIS-style collections, but these claims vary in strength depending on whether the result was:
- formally verified,
- independently checked,
- based on a human-edited proof,
- or only reported in a preprint or company post.deepmind1stproof
What remains unproven or uncertain
- There is not strong evidence that AI has autonomously solved the deepest famous open problems.
- Many results still involve human verification, reformulation, or cleanup.
- Some reported “open problem” successes are on specialized or lower-complexity questions, not major flagship conjectures.
- For several 2025–2026 claims, the public record is still mostly preprints, official blog posts, or controlled evaluation reports, so long-term mathematical acceptance may still evolve.deepmind1stproof
Bottom line
The best current reading is:
AI has already made real mathematical discoveries and has at least one well-supported open-problem disproof to its credit, but it has not yet broadly solved frontier mathematics on its own.
The evidence is strongest for formal proof search, counterexample discovery, and construction of improved bounds, and weakest for claims of general autonomous theorem-solving across major unsolved problems.openaideepminddeepmind
If you want, I can next turn this into a table of the strongest verified examples with one-line status labels such as “solved,” “disproved,” “improved bound,” or “still provisional.”
Where this comes from 38 sources in this check · 161 dated across the record · 21 new +
Best available sources
Show 33 more sources
Sources that left the record (123)
A source dropping out is part of the record too: it carried an earlier version of this answer and is not part of the current one. The models re-run their own web search on every check, so single links come and go. The 8 most recently dropped are shown here, dated to their last appearance.
Every position, model by model Direction Shift — · Not comparable +
Where the models actually split
Each question keeps its own dimensions. Ask one model and you get one of these positions with no sign that the others exist.
AI has materially advanced several previously unsolved mathematical problems
Supported by
- OpenAI
- Grok
- Gemini
disproves Erdős’s planar unit-distance conjecture
Supported by
- OpenAI
- Grok
- Gemini
AlphaProof and AlphaGeometry 2 solved 4 of 6 problems at the 2024 IMO
Supported by
- Grok
used an LLM-guided search to discover new constructions related to the
Supported by
- Grok
| Model | Jul 25 | Jul 27 | Jul 29 | Jul 30 | Jul 31 | Aug 01 |
|---|---|---|---|---|---|---|
| OpenAI | ||||||
| Gemini | ||||||
| Grok | — |
The record behind this page 6 checks kept in full · no material change yet +
The record
Unchanged through 5 checks, 7 days.No material change since the first check on Jul 25, 2026. 6 checks kept in full, 161 sources dated.
How this answer held up
64/100 agreement at this checkThe same question, re-asked 6 times. The score moves between a small set of grading levels, so read the steps as levels, not as measurements.
No check in this window was graded a material change. The steps in the curve are wording-level differences between two answers that say the same thing.