Skip to content

Topic

AI for Mathematics

Use of large language models to attack open problems in mathematics and theoretical computer science, and how the resulting output is classified.

Current stories

science1 publisher

Lorraine mathematician confirms Claude's 67.25% bound on Riemann zeta zeros by another route

Youness Lamzouri has independently confirmed Claude's result that at least 67.25% of Riemann zeta zeros lie on the critical line, against 40% known since 1989. The hypothesis itself is still open. The case shows a working split for AI in mathematics, in which a model finds a candidate result and a specialist rebuilds it.

Publishers:livescience.com

Reality

Evidence58
Adoption
Insufficient
Hype gap+15
Incentives40
Confidence60
science1 publisher

OpenAI's 10,000 agents propose a Navier-Stokes blowup that needs an external force

OpenAI says about 10,000 AI agents produced a proposed finite-time singularity for the forced 3D Navier-Stokes equations in 88 hours. The construction fits one route the Clay rules allow and leaves unforced smoothness open, while the mathematicians whose forced-Euler work came first ask whether their Codex drafts reached the model.

Publishers:kdnuggets.com

Reality

Evidence35
Adoption
Insufficient
Hype gap+30
Incentives70
Confidence40
science7 publishers

Terence Tao says automated proof checking is why the AI labs stopped consulting mathematicians

A proof checker confirmed the steps of OpenAI's Navier-Stokes result within days. The argument now running through mathematics is over who explains the 166 pages and who gets the credit for them.

Perspective Coverage

7 publishers
Builder
Builder 46%
Operator
Operator 33%
Investor
Investor 21%

Reality

Evidence60
Adoption
Insufficient
Hype gap+40
Incentives70
Confidence58
leadership5 publishers

OpenAI's Navier-Stokes claim sparks dispute over whether mathematicians' data was accessed

OpenAI says no specific user data was accessed to solve the problem, and that it cannot rule out that de-identified data derived from two mathematicians' use of its products improved its models. A procurement team has to read both sentences.

Perspective Coverage

5 publishers
Builder
Builder 36%
Operator
Operator 33%
Investor
Investor 31%

Reality

Evidence55
Adoption
Insufficient
Hype gap+45
Incentives70
Confidence60
build3 publishers

OpenAI gated an 88-hour, 10,000-agent proof search on a 17-hour Lean check

OpenAI's write-up gives the token counts, the agent count and the verification time for its Navier-Stokes result. The verification time is the number that decides whether the method transfers to anyone else's workload.

Perspective Coverage

3 publishers
Builder
Builder 52%
Operator
Operator 33%
Investor
Investor 15%

Reality

Evidence60
Adoption
Insufficient
Hype gap+30
Incentives70
Confidence55
science3 publishers

Scientific American puts Birch and Swinnerton-Dyer first in line to fall to AI

With Navier-Stokes claimed and five Millennium problems left, the ordering of what falls next turns on which conjectures a single counterexample would settle. Five million dollars of Clay prize money is still unclaimed.

Perspective Coverage

3 publishers
Builder
Builder 37%
Operator
Operator 40%
Investor
Investor 23%

Reality

Evidence55
Adoption35
Hype gap+35
Incentives70
Confidence50

Earlier coverage

  1. Andreas Thom wants OpenAI to prove his ChatGPT sessions stayed out of the training data

    Product · September 10, 2026 · 1 publisher

  2. OpenAI spent more than $1m of compute on a prize it says it will not claim

    Invest · September 9, 2026 · 1 publisher

  3. OpenAI's Navier-Stokes statement on user data has two parts, not a firm "we don't train on your data" promise

    Product · September 9, 2026 · 2 publishers

  4. Mathematicians set the verification bar for AI proofs by shipping the Lean file

    Product · September 8, 2026 · 1 publisher

  5. Eleven days of machine time turned Wiles's 100 pages into 13 million lines of Lean

    Science · September 5, 2026 · 2 publishers

  6. Anthropic put its Fermat proof's correctness check inside the default build target

    Leadership · September 5, 2026 · 3 publishers

  7. Vulnerability disclosures bent upward in 2026. Algorithm records did not.

    Security · August 25, 2026 · 1 publisher

  8. A Fields Medalist read the AI maths papers: most of the wins are counterexamples

    Product · August 17, 2026 · 1 publisher