Skip to content

Product1 publisher3 min readPublished

OpenAI puts a machine-checkable Lean proof of finite-time blowup in a public repo

OpenAI says the result resolves statement C of the official Millennium formulation, then says it will not claim the $1M prize. The smooth force applied to its fluid throughout is why that second sentence carries the weight.

The Product Desk · Product desk

Illustration accompanying OpenAI puts a machine-checkable Lean proof of finite-time blowup in a public repo

What happened

  • OpenAI has published a write-up, a PDF paper and a Lean formalisation of its Navier-Stokes argument in a public GitHub repository, so the mathematics can be checked by machine outside the company.
  • The write-up says the work resolves the Navier-Stokes Millennium Prize problem by establishing statement C, and also D, in the official Millennium Prize formulation.
  • Two paragraphs later the same page says the company does not intend to claim the Millennium Prize for the result.
  • The fluid in OpenAI's construction has a smooth force applied to it from rest until the singularity forms, where the Euler variant it credits to Levent Alpöge and Tristan Buckmaster was unforced.
  • The page now credits Alpöge and Buckmaster for concurrent work on the forced Euler problem and offers to recognise their priority in a joint announcement.

Compiled by The Product DeskSomething wrong?How this is made

Why it matters

  • decision The release rewards whoever has a Lean toolchain and time to run it; everyone else is back to reading the company's sentences about the company's result, which is a different quality of evidence.
  • constraint Checking by redoing is priced out of most of the field, so scrutiny of this class of result narrows to organisations that can fund a run of the same size.
  • exposure The provenance question stays open no matter how clean the formalisation is, because it is a claim about a training process nobody outside OpenAI can inspect.

At the announcement there was nothing outside the company to look at, which is the position TNW reported it from [15]. What changed is that there is now something to run. A Lean formalisation can be machine-checked by anyone with the toolchain, and per TNW that is the one part of this package that does not require trusting the vendor [2].

OpenAI says the formalisation and verification took a further 17 hours on GPT-6 Astra [3], the model that shipped last week. The search itself was run by an internal system the company describes as significantly more capable, with Astra demoted to checking the answer [11]. Set 17 hours against the roughly 88 hours the agents spent and checking came in near a fifth of finding: 17 divided by 88 is about 19 percent [21].

The finding side reads like a purchase order. Around 10,000 concurrent agents worked from 1 to 5 September, about 88 hours, consuming 2.7 million messages and roughly 130 billion output tokens [10]. Divided through, that is about 13 million output tokens per agent [19] and about 270 messages each [20]. OpenAI has previously put the compute bill in the millions [12]. This was not one model reasoning hard for a long time, the way teams often picture a mathematical result coming together; it was a swarm run on a three-and-a-half-day meter.

A Lean check certifies the theorem that was written down, force term included [2][8]. It does not certify that the theorem written down is the one the Clay Institute asked for, and TNW's read is that the forced-blowup question belongs to mathematicians rather than to a technology desk [16]. Turning down a $1M prize after 26 years of failed attempts signals something more than modesty [7], and that is what makes the declining sentence the most informative line on the page.

The credit position moved too. OpenAI denied Buckmaster's suggestion that it had pursued directions derived from his unpublished work with Alpöge [13]. The offer that has since appeared reads like an acknowledgement rather than a dismissal, and it costs nothing that a formalisation could ever verify.

Any result that arrives with a claim attached should be checked against two tests, kept in separate columns. One: is there an artifact you can check without the vendor's cooperation. Two: does that artifact cover the sentence in the headline. Both yes, and the headline is yours to repeat. If there is an artifact but coverage remains unresolved, which is where this sits, you take the theorem while holding the headline. Coverage asserted with no artifact is where this story stood in September. If neither holds, you are reading a press call. The discipline is refusing to let a passing score in column one settle column two.

What to watch

  • Whether the Clay Institute, or a named mathematician, rules on whether a forced blowup satisfies the official formulation.
  • Whether Alpöge and Buckmaster accept the joint priority announcement OpenAI has offered.
  • Whether a third party publishes the result of running the Lean artifact end to end, and how long it took them.
Loading claim ledger
Loading source directory links
Loading share composer
Loading topic controls
Loading related stories