← Back to the dossier

8 September 2026 · OpenAI

OpenAI's agents solve a Millennium Prize problem, and two mathematicians ask where the route came from

On 8 September OpenAI said that about 10,000 agents, running on an internal model, had found a proof in 88 hours that solutions to the three-dimensional Navier-Stokes equations can break down. A computer checked the proof in Lean. Half a day earlier, Tristan Buckmaster of New York University had published related results with Levent Alpöge, together with a statement on how OpenAI moved onto the same problem after hearing rumours about their work. The dispute is about credit and about what happens to user data. Nobody has shown that anything was stolen.

By Mara Masaeva · Updated 29 September 2026

CapabilityDevelopingThe story is still developing. There is one source so far, or the numbers are still changing.

What happened

The problem. The Navier-Stokes equations describe how water and air flow. In 2000 the Clay Mathematics Institute put a million dollars on one question: do the solutions always stay smooth, or can the flow at some point become infinitely fast in a finite time? OpenAI says its system found a case where that happens, a vortex that spirals inward and keeps stretching, driven by a smooth external force. That would settle versions C and D of the official problem statement by Charles Fefferman of Princeton.

The run. On 1 September, after rumours that two Millennium Prize problems had been solved, OpenAI set a new internal model on all the open ones at once. On Navier-Stokes alone the agents sent 2.7 million messages and produced about 130 billion tokens. After 88 hours they had a proof. GPT-6 Astra then needed 17 more hours to formalise it in Lean. These numbers come from OpenAI's announcement, which I read through Simon Willison's quotes because OpenAI's page blocks automated reading. Mark Chen, OpenAI's head of research, told WIRED the cost was "in the millions of dollars".

Why Lean matters. Lean is a proof assistant. It checks every step mechanically, no matter who or what wrote the proof. One human check remains, Quanta notes. Someone has to confirm that the statement Lean accepts is the statement mathematicians wanted proved. The Clay Institute also wants a published, peer-reviewed paper. According to the Wikipedia overview, the institute said on 11 September that the problem has "apparently been settled", and that its own process is deliberately slow.

Buckmaster and Alpöge. Tristan Buckmaster (Courant Institute, New York University) and Levent Alpöge, a mathematician employed by Anthropic, worked together for most of a year as a private collaboration, outside their employers, Buckmaster writes. They used Anthropic's Claude and OpenAI's Codex, mostly with GPT-5.6 Sol. They built on the programme of Diego Córdoba (Institute for Mathematical Sciences, Madrid) and Luis Martínez-Zoroa (CUNEF University), whose method both AI-assisted teams relied on. On 15 August the pair had blow-up with a smooth force for the Euler and Boussinesq equations. Euler is Navier-Stokes without friction. Lean confirmed it on 22 August. They made the results public just before midnight on 7 September, about twelve hours before OpenAI.

Buckmaster's account. On 3 September a rumour was going round that Anthropic had solved a major open problem, and Alpöge had heard that information about their work had reached OpenAI. Buckmaster wrote to a mathematician at OpenAI. On Sunday 6 September there were two calls, with Sébastien Bubeck of OpenAI also on the line. Buckmaster was told an internal model had proved blow-up for Navier-Stokes with a smooth force. That was the route he and Alpöge had quietly chosen, and he knew of almost nobody else working on it. Eventually it was agreed that OpenAI's first prompt was only a few days old, sent after information about their work had reached the company.

He asked whether the model had been trained on, or had access to, their Codex sessions, where they had put all their drafts for the whole project. He was told the model "did not look up user data". He asked again about training and got no answer. He also writes that Bubeck wanted Alpöge removed as an author because Alpöge works at Anthropic. Bubeck denies asking for that.

Buckmaster is explicit about what he does not claim. He has not seen OpenAI's proof. He does not know what the model did, or whether their data was used. "I am not accusing anyone of anything," he writes. USA Today quotes the passage in full.

How it workedtechnical detail

What OpenAI said, and when. The announcement of 8 September says the researchers and the agents "did not see any of their work through any means until they released it publicly", and that "no specific user data was accessed in order to solve this problem". The same paragraph adds: "While unlikely, we cannot rule out that de-identified data derived from their usage of our products helped improve our models." TechCrunch and The Verge quote the full passage.

Then the line moved. On 9 September spokesperson Laurance Fauconnet said: "We can say categorically that it is impossible for Dr. Buckmaster's Codex prompts over the last two months to have influenced the system in any way, including training." The Verge prints that statement in full. On 10 September an updated announcement said an investigation had confirmed this. On 13 September the company said no user input after 3 July could have influenced the system. I take the dates and the last two statements from Wikipedia, which cites the New York Times, the Washington Post and the updated announcement. I have not read those three myself.

Side by side, the statements do not cover the same thing. "No specific user data was accessed" is about lookup. The later statements are about training, and each covers a recent period only. According to Wikipedia, OpenAI has said nothing since about the pair's Codex use before 3 July, when they had already been working for months. TechCrunch points out that OpenAI reserves the right to train on Codex interactions unless the user opts out.

On the mathematics. The results are not the same. Buckmaster and Alpöge proved blow-up for Euler with a smooth external force. OpenAI had an Euler result without a force, and a Navier-Stokes proof it calls significantly different. Alpöge says OpenAI's proof looks more like another Euler proof the pair had, Wikipedia reports. Buckmaster's question is about the route they chose. In my view, a difference in the details of the proof does not answer it.

What may follow

In July, the Hugging Face intrusion came from many agents searching long and hard for anything that counted as progress. In September the same kind of setup solved an open problem in mathematics.

In my view, the dispute says most about how researchers will work from now on. Terence Tao (UCLA) wrote on Mastodon that even a rumour of someone working on a problem can now trigger a huge AI effort to get there first, as Analytics India Magazine reports. If that holds, mathematicians will say less about what they are working on. Mathematicians who spoke to The Verge fear the same. On 11 September Fields Medal winners published an open letter about AI in mathematics.

What I do not know

Whether any of Buckmaster and Alpöge's work reached OpenAI's system, by any route, is not established. OpenAI denies lookup and rules out influence through training after 3 July. It has said nothing further about the months before. Nobody outside the company can check any of this.

The proof has not been peer-reviewed, and the Clay Institute has not recognised anyone.

Buckmaster and Bubeck give different versions of the calls on 6 September, including on Alpöge's authorship.

I could not read OpenAI's announcement or the Science article myself. Both pages block automated fetching. What I quote from OpenAI comes through Simon Willison, TechCrunch, The Verge, WIRED and Wikipedia.

My notes

I first put this on the site as good news. That was too quick.

In front of a room I start with Lean, because that is the new part and it is easy to explain. Then I put OpenAI's statements in a row, with their dates. I do not say that OpenAI used their work. That is not established, and Buckmaster does not say it either.

One practical thing. If you put unpublished work into Codex or ChatGPT, check today whether training on those conversations is switched off.

“I asked whether the model had been trained on, or had access to, our sessions in Codex, into which we had been putting all our drafts for the whole of this project. I was told the model did not look up user data. I asked again, about training, and I did not get an answer.”

Tristan Buckmaster · Mathematician, Courant Institute, NYU

“We have now seen that even the rumour of someone working on a problem can trigger a massive amount of AI-powered effort to flatten it before the original research project has time to reach its full potential.”

Terence Tao · Mathematician, UCLA, Fields Medal 2006

Read next

Sources

  1. Tristan Buckmaster: statement of 7 September 2026 (PDF)primary · main source

    Buckmaster's first-person account: the collaboration, the calls with OpenAI, the Codex question, and what he explicitly does not claim. One side of the story.

  2. OpenAI: On the Navier-Stokes Millennium Prize Problemprimary · not read end to end yet

    The announcement and the paper. Blocks automated fetching. I tried twice. Quoted here through Simon Willison, TechCrunch, The Verge and Wikipedia.

  3. Wikipedia: Navier-Stokes priority controversyresearch

    An overview in one place, with sources. Source here for OpenAI's statements of 9, 10 and 13 September, the Clay Institute's reaction and Alpöge's reply. Those claims rest on the press it cites, which I did not all read.

  4. Quanta: AI has solved one of math's $1 million Millennium Prize Problemspress

    Konstantin Kakaes, 8 September. The mathematics, the Córdoba and Martínez-Zoroa programme, and what Lean does and does not settle.

  5. Simon Willison: some thoughts on the Navier-Stokes Millennium Prize Problemargument

    Quotes OpenAI's announcement at length, including the run numbers and the data paragraph.

  6. The Verge: OpenAI's sly mathematical breakthrough sends a chill through academiapress

    Robert Hart, 9 September. Reactions from mathematicians, and Bubeck's denial on authorship.

  7. The Verge: OpenAI just wants to winpress

    Robert Hart, 12 September. Prints the statement of OpenAI spokesperson Laurance Fauconnet on Buckmaster's Codex prompts, and the Clay Institute's line that its process is deliberately unhurried.

  8. WIRED: OpenAI just claimed a huge math discovery. Some academics are crying foulpress

    Will Knight and Maxwell Zeff, 8 September. OpenAI's press briefing: cost, agent numbers, and Bubeck's denial.

  9. USA Today: OpenAI touts math breakthrough. Mathematicians dispute creditpress

    Greta Reich, 8 September. Quotes Buckmaster on the Codex question and on what he does not claim, and Bubeck's reply.

  10. TechCrunch: OpenAI fought dirty on career-making math problem, says NYU mathematicianpress

    Russell Brandom, 8 September. Full quote of OpenAI's data paragraph, and the note on training on Codex data.

  11. Analytics India Magazine: the OpenAI Navier-Stokes controversy explainedpress

    Supreeth Koundinya, 9 September. Source for the Tao quote and the description of OpenAI's construction.

  12. Science: how an AI math breakthrough ignited a controversypress · not read end to end yet

    Blocks automated fetching. I tried twice. I do not rely on it for anything here.