"10,000 Agents, 88 Hours, 5 Million Messages"—Why is Plagiarism Allegations Surrounding OpenAI's Mathematical Achievement?
On September 8th, OpenAI announced that its unpublished internal model had solved one of the Millennium Prize Problems concerning the Navier-Stokes equations (equations describing fluid motion). However, just hours later, Tristan Buckmaster, a mathematician at New York University, accused OpenAI of potentially referencing his unpublished research without permission. As an engineer, I want to carefully examine the details of this technical achievement and the resulting conflict surrounding its recognition.
A Different Scale from Previous Case: "Approximately 10,000 Agents, 88 Hours"
First, let's examine the technical details. This approach differs from Astra's mathematical proof and Claude's formalization of Fermat's Last Theorem, which we previously discussed. According to OpenAI, an undisclosed internal model, said to be even more powerful than GPT-6 Astra, used approximately 10,000 agents in parallel for 88 hours to prove the "finite-time blowup" (a phenomenon where the solution diverges within a finite time) in the forced 3D Navier-Stokes equations.
In this process, the agents exchanged approximately 5 million messages and consumed approximately 300 billion output tokens. OpenAI itself describes this computational cost as "several million dollars," but external estimates based on publicly available token prices suggest a more specific figure of around $15 million to $22 million.
The Important Technical Distinction Between "Unsolved" and "Resolved"
There is a critically important distinction to accurately understand the content of this achievement. The Millennium Problems, which offer a $1 million prize, deal with the unforced Navier-Stokes equations, where no external forces are involved. However, OpenAI's proof addresses the forced version, where external coercion is applied. While related, this is a distinct problem of different nature.
The Clay Institute for Mathematics has not yet determined whether this result actually meets the criteria for the prize. Even if the proof is ultimately accepted, the prize money will not be paid until it is published as a peer-reviewed paper and undergoes two years of verification by the mathematical community. In other words, even if everything is correct, it will still take a long time before this achievement is "officially recognized." Fields Medal winner Terence Tao has praised the underlying approach as a "remarkable achievement," but has not yet formally endorsed the entire proof.
Buckmaster's Account of "Why Are You Ruining My Career?"
More than the technical achievement itself, the conflict over crediting the work has become a major topic of discussion within the industry. Buckmaster claims that as of August 22nd, he and Anthropic researcher Levento Alpege had already reached the proof of the relevant explosion phenomenon.
According to Buckmaster, on September 3rd, amidst rumors of the two companies' research, he contacted researchers at OpenAI. Initially, he received a friendly response, and they even offered to provide OpenAI's computing resources. However, in a call on September 6th, OpenAI researcher Sebastian Bubek revealed that their internal model had already completed a nearly 100-page forced Navier-Stokes proof, and allegedly asked Buckmaster to remove Alpege's name from the joint publication. When Buckmaster refused, he claims he was told, "Why are you ruining your career?"
Discrepancies between OpenAI's explanation and the timeline
OpenAI researcher Bübeck offers a different explanation for how this project began. According to him, OpenAI started the project after rumors spread on social media on September 1st that a competing lab had solved two Millennium Problems. He claims that they first tackled a related problem, the simpler "regularity of the unforcible Euler equation," with a 1,000-agent group working on it for 50 hours, before expanding resources to 10,000 agents for the Navier-Stokes problem.
A straightforward reading of this timeline suggests that OpenAI's project began on September 1st, after the time when Buckmaster and his colleagues allegedly reached their results (August 22nd). However, if OpenAI's subsequent claims that they already possessed a near-complete proof and requested the exclusion of co-authors are true, then it becomes possible that their actions were not based on independently achieved results, but rather on recognition of the existence of other research. This discrepancy remains unresolved at this point, with both sides' claims still at odds, and external verification is awaited.
What Engineers Should Consider
This incident demonstrates that the acceleration of mathematical research by AI is beginning to significantly shake the very framework of trust and credit within human research communities—the framework of "who knew what, when, and how"—not just the "speed" of research results. The tension between the results derived by 10,000 agents in 88 hours and the painstaking research accumulated by human researchers is likely to resurface repeatedly as AI-accelerated research becomes more widespread.
The 166-page paper and the formal verification code using Lean have already been made public for external review. It will be necessary to closely monitor how the technical correctness of this proof and the veracity of both sides' claims regarding credit will be verified.