OpenAI, the company behind ChatGPT, has claimed to have solved a decades-old advanced mathematics problem using a new artificial intelligence (AI) model and thousands of AI bots, according to BBC. The company said on Tuesday that it used around 10,000 AI agents to solve a notoriously difficult math problem in just 88 hours.

Problem in the Navier-Stokes Equations

The problem was part of the Navier-Stokes equations, which concern how fluids move. For 90 years, important aspects of the problem have lacked a proof, the argument underlying a correct math equation. OpenAI called the solution a “milestone” and evidence that AI tools are improving quickly.

Independent Verification Pending

OpenAI’s solution has yet to be verified independently or publicly accepted by The Clay Mathematics Institute, a math organization based in the US which runs the Millennium Prize that offers big money to the first to solve certain mathematical conundrums. The company said it started to train a new model in late August, which quickly showed it was adept at math. AI models are computer programs trained on huge amounts of data to recognize and predict patterns in information.

While the new OpenAI model remains a tool only used within the company, as it is “significantly more capable” than the company’s most recent AI model release, its researchers decided to use it on certain notable advanced math problems. OpenAI admitted that last week on 1 September, it had “heard rumors that two Millennium Prize problems had been resolved” and so decided to put thousands of AI bots trained on the new internal model to work attempting to solve some of the remaining problems.

By Sept 5, or roughly 88 hours after it had set 10,000 AI bots to the task, OpenAI had found a solution to what’s referred to as the Navier,Stokes existence and smoothness problem. The problem is core to turbulence, which is a phenomenon that is still not well understood.

Although it took the AI bots seemingly little time to reach a solution, OpenAI said the bots exchanged nearly 3 million messages and used up 130 billion output tokens, or the individual lines of text and code an AI model produces in answers, on Navier–Stokes alone. Such an effort would have cost roughly $10m (£7.3m), based on OpenAI’s own pricing for output from its most advanced models.

Controversy and Concurrent Work

The solution that OpenAI says it has now reached for the Navier,Stokes existence and smoothness problem resolved two out of the four statements in the proof that the Millennium Prize had demanded. The prize is worth $1m to a winner. “Our goal in releasing this result is to report on the substantial progress of our AI models,” OpenAI said on Tuesday. “We do not intend to claim the Millennium Prize for this result.”

Already the company’s claim is stirring some controversy. Tristan Buckmaster, a mathematics professor at New York University, said on Tuesday that he and Levent Alpöge, a mathematician working for OpenAI rival Anthropic, had also been working toward solutions for the problem. The duo had been using OpenAI’s tool Codex in their work. But Buckmaster said that on 3 September, he found out that “information about our progress had been passed to OpenAI.”

Buckmaster’s statement came the same day, but hours before, OpenAI published its Navier–Stokes work. He claimed that OpenAI did not begin working on the Navier,Stokes equations until “after information about our work had reached OpenAI.” He included text from emails exchanged with OpenAI over the work and his questions of the company’s timing and methods.

Buckmaster added that he had not yet read OpenAI’s full proof, but felt compelled to go public with “what I was told, when, and what was proposed to me…because the alternative is to let a sequence of announcements say something I know to be false.” OpenAI on Tuesday congratulated the “concurrent work” of Buckmaster and Alpöge, calling it “remarkable.” The company said it had not seen “any of their work through any means until they released it publicly” and that no user data was accessed in its work on the Navier,Stokes problem.

“While unlikely, we cannot rule out that de-identified data derived from their usage of our products helped improve our models,” the company added. “However, our proofs differ significantly and even the precise results proved are different.”

At a press briefing on Tuesday, OpenAI researcher Sebastien Bubeck denied that the company had used the pair’s work or accessed material shared with OpenAI’s servers. OpenAI’s announcement on the breakthrough gave more details, saying that the company was inspired to launch the effort after hearing rumors that two Millennium Prize Problems had been solved. It also said it could not rule out that data from the pair’s use of their products “helped improve our models.”

OpenAI said it tackled the Navier-Stokes problem with an internal OpenAI system that was more powerful than its latest GPT-6 Astra model. About 10,000 AI agents, AI systems that carry out tasks autonomously – worked on the problem at once and reached the solution in 88 hours. “A major goal of our work is to empower scientists to advance research and technology that benefits all of humanity,” OpenAI researchers wrote in a blog. “We believe it is important to inform the world about the pace of AI progress and what to expect from upcoming models”.

The breakthrough is the latest to demonstrate how advanced artificial intelligence models consuming enormous amounts of computing power are encroaching on and changing the field of mathematics. In May, OpenAI claimed progress in another mathematical problem first proposed 80 years ago, while Google DeepMind has also claimed maths achievements with its models.