
OpenAI just lately introduced one thing extraordinary: an inside AI system had produced a proposed resolution to the Navier-Stokes existence and smoothness downside, one in all arithmetic’ legendary Millennium Prize Issues.
Extra exactly, the system constructed a finite-time singularity for the three-dimensional Navier-Stokes equations with a easy exterior pressure, one of many routes permitted by the official Clay Arithmetic Institute formulation of the issue. It doesn’t settle the better-known query of whether or not the unforced equations all the time stay easy.
In line with OpenAI’s official report, the Navier-Stokes effort concerned on the order of 10,000 concurrent AI brokers, 2.7 million messages and roughly 130 billion output tokens. The brokers reached their proposed resolution about 88 hours after the experiment started, adopted by one other 17 hours of Lean formalization and verification.
At first look, it sounds just like the form of end result folks have lengthy imagined from extremely autonomous AI: give a system an issue mathematicians have struggled with for many years, let it work for just a few days, and get a proof again.
However the story of how OpenAI acquired there may be significantly extra attention-grabbing.
The Story Began With Two Human Mathematicians
Earlier than OpenAI launched these 10,000 brokers, mathematicians Tristan Buckmaster of NYU and Levent Alpöge, who works at Anthropic, had already been making main progress on intently associated fluid dynamics issues.
They weren’t working with out AI both.
Their analysis used instruments together with Claude and OpenAI Codex, and their outcomes have been formally verified utilizing Lean. Their work included developing finite-time blowup for the three-dimensional incompressible Euler equations with easy forcing.
Their end result did not clear up Navier-Stokes, but it surely pushed additional right into a intently associated space.
Then one thing attention-grabbing occurred.
In line with OpenAI, on September 1, 2026 it heard rumors that two Millennium Prize issues had been solved. These rumors, mixed with sturdy outcomes from its newly educated inside mannequin, prompted the corporate to check the system throughout the remaining open Millennium Prize issues.
OpenAI later realized the rumors have been linked to Buckmaster and Alpöge.
So AI didn’t merely get up one morning, independently select Navier-Stokes and clear up it from nowhere.
Human researchers have been already making important progress across the similar frontier.
However importantly, it was OpenAI’s personal Euler end result that later satisfied the corporate to pay attention its sources particularly on Navier-Stokes.
Then OpenAI Scaled It to 10,000 Brokers
That is the place OpenAI did one thing genuinely uncommon.
It basically created a big digital analysis lab.
The brokers have been divided into teams that might talk internally, run code and entry a cached model of the web.
Completely different teams explored totally different mathematical approaches.
Almost 100 brokers first labored for about 50 hours on an Euler-related downside and produced a end result that OpenAI thought of promising.
At that time, OpenAI redirected brokers away from different Millennium Prize issues and towards Navier-Stokes.
It additionally started sharing helpful findings between teams.
OpenAI describes this as cross-pollination: Codex consolidated promising intermediate outcomes from totally different agent teams, and people findings have been fed again into later prompts.
Ultimately, the profitable effort concerned roughly 10,000 concurrent brokers.
That will really be the most important story right here.
A single mathematician can discover a handful of concepts.
A analysis group can discover extra.
Ten thousand brokers can examine large numbers of doable paths concurrently, discard failures and preserve constructing on promising outcomes.
Maybe the breakthrough is not that AI out of the blue grew to become a mathematical genius.
Maybe it grew to become an extremely scalable analysis workforce.
Then the Controversy Began
After OpenAI’s end result emerged, Buckmaster raised an uncomfortable query.
He and Alpöge had been utilizing Codex whereas creating their very own analysis and, based on Buckmaster, had put drafts from the mission into the instrument.
So he requested OpenAI whether or not its new mannequin had been educated on, or had entry to, these classes.
As detailed in ABC Information’ account of the dispute, Buckmaster stated he was initially advised that the mannequin didn’t “search for” person information. When he requested particularly about coaching, he stated he didn’t instantly obtain a solution.
Buckmaster was cautious to not accuse OpenAI, saying, “I have no idea what their mannequin did, or how,” and that he didn’t know whether or not their information had been used.
OpenAI later investigated.
In an replace to its report, the corporate stated Buckmaster’s Codex prompts from the previous two months couldn’t have influenced the system in any method, together with by way of coaching. OpenAI additionally says its researchers and brokers had not seen Buckmaster and Alpöge’s unpublished work earlier than it grew to become public.
So based mostly on the proof at the moment accessible, there isn’t any foundation for saying OpenAI educated on their unpublished proof or copied their personal analysis.
However then one other disagreement emerged: who will get credit score?
There have been now two separate outcomes: Buckmaster and Alpöge’s Euler work and OpenAI’s Navier-Stokes end result.
In line with Buckmaster, OpenAI researcher Sébastien Bubeck introduced two doable methods ahead.
One was for Buckmaster and Alpöge to publish their Euler end result earlier than OpenAI launched Navier-Stokes.
The opposite was extra controversial.
Buckmaster says he was provided the prospect to write a paper presenting OpenAI’s Navier-Stokes proof, clearly acknowledging that an OpenAI mannequin had generated it.
However Alpöge wouldn’t be included as a result of he labored for Anthropic.
Buckmaster refused.
Bubeck later stated he had proposed Buckmaster as lead creator of a rewritten presentation of OpenAI’s proof, and that he thought of it inappropriate for an Anthropic worker to creator OpenAI’s work. He additionally clarified that he had by no means proposed eradicating Alpöge from Buckmaster and Alpöge’s personal Euler paper.
That disagreement factors to an issue educational publishing was by no means actually designed for.
If people develop the encircling concepts, AI instruments take part within the analysis, one other AI system generates the ultimate proof and people nonetheless have to interpret and publish it, who really will get the credit score?
So What Did AI Really Uncover?
OpenAI’s experiment was not 10,000 brokers ranging from a clean web page and inventing fluid dynamics from scratch.
It constructed on many years of arithmetic, latest human progress in the identical space, and analysis instructions that have been already turning into promising. People additionally determined the place to allocate compute, whereas OpenAI intentionally handed helpful intermediate findings between agent teams.
However that doesn’t make the end result any much less important.
What OpenAI demonstrated is one thing totally different: 1000’s of AI brokers can discover analysis paths in parallel, discard useless ends, mix promising concepts and compress an infinite quantity of labor into only a few days.
The Clay Arithmetic Institute has since stated that the Navier-Stokes downside “has apparently been settled,” whereas making clear that evaluating the end result and figuring out credit score will take time.
So maybe the essential query is just not whether or not AI has out of the blue reached synthetic normal intelligence.
What this experiment actually exhibits is that we are able to now give a really tough analysis downside to 1000’s of AI brokers, allow them to discover many alternative concepts on the similar time, and doubtlessly end in days what may take human researchers for much longer.
That may be a main shift.
Nevertheless it additionally creates a brand new query: if AI is constructing on many years of human analysis after which exploring these concepts at huge scale, what a part of the ultimate end result ought to we name a real AI discovery?
Ultimate Ideas
For me, essentially the most attention-grabbing a part of this experiment is just not that AI solved a well-known arithmetic downside in 88 hours.
It’s the way it acquired there.
OpenAI didn’t ask one mannequin to take a seat and assume tougher. It created 1000’s of brokers, despatched them down totally different analysis paths, moved extra compute towards promising concepts, shared helpful discoveries between teams, and saved going till one thing labored.
That begins to look much less like a chatbot and extra like a analysis group working at machine velocity.
Actually, OpenAI has already described its broader objective as constructing an automated AI researcher.
However Navier-Stokes additionally exhibits why we must be cautious with the phrase discovery.
The brokers have been engaged on high of many years of arithmetic, current analysis, human choices and concepts that have been already creating on the frontier.
So possibly the true milestone right here is just not that AI out of the blue discovered the way to uncover issues by itself.
It’s that analysis itself might now be scalable.
And if 10,000 AI brokers can already do that at the moment, the extra attention-grabbing query is what occurs when the identical strategy is utilized to 1000’s of different unsolved issues throughout arithmetic, science and engineering.
Abid Ali Awan (@1abidaliawan) is an authorized information scientist skilled who loves constructing machine studying fashions. At present, he’s specializing in content material creation and writing technical blogs on machine studying and information science applied sciences. Abid holds a Grasp’s diploma in know-how administration and a bachelor’s diploma in telecommunication engineering. His imaginative and prescient is to construct an AI product utilizing a graph neural community for college kids fighting psychological sickness.

OpenAI just lately introduced one thing extraordinary: an inside AI system had produced a proposed resolution to the Navier-Stokes existence and smoothness downside, one in all arithmetic’ legendary Millennium Prize Issues.
Extra exactly, the system constructed a finite-time singularity for the three-dimensional Navier-Stokes equations with a easy exterior pressure, one of many routes permitted by the official Clay Arithmetic Institute formulation of the issue. It doesn’t settle the better-known query of whether or not the unforced equations all the time stay easy.
In line with OpenAI’s official report, the Navier-Stokes effort concerned on the order of 10,000 concurrent AI brokers, 2.7 million messages and roughly 130 billion output tokens. The brokers reached their proposed resolution about 88 hours after the experiment started, adopted by one other 17 hours of Lean formalization and verification.
At first look, it sounds just like the form of end result folks have lengthy imagined from extremely autonomous AI: give a system an issue mathematicians have struggled with for many years, let it work for just a few days, and get a proof again.
However the story of how OpenAI acquired there may be significantly extra attention-grabbing.
The Story Began With Two Human Mathematicians
Earlier than OpenAI launched these 10,000 brokers, mathematicians Tristan Buckmaster of NYU and Levent Alpöge, who works at Anthropic, had already been making main progress on intently associated fluid dynamics issues.
They weren’t working with out AI both.
Their analysis used instruments together with Claude and OpenAI Codex, and their outcomes have been formally verified utilizing Lean. Their work included developing finite-time blowup for the three-dimensional incompressible Euler equations with easy forcing.
Their end result did not clear up Navier-Stokes, but it surely pushed additional right into a intently associated space.
Then one thing attention-grabbing occurred.
In line with OpenAI, on September 1, 2026 it heard rumors that two Millennium Prize issues had been solved. These rumors, mixed with sturdy outcomes from its newly educated inside mannequin, prompted the corporate to check the system throughout the remaining open Millennium Prize issues.
OpenAI later realized the rumors have been linked to Buckmaster and Alpöge.
So AI didn’t merely get up one morning, independently select Navier-Stokes and clear up it from nowhere.
Human researchers have been already making important progress across the similar frontier.
However importantly, it was OpenAI’s personal Euler end result that later satisfied the corporate to pay attention its sources particularly on Navier-Stokes.
Then OpenAI Scaled It to 10,000 Brokers
That is the place OpenAI did one thing genuinely uncommon.
It basically created a big digital analysis lab.
The brokers have been divided into teams that might talk internally, run code and entry a cached model of the web.
Completely different teams explored totally different mathematical approaches.
Almost 100 brokers first labored for about 50 hours on an Euler-related downside and produced a end result that OpenAI thought of promising.
At that time, OpenAI redirected brokers away from different Millennium Prize issues and towards Navier-Stokes.
It additionally started sharing helpful findings between teams.
OpenAI describes this as cross-pollination: Codex consolidated promising intermediate outcomes from totally different agent teams, and people findings have been fed again into later prompts.
Ultimately, the profitable effort concerned roughly 10,000 concurrent brokers.
That will really be the most important story right here.
A single mathematician can discover a handful of concepts.
A analysis group can discover extra.
Ten thousand brokers can examine large numbers of doable paths concurrently, discard failures and preserve constructing on promising outcomes.
Maybe the breakthrough is not that AI out of the blue grew to become a mathematical genius.
Maybe it grew to become an extremely scalable analysis workforce.
Then the Controversy Began
After OpenAI’s end result emerged, Buckmaster raised an uncomfortable query.
He and Alpöge had been utilizing Codex whereas creating their very own analysis and, based on Buckmaster, had put drafts from the mission into the instrument.
So he requested OpenAI whether or not its new mannequin had been educated on, or had entry to, these classes.
As detailed in ABC Information’ account of the dispute, Buckmaster stated he was initially advised that the mannequin didn’t “search for” person information. When he requested particularly about coaching, he stated he didn’t instantly obtain a solution.
Buckmaster was cautious to not accuse OpenAI, saying, “I have no idea what their mannequin did, or how,” and that he didn’t know whether or not their information had been used.
OpenAI later investigated.
In an replace to its report, the corporate stated Buckmaster’s Codex prompts from the previous two months couldn’t have influenced the system in any method, together with by way of coaching. OpenAI additionally says its researchers and brokers had not seen Buckmaster and Alpöge’s unpublished work earlier than it grew to become public.
So based mostly on the proof at the moment accessible, there isn’t any foundation for saying OpenAI educated on their unpublished proof or copied their personal analysis.
However then one other disagreement emerged: who will get credit score?
There have been now two separate outcomes: Buckmaster and Alpöge’s Euler work and OpenAI’s Navier-Stokes end result.
In line with Buckmaster, OpenAI researcher Sébastien Bubeck introduced two doable methods ahead.
One was for Buckmaster and Alpöge to publish their Euler end result earlier than OpenAI launched Navier-Stokes.
The opposite was extra controversial.
Buckmaster says he was provided the prospect to write a paper presenting OpenAI’s Navier-Stokes proof, clearly acknowledging that an OpenAI mannequin had generated it.
However Alpöge wouldn’t be included as a result of he labored for Anthropic.
Buckmaster refused.
Bubeck later stated he had proposed Buckmaster as lead creator of a rewritten presentation of OpenAI’s proof, and that he thought of it inappropriate for an Anthropic worker to creator OpenAI’s work. He additionally clarified that he had by no means proposed eradicating Alpöge from Buckmaster and Alpöge’s personal Euler paper.
That disagreement factors to an issue educational publishing was by no means actually designed for.
If people develop the encircling concepts, AI instruments take part within the analysis, one other AI system generates the ultimate proof and people nonetheless have to interpret and publish it, who really will get the credit score?
So What Did AI Really Uncover?
OpenAI’s experiment was not 10,000 brokers ranging from a clean web page and inventing fluid dynamics from scratch.
It constructed on many years of arithmetic, latest human progress in the identical space, and analysis instructions that have been already turning into promising. People additionally determined the place to allocate compute, whereas OpenAI intentionally handed helpful intermediate findings between agent teams.
However that doesn’t make the end result any much less important.
What OpenAI demonstrated is one thing totally different: 1000’s of AI brokers can discover analysis paths in parallel, discard useless ends, mix promising concepts and compress an infinite quantity of labor into only a few days.
The Clay Arithmetic Institute has since stated that the Navier-Stokes downside “has apparently been settled,” whereas making clear that evaluating the end result and figuring out credit score will take time.
So maybe the essential query is just not whether or not AI has out of the blue reached synthetic normal intelligence.
What this experiment actually exhibits is that we are able to now give a really tough analysis downside to 1000’s of AI brokers, allow them to discover many alternative concepts on the similar time, and doubtlessly end in days what may take human researchers for much longer.
That may be a main shift.
Nevertheless it additionally creates a brand new query: if AI is constructing on many years of human analysis after which exploring these concepts at huge scale, what a part of the ultimate end result ought to we name a real AI discovery?
Ultimate Ideas
For me, essentially the most attention-grabbing a part of this experiment is just not that AI solved a well-known arithmetic downside in 88 hours.
It’s the way it acquired there.
OpenAI didn’t ask one mannequin to take a seat and assume tougher. It created 1000’s of brokers, despatched them down totally different analysis paths, moved extra compute towards promising concepts, shared helpful discoveries between teams, and saved going till one thing labored.
That begins to look much less like a chatbot and extra like a analysis group working at machine velocity.
Actually, OpenAI has already described its broader objective as constructing an automated AI researcher.
However Navier-Stokes additionally exhibits why we must be cautious with the phrase discovery.
The brokers have been engaged on high of many years of arithmetic, current analysis, human choices and concepts that have been already creating on the frontier.
So possibly the true milestone right here is just not that AI out of the blue discovered the way to uncover issues by itself.
It’s that analysis itself might now be scalable.
And if 10,000 AI brokers can already do that at the moment, the extra attention-grabbing query is what occurs when the identical strategy is utilized to 1000’s of different unsolved issues throughout arithmetic, science and engineering.
Abid Ali Awan (@1abidaliawan) is an authorized information scientist skilled who loves constructing machine studying fashions. At present, he’s specializing in content material creation and writing technical blogs on machine studying and information science applied sciences. Abid holds a Grasp’s diploma in know-how administration and a bachelor’s diploma in telecommunication engineering. His imaginative and prescient is to construct an AI product utilizing a graph neural community for college kids fighting psychological sickness.















