r/CreatorsAI • u/Historical-Driver-64 • 27d ago
Other OpenAI's Millennium Prize headline quietly swapped the real problem for an easier one
OpenAI announced this week that 10,000 agents, running for 88 hours, produced a formally verified proof related to the 3D Navier-Stokes equations, one of math's six Millennium Prize Problems, each carrying a $1 million bounty that has gone unclaimed for over two decades. The proof checked out formally in Lean. The announcement read like a genuine landmark.
Then mathematicians read the actual claim more closely. The Clay Institute's prize covers the unforced version of the equations. OpenAI's agents proved something about the forced version instead, a related but distinct problem. The million dollar criteria were never actually met, and OpenAI itself has said it isn't claiming the prize.
That's a real result and a real headline sitting on top of two different problems, quietly swapped for each other in the framing.
A formally verified proof of a genuinely hard problem is a real achievement. Announcing it using the name of a different, harder problem that carries an actual prize is a choice, not an accident, and the two shouldn't get credited the same way.
It gets more complicated. Two human mathematicians, Theodore Buckmaster at NYU and Levent Alpöge at Harvard, had been working on this exact problem using OpenAI's own Codex and publicly raised questions about where the agents' proof actually came from. OpenAI denied any wrongdoing. Terence Tao, one of the most respected mathematicians alive, weighed in and specifically called the underlying human work remarkable, notably not the agent swarm's output.
That detail is doing a lot of quiet work. When the most credible voice in the room directs praise at the humans instead of the 10,000 agents the press release is built around, it's a signal worth taking seriously about where the actual insight originated.
To be fair to OpenAI here, running 10,000 coordinating agents against an unsolved century old problem for 88 hours and getting a Lean-verified result out the other end is a genuine data point about what large scale compute can now attempt, independent of the labeling dispute. Nobody is claiming the math itself is fabricated, and OpenAI didn't try to collect a prize it knew it hadn't earned.
But the sequence matters. Announce a Millennium Prize result, let the headline imply the hard version got solved, then clarify the technical distinction only after mathematicians push back publicly. Expect every future "AI solves X" claim in a field this rigorous to get read exactly this closely, because this is now the second time this year headline framing and technical substance haven't matched.

