How An AI math breakthrough ignited a controversy
pseudolus
213 points
228 comments
September 09, 2026
Related Discussions
Found 5 related stories in 73.9ms across 6,054 title embeddings via pgvector HNSW
- Controversy over OpenAI's Maths Breakthrough doubledamio · 34 pts · September 08, 2026 · 81% similar
- OpenAI fought dirty on career-making math problem jonbaer · 161 pts · September 08, 2026 · 66% similar
- Mathematics in the age of AI jonbaer · 142 pts · August 19, 2026 · 64% similar
- Major Math Breakthrough by AI mvrandswe · 11 pts · September 08, 2026 · 62% similar
- OpenAI Says It Has Cracked One of Math's 'Millennium Problems' doener · 13 pts · September 08, 2026 · 60% similar
Discussion Highlights (20 comments)
pseudolus
Extensive discussion on OpenAI's blog post on Navier-Stokes: https://news.ycombinator.com/item?id=49613262 . Quanta Magazine article that also discusses some of the controversy: https://www.quantamagazine.org/ai-has-solved-one-of-maths-1-...
paxys
> Navier-Stokes is one of six “Millennium Problems” on a list compiled by the Clay Mathematics Institute in 2000. Seven, not six. One is solved already, but is still a millennium problem.
jibal
> OpenAI, meanwhile, says its experience with Navier-Stokes could open the door to solving puzzles with more practical relevance. “We are now able to spend millions of dollars on a problem that we really care about and that really matters: developing new materials, finding cures to diseases,” Bubeck said. “All of those things that we have been talking about for a long time—now they seem to be at our fingertips.” Eh? There's no connection at all between the Navier-Stokes work and those things.
zero-sharp
Moving forward, I can't imagine other mathematicians wanting to have this kind of experience. So there has to be a shift away from these services.
snsr
OpenAI apparently used Buckmaster and Alpöge‘s work w/ Codex to bootstrap “their” dis-proof. https://cims.nyu.edu/~tristanb/statement.pdf
afavour
The core section: > However, communications quickly became contentious. According to Buckmaster, OpenAI offered to give him sole authorship on the Navier-Stokes solution—but only if Alpöge’s name was removed from the work and if the write-up would acknowledge the problem had been resolved by an internal OpenAI model. Buckmaster refused, in part because he was troubled by the question of what OpenAI's system had actually seen. For example, Buckmaster said the company did not initially give him a clear answer about whether its agents had access to the pair's logs on Codex (which is an OpenAI product). > OpenAI executives have denied that any employee or AI agent saw the pair’s work before the researchers released it publicly on 7 September. But there still remains a separate question: Could the pair's work have reached OpenAI's models through its training data? > OpenAI’s blog announcing the Navier-Stokes solution does not dismiss the possibility: “While unlikely, we cannot rule out that de-identified data derived from [Buckmaster and Alpöge’s] usage of our products helped improve our models .”
rrhjm53270
My impression: the re-aristocratization of scientific research seems inevitable.
ltononro
Does it matter who gets the credit at this point? Both used AI to do 99%+ of the work. So... do machines have ego?
rsfern
Regardless of what you think of the priority dispute issue discussed on sibling threads, I’m highly skeptical of the closing quote that this Navier Stokes result means that the same approach of casually spending a few million on agentic computation is going to solve end to end materials design or drug development. Those problems can’t be formally verified with an automated theorem prover. We have a lot of physics based simulation tools, but they tend to focus on small subsets of the full design problem and they make limiting approximations because otherwise they’d be too computationally expensive, or we just don’t have the right data to parameterize them beyond describing qualitative behavior. Agents are helping accelerate research in these fields but I think it’s mostly a different class of problem that’s a lot harder to specify and verify
elternal_love
Hmm, is the formal verification through? Lean just asserts no errors in the proof, but like can prerequisites not be fullfilled?
timmg
I think it was a pretty questionable thing to do by trying to front-run these researchers even if they didn’t make use of their techniques. The fact that they may have inadvertently “borrowed” their work via training data makes it much worse. OpenAI’s behavior here — even if you only consider [their] side of the story — was (at best) in bad taste.
lolakutty
Who should get credit? All the humans who ever worked to create the data. And the AI company for making the search program that searched through the data and found the solution.
sherburt3
"Breakthrough" to me would be like Isaac Newton inventing calculus to calculate pi. This feels more like $12MM in tokens was spent to add another digit to pi using the old way.
lemoncookiechip
This right here, or at least the thought of this, is why in the not so far future, businesses can't (won't?) be using these LLMs services. You cannot risk companies like Anthropic, OpenAI or their business partners like Microsoft having unfettered access to proprietary data on your company/businesses. It it likely that they or rogue employees will use the information to make a profit? It's pure speculation, but I'd say more than likely, and we will never hear about it or read it on the news unless there's whistleblowers in high enough positions to know about it. Assuming you and your employees aren't careful with what data you share, they will have intimate knowledge about your company from files and conversations logs. Likely personal user data too which they'll gladly create databases to link to and create extensive profiles on you, your employees and your businesses. It's not far-fetched to see them leveraging insider information shared with LLMs to play the stock market, leveraging data against competing businesses in other markets they might want to explore, and likely a bunch of other things that are escaping me right now as I write this. At the end of the day it's on those people for sharing such sensitive data, but it's not like these AI companies are innocent and won't gladly exploit every little byte of data without telling you, we know it happens.
harhargange
OpenAI has messed up big time here by competing with their customers. It would have become the norm for humans and mathematicians to use the tools and publish bigger results any way. If OpenAI didn’t run for credit, this theorem itself may have been proven by Buckmaster OR others in maybe a year or two. But now the bigger issue than AI solving problems is the issue of chat privacy, at the end of the day.
pama
Other than the undeniable breakthrough in math, the important point is the ability to orchestrate 10k agents to productively work on a single problem, which creates options: > OpenAI, meanwhile, says its experience with Navier-Stokes could open the door to solving puzzles with more practical relevance. “We are now able to spend millions of dollars on a problem that we really care about and that really matters: developing new materials, finding cures to diseases,” Bubeck said. “All of those things that we have been talking about for a long time—now they seem to be at our fingertips.”
jeanmichelselli
AI models and in particular LLMs are not capable of logic reasoning. See for example this paper: https://arxiv.org/abs/2506.06941 Ergo, they can't prove any theorem whatsoever. How do people at OpenAI expect that we believe in claims like that? This is yet another before-the-IPO stunt in my opinion.. Personally, I won't believe any of these claims until the community of mathematicians says otherwise.
VyseofArcadia
Regardless of the end result, OpenAI's behavior would be a career-ending ethics scandal for a human mathematician. This bit alone would be a career-ender. > According to Buckmaster, OpenAI offered to give him sole authorship on the Navier-Stokes solution—but only if Alpöge’s name was removed from the work and if the write-up would acknowledge the problem had been resolved by an internal OpenAI model. I wonder if an appropriate response from the mathematical community would be a good old-fashioned shunning. Mathematicians are allowed to use OpenAI's tools as much as they want, but no one with any current or prior OpenAI affiliation gets published in a reputable journal, ever.
elgertam
> “I certainly don't expect the industry to continue to spend millions of dollars to solve problems in mathematics, because there is no profit in it,” Columbia University mathematician Michael Harris wrote in an email to Science. But he worries the highly publicized achievement will be “extremely damaging to mathematics; it convinces decision makers that human mathematicians are obsolete, and it convinces young people that their passion for mathematics has no future.” LLMs seem particularly suited toward these existence-proof problems. Working mathematicians seem absolutely essential for universally quantified results, still. I strongly doubt, for example, that if Fermat's Last Theorem hadn't been proven three decades ago, that an LLM would be able to do work equivalent to inventing the mathematics as Andrew Wiles did to solve the problem. I have similar doubts about P vs NP, the twin prime conjecture, even the Riemann Hypothesis (unless the latter has at least one counterexample). And I want to be clear: I'm not downplaying the achievements of these models. This is remarkable! I simply think that the pattern of success is in existence proofs or finding counterexamples, which makes sense based on how LLMs function and are trained.
icepush
I have started to feel the sense lately, that first with the HF breach and now this plagiarism scandal, it is the straw that has broken the camel's back (so to speak). We have turned the corner and clearly entered the endgame - everything is going to unravel astonishingly quickly.