Navier-Stokes – Tristan Buckmaster [pdf]
Posted by procedurecall 8 hours ago
Comments
Comment by n2d4 5 hours ago
- Aug 15th: Tristan Buckmaster & Levent Alpöge make progress on a few important math problems, "finite-time blowup with smooth forcing for incompressible porous media, for Boussinesq, and for 3d incompressible Euler."
- they do NOT have a proof for the $1,000,000 Millenium Prize problem. BUT, they do claim to have a proof for a similar (non-Millenium) Navier Stokes problem that could help lead the way there
- Levent works at Anthropic, but this research was independent of his work there, with a mix of GPT and Claude models. Tristan is not related to Anthropic.
- Early Sep: Rumor spreads to OpenAI that Anthropic solved a major problem. Tristan emails OpenAI to clarify, without revealing the problem they solved or how they did it.
- After hearing of the rumor, OpenAI started researching Navier Stokes with a new internal model.
- Sep 6th: OpenAI's Sebastien Bubeck tells Tristan that they solved the $1,000,000 Millenium Prize Navier Stokes problem. The approach is very similar to Tristan & Levent's approach to the non-Millenium problem.
- Tristan is suspicious of the timing, as only few others were trying this approach. OpenAI says the model didn't access his user data directly, but leaves unanswered whether Tristan's chat conversations were part of the training.
- OpenAI says they would partially credit Tristan for the $1,000,000 discovery (even though Tristan did not solve the $1,000,000 problem) — but only if they remove Levent as an author, as he works for Anthropic.
- Sep 8th: Tristan refuses to remove Levent, and rushes to publish their results independently.
Comment by singularity2001 1 hour ago
Comment by bhouston 1 hour ago
But the drama here is a little important. Stealing the millennium prize for N-S is sort of a big deal, especially to those who had been working on it for the last few years.
Comment by dandanua 45 minutes ago
Comment by tecleandor 3 hours ago
Comment by onidj 2 hours ago
Comment by simianwords 21 minutes ago
Comment by gcr 7 minutes ago
1. Less than a hundred people in the world are working at this problem, 2. A significant fraction of those happen to work at competing hyperscalers, 3. Those hyperscalers repeatedly show themselves not to take user privacy seriously
Comment by simianwords 32 seconds ago
> Those hyperscalers repeatedly show themselves not to take user privacy seriously
where? Any examples?
Comment by qnleigh 6 hours ago
I'm not even sure what to say to this, but I think this should be widely known if it is indeed what happened.
Comment by Keyframe 6 hours ago
Comment by pred_ 5 hours ago
Comment by xdavidliu 7 minutes ago
Comment by Keyframe 4 hours ago
Comment by Certhas 3 hours ago
Comment by robinhouston 3 hours ago
> A series of false and inflammatory allegations against me are currently circulating on social channels. To clarify, I came into the discussion following academic norms, and I'm disappointed that it has come to this. Anyone who knows me knows that academic standards are of the highest importance to me. Will have more to say tomorrow.
Comment by oldgradstudent 14 minutes ago
Comment by irthomasthomas 1 hour ago
Comment by macleginn 6 hours ago
Comment by tristanj 2 hours ago
First, this is an unnamed OpenAI employee speaking, not OpenAI the organization.
Second, you miscomprehended the article. The employee did not "threaten to ruin a prominent researcher's career". The actual quote is "Why would you ruin your career?", which implies the researcher would damage their own career, i.e. via self-sabotage.
Then the actual "threat" is "If you don’t want me to be nice, then I don’t have to be nice” which is an entirely different statement.
Comment by garyfirestorm 1 hour ago
Proceeds to ask removal of another coauthor or else we totally discredit you - phrased as why would you do this to yourself.
Comment by tristanj 21 minutes ago
From the article, it seems OAI wanted to continue discussing the situation with Buckmaster and reach a resolution, but Buckmaster did not want to, declined to respond, and published first.
Also keep in mind we've only heard one side of the story, so any interpretation of events so far is incomplete. There should be a lot more information from OAI's side coming out later today.
Comment by gcr 11 minutes ago
Comment by macleginn 1 hour ago
Comment by Davidzheng 2 hours ago
Oai offers two options, the second Tristan views as dishonest. But Tristan rejects the first, why? Because he thinks it's theft? But then why would OAI threaten him?
Comment by tristanj 1 hour ago
OpenAI claims to already have a full proof (which they produced in the past 5 days after the rumors leaked). Hence the dispute.
What I find interesting is the timeline of when he found counterexample for NS is very unclear. Did Tristan find a counterexample weeks ago or was it very recently? Was it after OpenAI solved it? The wording is intentionally vague.
Either way, there was a massive rush to publish these results.
Comment by Ar-Curunir 1 hour ago
Either put up some evidence-backed arguments, or shut up.
Comment by ummonk 5 hours ago
Whether the Codex sessions could have indeed made their way into Astra training data is something I can only speculate on though.
Comment by Semkas 5 hours ago
Comment by curt15 39 minutes ago
Comment by az226 5 hours ago
No looksies, wink.
No trainsies, wink.
Comment by ngomez 4 hours ago
[0] https://openai.com/policies/how-your-data-is-used-to-improve...
Comment by tempfile 3 hours ago
If OpenAI did use the conversations from Buckmaster and Alpoge, then not disclosing it, explicitly, is plagiarism. If they planned to use that plagiarism to pressure the authors to publish, that is even more unethical. What the terms of use say does not make it any more or less ethical.
Comment by YeGoblynQueenne 3 hours ago
Comment by tempfile 1 hour ago
> Why the downvotes?
I think there was only ever one. Not sure why.
Comment by YeGoblynQueenne 1 hour ago
About the plagiarism issue, I model it as OpenAI being an advisor and their AI a PhD student. If the advisor puts their name on a paper behind that of their PhD and it turns out the PhD copied the text of the paper from somewhere else the advisor is also responsible of plagiarism, not just the student. The least the advisor can do is withdraw their authorship from the paper.
But, yeah, point well made: it could be much worse than that. Like an advisor instructing a student to copy someone else's paper.
Comment by tempfile 49 minutes ago
Personally I don't like thinking of LLMs like a PhD student, because most PhD students remember where they learned things from, while LLMs essentially cannot. I think of it a bit more like someone using a search tool carelessly. Although in this case it is apparently more like deliberate misuse than carelessness.
Comment by senordevnyc 4 hours ago
Comment by hodgehog11 3 hours ago
Comment by Certhas 3 hours ago
Comment by dandanua 21 minutes ago
Comment by wheresyourat 1 hour ago
That dude shows up in every thread that even implies genai uses user input for “market validation” to defend the poor defenseless genai corpos. (see figma/claude)
Must be circling the wagons to get their stories straight.
Comment by calf 3 hours ago
Call this behavior what it is, technofascism. Another comment compared it to the Godfather. To think this is the 21st century and academics are still horrible human beings.
Comment by Ar-Curunir 1 hour ago
Comment by IAmBroom 1 hour ago
Or are you stating that the researcher is a "horrible human being" for refusing to allow AI?
Comment by mayakacz 7 hours ago
OpenAI looked at user data, stole world class researchers' work, and then tried to threaten those researchers to do what would make their corporation profit (which they would anyways!).
Imagine you have been working on a terribly difficult math problem for a decade. This is a result you have spent years on, and what you will likely be remembered for. And to have some punk from OpenAI lie to you, threaten you, and tell you that they are willing to go on the record that you "deserved" it? What is this, the Godfather?
If OpenAI solved Navier-Stokes, that is an astounding result! - yet they'll still be remembered as those who thought credit was more important than results. That winning was more important than collaboration. If this is true, they're burning any trust left with academia.
Comment by postalcoder 6 hours ago
OpenAI is no stranger to rivalry with Anthropic but 1. it's not like user data is sitting around on some kitchen table somewhere and 2. I consider OpenAI to be as economically motivated as any other actor in this space and playing around with user data like that would destroy their business.
There are things that Buckmaster alleged and things that he speculated. The entire training data thing is speculation. If this is pissing you off, then you ought to evaluate how you ingest information.
Comment by dash2 5 hours ago
> I asked whether the model had been trained on, or had access to, our sessions in Codex, into which we had been putting all our drafts for the whole of this project. I was told the model did not look up user data. I asked again, about training, and I did not get an answer.
The shocking/interesting thing would be if it was trained on the sessions. I think it's very implausible that they gave the model access to someone else's sessions as input. That would be a huge privacy violation and would probably blow up a large proportion of their enterprise business.
Does openAI train on user conversations in general? I assume so. But so fast as that? That seems unlikely in general. I expect OpenAI will come out denying this.
Comment by Traster 5 hours ago
> I was told the model did not look up user data.
The naive way to read this is "Nothing you guys did influenced the way our model got to the solution".
The less naive way to read this is "Of course the model isn't looking up your user data. I (the guy trying to blackmail you to remove the Anthropic employee from credit on your paper) looked up your sessions, and tipped our model off on how to solve this problem".
Comment by YeGoblynQueenne 2 hours ago
It's a bit like sending unencrypted messages through a messaging app and the developer having a TOS that says they don't look at your messages. They might not, but they are fully capable of doing so. If they have a reason to do it, they will. Nobody's stopping them.
Comment by YeGoblynQueenne 2 hours ago
How "fast" does it have to be? Buckmaster and Alpoge have been working on this for just a day short of a year. See Alpoge's tweet announcing his collaboration with Bukmaster dated 9/19/25:
https://x.com/__alpoge__/status/2097206973418611054
It takes a few months to train a model these days but not a whole year. OpenAI had all the time to train on Buckmaster and Alpoge's results of just a few months earlier at which point they must have been well on the path to their result.
Comment by revolvingthrow 5 hours ago
There’s potentially trillions on the line, do you seriously expect those companies to adhere to laws and regulations any more than, say, uber?
The only unlikely part is the timeline - your sessions from a week ago probably haven’t made their way into the model. It’ll just take a while longer, and will be massaged just enough so that it isn’t really your exact session word for word so you can’t sure as easily.
Comment by blini-kot 1 hour ago
I also don't get why it was downvoted, other than due to people not reading past the first sentence - although in the modern world's attention deficit that is understandable too
Comment by actionfromafar 5 hours ago
Comment by johnnienaked 5 hours ago
Enterprises are well aware of it and are fully on board. You didn't think every corporation in America has an OpenAI subscription because the models were good, did you?
The whole reason they have subs is to train them on YOUR WORKFLOWS lol
Comment by howdareme 5 hours ago
Comment by irthomasthomas 1 hour ago
Comment by xdavidliu 2 hours ago
Comment by paxys 2 hours ago
Comment by FiberBundle 5 hours ago
Comment by defmacr0 1 hour ago
Comment by thereitgoes456 5 hours ago
Comment by johnnienaked 5 hours ago
Comment by az226 5 hours ago
Comment by Sol- 3 hours ago
This doesn't seem to be clear and is very implausible for a large company. Be as cynical as you want, but a normal researcher will simply not have access rights to this data, which will be siloed away somewhere else.
It might very well be somewhat unfair to catch wind of a promising approach and then try to frontrun them by throwing compute at the problem, but this isn't really the same.
Comment by cududa 2 hours ago
Comment by alas44 3 hours ago
"I asked whether the model had been trained on, or had access to, our sessions in Codex, into which we had been putting all our drafts for the whole of this project. I was told the model did not look up user data. I asked again, about training, and I did not get an answer."
Let's see what statement OpenAI will come up with for their side of the story
EDIT: precised my thought on user data vs session data
Comment by tristanj 6 hours ago
And simply knowing a problem can be solved is half the battle.
Comment by dbdr 5 hours ago
The route to the Clay problem through a
smooth force, options c and d in Fefferman’s statement of the problem, is the
route Luis and Diego opened and the one Levent and I had quietly chosen to
attack. Almost nobody else I know of was working on it. It is not the direction
one arrives at in a few days by giving a model the problem statement. When I
heard “forced,” it was a bright red flag.
This is much more than the knowledge than the problem can be solved, it's also the specific, non-obvious approach to solving it. That's much more damning for OpenAI, if confirmed.Comment by tristanj 5 hours ago
And the article states "an insane amount of compute had been used," which implies OpenAI brute-forced their way to a solution. I.e. they searched for every paper published on Navier-Stokes and exhaustively attempted every approach. Such an approach would lead them to a solution.
There is not enough information at this time to reach a conclusion. The best option is to wait for statements from both sides, then reevaluate.
Comment by mishellaneous 2 hours ago
the post you were replying to quotes Buckmaster specifically denying this: "It is not the direction one arrives at in a few days by giving a model the problem statement."
> And the article states "an insane amount of compute had been used," which implies OpenAI brute-forced their way to a solution. I.e. they searched for every paper published on Navier-Stokes and exhaustively attempted every approach. Such an approach would lead them to a solution.
"implies" is a surprising choice of word here. that's certainly one interpretation of "an insane amount of compute had been used". what came to my mind, considering Buckmaster's statement that the AI would not head down this specific path on its own, is, though, that they prompted it in this specific direction and then used an insane amount of compute. this seems consistent as well with these other statements:
> Over the course of the call, as members of their team sent Sebastien corrections and details over their internal chat, it emerged that an entire team had been working on the problem, that this was one of a number of things that was tried, that work had started on the unforced problem, that the team first set the model on easier problems, including Euler (...)
Comment by golol 1 hour ago
To be honest, he was referring to routes (c) and (d) to the millenium problem, as far as I understand no more specific. Which is 2/4 routes.
Comment by mishellaneous 15 minutes ago
> The route to the Clay problem through a smooth force, options c and d in Fefferman’s statement of the problem, is the route Luis and Diego opened and the one Levent and I had quietly chosen to attack. Almost nobody else I know of was working on it. It is not the direction one arrives at in a few days by giving a model the problem statement. When I heard “forced,” it was a bright red flag.
so, you're saying that "the direction one arrives at in a few days" is (A) "the route through a smooth force, options c and d in Feffermann's statement of the problem", and no more specific than that, i.e. does not necessarily include (B) "the same path route as Luis and Diego" (quoted from the post i was replying to) (which, as i understand, is a subset of A -- directly from Buckmaster's quote: "the route A is is the route Luis and Diego opened")?
but the post i was replying to claims that "An AI model could independently choose the same path route as Luis and Diego, without access to Buckmaster and Alpöge’s work", i.e. that the AI model could independently chose B. but choosing B implies choosing A, since B is a subset of A. and in this case it is irrelevant whether Buckmaster claimed that an AI could not independently choose A or B -- the point which you seem to be contesting.
EDIT: my understanding is that "solving the Navier-Stokes existence and smoothness problem" consists of proving at least 1 of 4 precise statements ("options (a) through (d) of Fefferman's statement of the problem"), and Luis and Diego's work were developments towards a proof of statements (c) and (d), which have been recently further expanded by Alpoge and Buckmaster
Comment by shrx 2 hours ago
Can they back up this statement somehow?
Comment by mishellaneous 2 hours ago
Buckmaster did also mention, for example, that a team of people was employed to solve the problem, which supports this claim. but that is another claim whose veracity could also be questioned. but at some point we must trust other people, unless we can be satisfied with only believing what we personally see.
(also, IMO, the coincidence of both discoveries in time is pretty suspicious. this one doesn't need you to trust many people i guess)
Comment by Davidzheng 2 hours ago
Sorry for random rant but I don't think these statements help your point
Comment by seeall 1 hour ago
OpenAI said there [1]: > The method by which the problem was solved is also notable. The proof brings unexpected, sophisticated ideas from algebraic number theory to bear on an elementary geometric question.
[1] https://openai.com/index/model-disproves-discrete-geometry-c...
Comment by Keyframe 5 hours ago
Comment by Ar-Curunir 1 hour ago
Have you done any mathematical research? If not, then no, knowing that a problem is solvable is not “half the battle”.
Homework problems are all designed to be solvable, yet they can vary greatly in difficulty. Research mathematics is even more extreme, because, unlike with homework, you don’t know that it is solvable with the extant mathematics, and you might need to invent new maths.
Comment by tristanj 49 minutes ago
If not for the rumors that A/ had already solved NS, OAI would likely never have pursued solving the problem with such fervor. The rumors drove OAI to assemble an entire team to crack this.
Comment by johnnienaked 5 hours ago
Comment by scurnus 6 hours ago
OpenAI tried to collaborate and share the results together with a fixed timeline, to avoid this mess but it was inevitable. There is a conflict of interest, where the other researcher works at Anthropic, who will also try to take credit.
Where they may be in the wrong is if they took user data regarding the problem, how will we know if they did or not?
Comment by Ar-Curunir 1 hour ago
That is not collaboration, and is not an academic norm.
Comment by sk4rekr0w 6 hours ago
Comment by tigershark 6 hours ago
Comment by famouswaffles 6 hours ago
Comment by PowerElectronix 5 hours ago
Comment by famouswaffles 4 hours ago
That OpenAI heard about the result and decided to throw a lot of money at it knowing it was within reach ?
That the model took an approach that was published years ago ?
Comment by calf 4 hours ago
Comment by tristanj 2 hours ago
Comment by famouswaffles 4 hours ago
So either the rumors or Tristan's contact would explain it fine.
Comment by calf 3 hours ago
The whole point of Tristan's first email was to address the issue of the rumors so in fact this explanation confounds several things (in terms of the mutual knowledge of the conversants and their intentions).
Based on this lack of understanding it is pointless to continue this thread
Comment by tigershark 6 hours ago
Comment by famouswaffles 6 hours ago
Comment by tigershark 5 hours ago
Yes, I read fully the statement, what about you? Do you know what is a threat? What is this in your super-humble opinion if not a threat:
> The reply was, “Why would you ruin your career?” I replied that I am an academic, and asked why he thought going public would ruin my career. The reply was, “If you don’t want me to be nice, then I don’t have to be nice.”
And you are saying that I don't have reading comprehension...
Comment by sk4rekr0w 5 hours ago
Comment by tigershark 5 hours ago
> The reply was, “Why would you ruin your career?” I replied that I am an academic, and asked why he thought going public would ruin my career. The reply was, “If you don’t want me to be nice, then I don’t have to be nice.”
Comment by thereitgoes456 6 hours ago
Comment by johnnienaked 5 hours ago
Comment by paxys 2 hours ago
Comment by traes 5 hours ago
[0] https://xcancel.com/SebastienBubeck/status/20972141224714323...
[1] https://xcancel.com/polynoamial/status/2097215233119211902
[2] https://xcancel.com/danintheory/status/2097214838003138603
[3] https://xcancel.com/_sholtodouglas/status/209721833169057800...
Comment by Recursing 4 hours ago
Comment by mitchellgoffpc 4 hours ago
Comment by keeeba 5 hours ago
Comment by pdantix 5 hours ago
Comment by andrewchambers 2 hours ago
Comment by pred_ 2 hours ago
Also note the lower post ID in the URL.
Comment by paxys 2 hours ago
Comment by pred_ 1 hour ago
Comment by timmg 1 hour ago
If I needed to get my story straight, well, it might take a little time and coordination…
Comment by Semkas 5 hours ago
If compute is cheap, and the difficult thing with scientific discovery is now mostly in steering agents into promising areas, there's an obvious incentive for OA mathematicians to simply monitor closely which researchers are close to releasing exciting results, make some assumptions about their prompts based on their past work, and quickly prompt their own (stronger) model to look into the same areas.
Comment by AJRF 5 hours ago
Why leave that aside? That is _the_ story.
If a Chinese research lab did this we'd call it espionage.
Comment by Semkas 5 hours ago
If they're going to try to beat researchers to discoveries like this it disincentives researchers to talk about their progress publicly, and basically breaks the ecosystem of scientific cooperation / discovery. It's also immoral.
Comment by modeless 5 hours ago
Comment by FiberBundle 5 hours ago
Comment by Semkas 5 hours ago
Comment by ehhthing 2 hours ago
For math, a field that is built on incremental research it feels like AI labs will do nothing but discourage publishing research at all for fear that they will be able to spend the money for compute that publicly funded academia simply cannot afford.
It feels like publishing anything at this point just means that your work will be fed to a machine that will make sure your work will never been seen by anyone else because it will always be the ones making the “true advancements”.
Perhaps I’d feel better about this if AI labs really existed for humanity’s benefit, but for some reason I don’t think that comes up in their investor slide decks.
Comment by stabbles 3 hours ago
Comment by ummonk 5 hours ago
Comment by taylorfinley 7 hours ago
"Oops! We really did mean it when we said we wouldn't train on your data. Our models are just so good they decided to anyway."
Comment by hnfong 1 hour ago
"Let's crawl the social media of prominent mathematicians in this field to see if we can copy/steal any ideas for low hanging fruits"
That actually might get you quite far already.
Comment by JuniperMesos 6 hours ago
Comment by Prophet6-0091 5 hours ago
Comment by margorczynski 4 hours ago
Comment by Maken 3 hours ago
Comment by tristanj 7 hours ago
Mathematical explanation by Terrance Tao: https://mathstodon.xyz/@tao/117233527638291447
It seems there is much background drama behind this, and this is what I've pieced together of what happened:
Over the past year, Buckmaster and Alpöge have been using AI to work on fluid dynamics maths problems. Alpöge works at Anthropic, which will cause future issues.
In mid-August, they found a counterexample for a simpler version of the Navier-Stokes problem. They spend the next few weeks preparing their paper.
In early September, rumors start spreading on X that Anthropic has solved a Millennium prize problem (and that it's Navier-Stokes). Buckmaster reaches out to OpenAI to explain this is their own personal research, not an Anthropic project.
A few days later, OpenAI gets back to him, and tells him an internal model found has a counterexample for Navier–Stokes, potentially worth the $1 million Millennium prize. The proof uses the same method that Buckmaster and Alpöge chose to work on. They don't show him the proof.
Buckmaster pressed them for more details. OpenAI reveals they had an entire team had been working on the problem, and that they started work in the past few days, after the rumors that Anthropic had solved a Millennium prize problem.
Buckmaster says OpenAI talked about a shared publication timeline. They want to Buckmaster to publish first, then give Buckmaster shared credit for the Millennium Prize when they publish the full result. But they want to exclude Alpöge as an author because he works at Anthropic. An agreement is not reached. Buckmaster had been using OpenAI Codex to draft/check his work, and asks if his private AI chats were used to accelerate OpenAI's result.
Buckmaster and Alpöge think they have found a counterexample for Navier-Stokes, but the paper is not yet presentable. It's unclear what date they found this result.
Because of the situation with OpenAI, they published their existing papers earlier than planned (today), alongside this statement announcing they have a tentative result on Navier-Stokes and revealing the OpenAI drama.
The post is missing context from both sides, and this isn't my field, so hopefully someone else can unpack what's happening here.
Comment by akersten 7 hours ago
Skimming the PDFs it seems much more dramatic than that? It sounds like at least one of them is concerned OpenAI "solved" the problem by having their internal model use the chats of the independent researchers and want to claim the credit instead? I don't know. The tone is pretty accusational though:
> the one Levent and I had quietly chosen to attack. Almost nobody else I know of was working on it. It is not the direction one arrives at in a few days by giving a model the problem statement. When I heard “forced,” it was a bright red flag.
> I was shown a prompt and told the internal research model had simply been given the problem statement. Levent had been told by Sebastien “very little human input” had been used. This turned out not to be true. Over the course of the call, as members of their team sent Sebastien corrections and details over their internal chat, it emerged that an entire team had been working on the problem, that this was one of a number of things that was tried, that work had started on the unforced problem, that the team first set the model on easier problems, including Euler, that even the prompt that had been shown to me had been written by prompting Codex, and that an insane amount of compute had been used.
> I asked when the first prompt had been sent by them. This question was not answered directly by OpenAI for some time. Eventually it was agreed that it had been sent in the past few days, after information about our work had reached OpenAI.
> I asked whether the model had been trained on, or had access to, our sessions in Codex, into which we had been putting all our drafts for the whole of this project. I was told the model did not look up user data. I asked again, about training, and I did not get an answer. [0]
Comment by tristanj 7 hours ago
Comment by lhd1 2 hours ago
Comment by rf_physics 1 hour ago
If what he wrote is accurate, it does suggest that OAI is effectively extremely hostile to cutting edge researchers (eg, if we hear rumors about your partial success on a problem that has huge PR benefits, then we'll assemble a strike team of researchers with unlimited compute to claim the win for ourselves, possibly by training on your data). It's also not a good look for them to request author removals based on company affiliations.
I think what you have in mind is more appropriate for more normal corporate projects and the like. But academic collaborations are not usually so political/'profit' driven, if that makes sense.
Comment by defmacr0 1 hour ago
Presumably because this was something Levent did in his spare time and because it was not obvious that this work would eventually lead to a breakthrough.
> Why did Tristan use OpenAI's models when it should have been known was a potential outcome?
I'm sure in the past he had less cynical feelings about OpenAI and their penchant for academic fraud.
> I understand they wanted a normal math collaboration but presumably what Levent brought was his resources (as far as I can see Navier-Stokes is not his speciality)
I think you're not giving the guy enough credit in saying that his contribution came down to having an API key for Anthropic models.
> Normally these things are hashed out formally beforehand to avoid the sort of thing now happening.
How would that have helped? That agreement (which may well still exist) would not have involved OpenAI.
Comment by cma 6 hours ago
Comment by kzrdude 5 hours ago
While common sense reminds us here that if you send your data to an external entity’s computer, you are no longer in control of said data. The lines have blurred here clearly over the last decade, but that should have made the theory yet more clear to everyone involved: your data will be vacuumed up unless you keep it sealed. Use your own computer if you want to be in control.
Comment by cma 13 minutes ago
I'd prefer things be opt in, and especially not start opt out, then try to trick you opt in with a popup defaulting to opt-in, like Anthropic did on consumer plans, but if they submitted anything on an opted-in plan it's not reasonable to be mad it trained on it.
Even still, I also believe for significant reasons that OpenAI would ignore the opt-out in selective cases and could be in the wrong here.
And the threats and terms they offered seem wrong either way, pending more context.
Comment by instagraham 5 hours ago
Why is OpenAI chatting with him at all at this stage? Is the discussion along the lines of "hey we used the work you are famous for to do a bigger piece of work, just thought you should know" or "heyyy....so we kinda liked what you were typing in your private chat, and thought we'd develop those ideas a bit. and yeah we solved Navier-Stokes in the process. But it's our finding, so do you want like an honorary acknowledgement or do you want to go to court?"
Comment by traes 5 hours ago
It seems like the timeline according to OpenAI is that:
1. Buckmaster developed a counterexample to a reduced version of Navier-Stokes with Anthropic employee Levent
2. Rumors start spreading that Anthropic has solved Navier-Stokes
3. OpenAI learns this and starts throwing a ridiculous amount of compute at it, now knowing it's within reach of LLMs
4. Their LLMs (with human assistance) get FARTHER than Buckmaster, using the exact same method.
5. OpenAI reaches out to Buckmaster to negotiate a fair way to publish both results and properly assign credit
Perhaps they simply and honestly feel he is owed credit. I can't help but imagine it at least played some small part. I expect that's what Bubeck is going to claim: https://xcancel.com/SebastienBubeck/status/20972141224714323...
Comment by denverllc 1 hour ago
That is absolutely *ridiculous* in academia to deny authorship because of affiliation of the author worked on a substantial portion. You’d be ostracized because nobody would ever want to work with you again.
Comment by dumberquestions 7 hours ago
Someone correct me if I'm wrong, but the work involved here is not the actual millennium problem, but it concerns versions with an added external force that the author thinks is a path that may help toward solving the harder unforced problem.
Comment by modeless 6 hours ago
The question is whether OpenAI's pursuit of this direction happened spontaneously, or as a result of them learning about Tristan's work somehow. To be clear, while the tone of this post seems quite accusatory, Tristan does not claim to know for sure whether OpenAI unfairly benefited from his work. Sholto Douglas from Anthropic is also on record saying the suggestion that OpenAI used Tristan's codex transcripts somehow is extremely unlikely to be true[1], which I agree with, though it doesn't rule out them learning of Tristan's work some other way. I am sure OpenAI will have a statement out tomorrow clarifying their position.
Comment by dumberquestions 5 hours ago
Comment by tristanj 5 hours ago
Comment by defmacr0 52 minutes ago
Comment by robotpepi 3 hours ago
That's beyond naive. The money this would mean for OpenAI (and the money they've already spent)...
Comment by dumberquestions 4 hours ago
Comment by MathmoKiwi 4 hours ago
Comment by MathmoKiwi 4 hours ago
So to be fair, if Anthropic is *also* doing this (quite likely!) then Sholto would have a very strong incentive to try and spin it as highly unlikely that any of the big AI labs are possibly doing this.
Comment by Recursing 6 hours ago
I guess this is true in more ways than one. Kasparov famously accused IBM of cheating during the match, by spying on his preparation (edit: though the main cheating accusation was live human intervention during the games, on top of IBM downplaying the heavy human involvement behind the AI, which also mirrors this situation)
Comment by sk4rekr0w 5 hours ago
Comment by 20k 6 hours ago
If you think these companies are not training on your prompts you are incredibly naive. These models were built by stealing and pirating literally everything they can get their hands on no matter the legality. AI companies are always very specific about what they're not doing - in a way that you can drive a truck through the loopholes
Comment by coliveira 5 hours ago
Comment by wrasee 6 hours ago
But I would love to see more informed insight/discussion on this.
Comment by chvid 6 hours ago
The money in nerdy frontier math is very little. The money in Big AI is very very much.
So the deal is this: We will pay an army of you guys very well and you will get to work on your favorite problems. The only thing is if you find something you will have to credit the Machine God.
Do you think you can handle that?
Comment by 0x5FC3 5 hours ago
I said something similar a month ago.
Comment by nxobject 5 hours ago
Comment by coliveira 5 hours ago
Comment by pred_ 5 hours ago
It seems clear now that mathematical results can be traded on some kind of obscure market made by the frontier AI labs.
I suppose it could go the other way too: “Dear Bubeck, how much will you pay me to not write that I did this with GLM-5.3?”
Comment by simpetre 44 minutes ago
Comment by instagraham 5 hours ago
For eg: "Hey ChatGPT my name is X and I am 6 and a half feet tall. Am I anaemic?" This is a query, and while it might suggest to an AI model that tall people may worry about iron deficiencies, it's not really necessary to include in training. The user may be tall or short, but the idea that one may randomly ask about anaemia is not exclusive to this dataset. At best, this chat is an example of linguistics, not anything else, and the models figured out how to write and answer such questions years ago. It is ignored in training.
But when your work involves solid complex and unique mathematical proofs, the data is suddenly worth training upon. If I understand it correctly, the LLM may view your approach as a brand new path to take to solve an otherwise intractable problem. Its reinforcement training emphasises that it should do this in order to improve. And since it leads to results - large internal teams likely flag the model that reached this stage, the model is rewarded and given compute and attention - it is a desireable outcome both for the model and for OpenAI.
OFC, OpenAI becoming an advertising company will suddenly have incentive to treat all data as valuable. But while they are a "we need to make headlines" company, it's more rational that they view these examples of data as more valuable than others.
I don't doubt that they trained on his chats. This seems like the ideal usecase for "mass surveillance but using training" as a sort of filter.
But even so, one wonders how the model differentiates. If the researcher entered proofs into ChatGPT every day that mentioned "strawberries", while no other math paper on the topic did so, does that mean their chats would be audited?
Comment by defmacr0 29 minutes ago
Comment by coliveira 5 hours ago
Comment by camel-cdr 4 hours ago
Comment by harhargange 6 minutes ago
Comment by aaraujo002 1 hour ago
https://mastodon.social/@tristanbuckmaster/11723341370570119...
And here are Terence Tao’s comments on the results: https://mathstodon.xyz/@tao/117233527638291447
Comment by nopinsight 3 hours ago
"Strong agree. I know that there is rivalry between the labs but it's important that we learn to work together given what's coming. <quote tweet [1] above>" -- Noam Brown, an OpenAI researcher [2]
We all should heed the implied warnings of these top researchers about what's coming. The world is far from ready and everyone who can should pitch in.
[1] https://x.com/_sholtodouglas/status/2097224624274911368 [2] https://x.com/polynoamial/status/2097225279366414541
Comment by goldenarm 5 hours ago
Key quote : "Solving the problem by purely AI-powered methods [would be a] net negative for the progress of mathematics."
Comment by porridgeraisin 2 hours ago
We are basically mass manufacturing math. Just like you have just 100 designers for a product selling millions of units, you will now need 100 mathematicians to make millions of advancement. Yes you have factory workers, but if we are being realistic they have negative leverage in the world and the analogue of that is not something most of today's mathematicians would want to do. They would want to be in the 100.
Like Tao says, each advancement is now significantly less useful since it yields fewer usable objects. However, we will get many many advancements. Is the tower made with many worse bricks better or worse than the tower made with a few amazing bricks? Depends on the tower. And time will tell.
For some fields of math and some of it's usecases, economies of scale will be positive ROI overall. In others it won't. But we will know which is which only after it's been fully scaled up, which will take 10-15y in my estimate.
Some feel that in the majority of usecases it is negative ROI, some feel the other way, but that opinion is for practicing mathematicians like Tao to hold. Also, some opinions on either side are held in the context of a particular field or practice, and should not be interpreted generally.
Comment by martijn_himself 4 hours ago
Comment by mishellaneous 2 hours ago
i always found it fascinating how existence and uniqueness of solutions for the basic types of PDEs (Laplace, wave, heat...) follows from boundary conditions of just the right type intuition tells us, i.e. either value or derivative for Laplace (corresponding to fixing voltage or charge on the conductors), both value and derivative for wave (corresponding to initial position and velocity of the parts of the string, as we'd expect from classical mechanics), and also something about the solutions for the heat equation being unstable for negative times (which totally makes sense when you think of "diffusion" -- can't unmix it).
Comment by redox99 2 hours ago
Comment by contubernio 3 hours ago
For those who know nothing about the context - the Diego mentioned was a student of Fefferman and Luis was a student of Diego's - these people have all worked hard on these problems for a long time and are genuine experts. The mathematicians at OpenAI are strong mathematicians, but not expert on these particular problems. The particular approach is claimed to be the key to the whole thing.
The allegation is not different in spirit to alleging that a particular group of astronomical researchers "discovered" a new planet because they had access to the logs of another group that had already pointed its telescope at the planet.
This post is not intended to assess the correctness of the allegation.
Comment by jarbus 7 hours ago
Comment by YeGoblynQueenne 3 hours ago
This sounds like a very big coincidence and it looks really bad for OpenAI but there is an alternative explanation that I can only state as a conjecture.
Suppose that the ability of LLMs to generate mathematical proofs is like a quiver full of arrows: each arrow, one proof. The same quiver is shared between all instances of one model and substantially similar models share substantial subsets of the arrows in the same quiver.
That would allow two independent teams to converge on the same LLM-aided solutions to the same problems. Even more likely so if the quivers were small and finite and their arrows were specific to a distinct class of problems (without being able to suggest a particular class from what we've seen so far).
This would explain the kind of LLM-mediated results we've seen so far that tend to be ... sparse. By which I mean that every time there's a new model release we get some new results and then they seem to dry out, until the next release.
It would also explain how OpenAI was about to prove the same result as Buckmaster and Alpoge, while absolving OpenAI of any misconduct. And this is one reason to prefer this explanation: one should not favour accusations of misconduct as long as there are conceivable alternatives.
But, that's just a conjecture that I can't prove.
Comment by wwind123 5 hours ago
Comment by antonmks 6 hours ago
Comment by ggcr 3 hours ago
Stadlmann improved it from 246 to 240, OpenAI later claimed 186 I think?
Maybe someone can help clarify? I am no expert at all, but I can't help but see similarities.
Comment by kzrdude 3 hours ago
Comment by kingkandu 1 hour ago
but they probably won't and if it's happened on some obscure math research it's happening everyday everywhere else.
Fully local AI compute can't come fast enough, these guys have IP theft baked into their bones.
Comment by lz400 3 hours ago
Comment by harhargange 6 hours ago
This is significant.
Comment by world2vec 5 hours ago
Comment by amai 1 hour ago
https://www.quantamagazine.org/computer-helps-prove-long-sou...
Comment by bjenik 6 hours ago
Terry also talks about it https://mathstodon.xyz/@tao/117234157753860650
Comment by paxys 2 hours ago
If I publish something, and disclose that I used AI for assistance, do I have to credit everyone who previously used the same AI to try the same problem? Because their prompts inevitably made it to the training data for my prompts?
Comment by Davidzheng 2 hours ago
Comment by tecleandor 1 hour ago
Comment by ggcr 6 hours ago
Comment by anon109 1 hour ago
Comment by aqsnow 12 minutes ago
Comment by Traster 4 hours ago
"I don't want to live in a world where someone makes the world a better place, better than we do."
It's amazing how transparently OpenAI is running the standard silicon valley playbook.
Comment by Tycho 1 hour ago
Comment by auggierose 3 hours ago
Comment by epsteingpt 6 hours ago
Time will almost certainly reveal a lot more about the drama and the related ethics, but let's get excited about the actual breakthrough as well!
Comment by nullbio 5 hours ago
What I'm curious to know is whether this was a manual snooping, or automated farming that occurs for anything of value that happens in chats.
Comment by margorczynski 4 hours ago
I'm just wondering how much real input Buckmaster gave here that he thinks the proof is his. I guess at the end of the day OAI still wins if ChatGPT was used to prove this successfully.
Comment by achierius 7 hours ago
- the OpenAI researchers claimed that they had "just told it to work on the problem" with little human input
- in fact, they had a whole team working on it
- and used, among other things, the work of third party human researchers to drive the work
- then threatened? a researcher who tried to go against theit planned narrative
Just from this document (which is of course only one side of the story) it really sounds like OpenAI was hoping to publish and say "we just told the model to try harder and it solved a Millennium problem!". Not great if true.
This part in particular was especially egregious:
> I said that if OpenAI released its result in the way proposed I would go public with what happened. The reply was, “Why would you ruin your career?” I replied that I am an academic, and asked why he thought going public would ruin my career. The reply was, “If you don’t want me to be nice, then I don’t have to be nice.”
Comment by sk4rekr0w 6 hours ago
Comment by tigershark 6 hours ago
Comment by sk4rekr0w 6 hours ago
Comment by monster_truck 2 hours ago
Comment by dist-epoch 2 hours ago
Comment by kzrdude 1 hour ago
While I support their argument - push for stronger data and privacy protections from OpenAI and similar - it is naive to believe we can have privacy while sending our data to third parties. It's clearly better to be safe than to be sorry here. Well, clearly better in terms of privacy. In terms of the maths gold rush, who can say what's better, that probably favours those taking more risk.
Comment by ks1723 6 hours ago
Telling OpenAI that Anthropic has apparently solved an important problem but most likely that refers to him and he is using OpenAI models (not Anthropic's)?
And he wants to clarify that with OpenAI in advance? And get a pardon for Anthropic's likely but false press statements?
I dont get it.
[edited] needless to say, the behavior of the OpenAI employee is really despicable
Comment by traes 5 hours ago
https://xcancel.com/ElliotGlazer/status/2096298696438906934#
Comment by calf 4 hours ago
Comment by dbdr 5 hours ago
That seems like a valid reason to contact OpenAI.
Comment by sashank_1509 7 hours ago
Comment by traes 5 hours ago
> We used several LLMs throughout: Anthropic’s Claude, OpenAI’s Codex, especially with GPT-5.6 Sol and, more recently, Astra. The latter was only used for writeups and auditing our arguments.
Comment by supriyo-biswas 7 hours ago
Comment by johnnienaked 5 hours ago
Comment by monster_truck 2 hours ago
Comment by happa 6 hours ago
Comment by phorkyas82 6 hours ago
(some drama from good ol' William)
Comment by dash2 5 hours ago
Comment by cherryteastain 5 hours ago
\nu d^2 u_i / dx_j dx_j - Viscosity
-1/\rho dp/dx_i - Pressure gradient
u_j du_i / dx_j - Advection. Kinda like momentum transfer from the motion of the fluid itself. Nonlinear, which makes the N-S equations hard to solve
du_i/dt - Rate of change of velocity. Note that this is in an Eulerian framework so it's not the acceleration of a packet of fluid, rather it's just the change in velocity at a particular location in space
Euler is when you omit some terms. Forcing is when you add some other terms to account for phenomena external to the fluid like gravity or flow through a porous medium like in the article.