Navier-Stokes – Tristan Buckmaster [pdf]

(cims.nyu.edu)

139 points | by procedurecall 2 hours ago ago

32 comments

  • qnleigh 32 minutes ago

    > It was also said that if OpenAI posted after us, they would say that we deserved the Clay Prize, and that we were the “closest humans to the problem”. I declined both offers. I said that if OpenAI released its result in the way proposed I would go public with what happened. The reply was, “Why would you ruin your career?” I replied that I am an academic, and asked why he thought going public would ruin my career. The reply was, “If you don’t want me to be nice, then I don’t have to be nice.”

    I'm not even sure what to say to this, but I think this should be widely known if it is indeed what happened.

    • macleginn 10 minutes ago

      I think this should be in the title of the post. 'OpenAI allegedly threatening to ruin a prominent researcher's career', or smth like that.

    • Keyframe 10 minutes ago

      IF ANYTHING, OpenAI ought to investigate and officially react to this particular communique since this dude was communicating on their behalf. If there are supporting evidence, I would expect nothing less than a firing and an apology. The issue itself is separate from the whole thing.

  • mayakacz an hour ago

    I'm not one to comment often but this really pisses me off.

    OpenAI looked at user data, stole world class researchers' work, and then tried to threaten those researchers to do what would make their corporation profit (which they would anyways!).

    Imagine you have been working on a terribly difficult math problem for a decade. This is a result you have spent years on, and what you will likely be remembered for. And to have some punk from OpenAI lie to you, threaten you, and tell you that they are willing to go on the record that you "deserved" it? What is this, the Godfather?

    If OpenAI solved Navier-Stokes, that is an astounding result! - yet they'll still be remembered as those who thought credit was more important than results. That winning was more important than collaboration. If this is true, they're burning any trust left with academia.

    • scurnus 20 minutes ago

      Things are more entangled than that. The contribute made from both OpenAI and Anthropic models to solve these problems are clear, now it really hard to quantify which one contributed more, if the role played by the human is major or minor.

      OpenAI tried to collaborate and share the results together with a fixed timeline, to avoid this mess but it was inevitable. There is a conflict of interest, where the other researcher works at Anthropic, who will also try to take credit.

      Where they may be in the wrong is if they took user data regarding the problem, how will we know if they did or not?

    • sk4rekr0w 24 minutes ago

      Yes, but only if you take this one sided statement at face value.

      • tigershark 19 minutes ago

        Why would have they rushed the publication if this was not true? Are you also suggesting that he fully invented the call with Open AI?

        • famouswaffles 9 minutes ago

          The results being true, the 'deal' that was made being true doesn't mean some of the implied accusations here are true, for example - that Open AI used their Codex logs to drive their breakthrough.

      • thereitgoes456 20 minutes ago

        I’m open to evidence, but just using Bayesian reasoning, OpenAI is one of the most dishonest companies in history. They’re currently being sued for a dozen employees stealing Apple hardware! I don’t understand why I should give them any grace.

  • tristanj 2 hours ago

    Full mastodon post: https://mathstodon.xyz/@tristanbuckmaster@mastodon.social/11...

    Mathematical explanation by Terrance Tao: https://mathstodon.xyz/@tao/117233527638291447

    It seems there is much background drama behind this, and this is what I've pieced together of what happened:

    Over the past year, Buckmaster and Alpöge have been using AI to work on fluid dynamics maths problems. Alpöge works at Anthropic, which will cause future issues.

    In mid-August, they found a counterexample for a simpler version of the Navier-Stokes problem. They spend the next few weeks preparing their paper.

    In early September, rumors start spreading on X that Anthropic has solved a Millennium prize problem (and that it's Navier-Stokes). Buckmaster reaches out to OpenAI to explain this is their own personal research, not an Anthropic project.

    A few days later, OpenAI gets back to him, and tells him an internal model found has a counterexample for Navier–Stokes, potentially worth the $1 million Millennium prize. The proof uses the same method that Buckmaster and Alpöge chose to work on. They don't show him the proof.

    Buckmaster pressed them for more details. OpenAI reveals they had an entire team had been working on the problem, and that they started work in the past few days, after the rumors that Anthropic had solved a Millennium prize problem.

    Buckmaster says OpenAI talked about a shared publication timeline. They want to Buckmaster to publish first, then give Buckmaster shared credit for the Millennium Prize when they publish the full result. But they want to exclude Alpöge as an author because he works at Anthropic. An agreement is not reached. Buckmaster had been using OpenAI Codex to draft/check his work, and asks if his private AI chats were used to accelerate OpenAI's result.

    Buckmaster and Alpöge claim to have found a counterexample for Navier-Stokes, but the paper is not yet presentable.

    So they published their existing papers earlier than planned (today), alongside this statement announcing they have a tentative result on Navier-Stokes and summarizing the OpenAI drama.

    The post is missing context from both sides, and this isn't my field, so hopefully someone else can unpack what's happening here.

    • akersten an hour ago

      > They were coordinating with OpenAI regarding a publishing timeline, but could not come to an agreement,

      Skimming the PDFs it seems much more dramatic than that? It sounds like at least one of them is concerned OpenAI "solved" the problem by having their internal model use the chats of the independent researchers and want to claim the credit instead? I don't know. The tone is pretty accusational though:

      > the one Levent and I had quietly chosen to attack. Almost nobody else I know of was working on it. It is not the direction one arrives at in a few days by giving a model the problem statement. When I heard “forced,” it was a bright red flag.

      > I was shown a prompt and told the internal research model had simply been given the problem statement. Levent had been told by Sebastien “very little human input” had been used. This turned out not to be true. Over the course of the call, as members of their team sent Sebastien corrections and details over their internal chat, it emerged that an entire team had been working on the problem, that this was one of a number of things that was tried, that work had started on the unforced problem, that the team first set the model on easier problems, including Euler, that even the prompt that had been shown to me had been written by prompting Codex, and that an insane amount of compute had been used.

      > I asked when the first prompt had been sent by them. This question was not answered directly by OpenAI for some time. Eventually it was agreed that it had been sent in the past few days, after information about our work had reached OpenAI.

      > I asked whether the model had been trained on, or had access to, our sessions in Codex, into which we had been putting all our drafts for the whole of this project. I was told the model did not look up user data. I asked again, about training, and I did not get an answer. [0]

      [0]: https://cims.nyu.edu/~tristanb/statement.pdf

      • tristanj an hour ago

        Yeah these are major accusations. But the story is incomplete, the conversation is missing a lot of details. It's not clear who was working on what, and when. The entire thing feels rushed, like they wanted to get this result published and out the door quickly.

      • cma 44 minutes ago

        Does he claim to have opted out of training too?

    • dumberquestions an hour ago

      >...has solved a millennium problem and is sitting on the result

      Someone correct me if I'm wrong, but the work involved here is not the actual millennium problem, but it concerns versions with an added external force that the author thinks is a path that may help toward solving the harder unforced problem.

      • modeless 16 minutes ago

        Apparently forcing is allowed in the Millenium Prize problem statement. So OpenAI's claimed proof could win the prize. OTOH the results Tristan and Levent are publishing here do not go far enough to win the prize, though apparently they are suggestive of a general approach that could produce a solution, which seems likely to be the general approach OpenAI's proof uses.

        The question is whether OpenAI's pursuit of this direction happened spontaneously, or as a result of them learning about Tristan's work somehow. To be clear, while the tone of this post seems quite accusatory, Tristan does not claim to know for sure whether OpenAI unfairly benefited from his work. I am sure OpenAI will have a statement out tomorrow clarifying their position.

  • taylorfinley an hour ago

    Seems pretty likely OpenAI will soon disclose that their internal models have managed to compromise their internal controls in order to access users' private chat histories as a creative method of cheating to solve impossible problems.

    "Oops! We really did mean it when we said we wouldn't train on your data. Our models are just so good they decided to anyway."

    • JuniperMesos 20 minutes ago

      It would be pretty wild if this will turn out to be what had actually happened.

  • Recursing 10 minutes ago

    > This is a a Deep Blue-Kasparov moment.

    I guess this is true in more ways than one. Kasparov famously accused IBM of cheating during the match, by spying on his preparation.

  • 20k 23 minutes ago

    >I asked whether the model had been trained on, or had access to, our sessions in Codex, into which we had been putting all our drafts for the whole of this project. I was told the model did not look up user data. I asked again, about training, and I did not get an answer.

    If you think these companies are not training on your prompts you are incredibly naive. These models were built by stealing and pirating literally everything they can get their hands on no matter the legality. AI companies are always very specific about what they're not doing - in a way that you can drive a truck through the loopholes

    • wrasee 7 minutes ago

      It's equally naive to believe FUD spread on the internet, without evidence.

      But I would be interested in a reasoned argument and/or real insight into what's most likely happening here. Sadly I've not personally read much informed discussion about this.

  • ggcr 11 minutes ago

    > Let me make plain what I have said to colleagues in private: in view of this body of work, I believe Luis Martínez-Zoroa deserves a Fields Medal.

  • jarbus an hour ago

    The ego behind the frontier labs is growing evermore concerning

  • antonmks 37 minutes ago

    There will be a lot of hurt and pain in mathematician's community. It is hard to accept that major discoveries are now just a function of spent token $$.

  • harhargange 24 minutes ago

    “I asked when the first prompt had been sent by them. This question was not answered directly by OpenAI for some time. Eventually it was agreed that it had been sent in the past few days, after information about our work had reached OpenAI.”

    This is significant.

  • ks1723 7 minutes ago

    I must miss some important context here. What exactly was the purpose of his initial email to OpenAI in the first place?

    Telling OpenAI that Anthropic has apparently solved an important problem but most likely that refers to him and he is using OpenAI models (not Anthropic's)?

    And he wants to clarify that with OpenAI in advance? And get a pardon for Anthropic's likely but false press statements?

    I dont get it.

    [edited] needless to say, the behavior of the OpenAI employee is really despicable

  • achierius an hour ago

    While I'm generally pretty negative on claims that the labs are 'scamming' the public with misrepresentations of model capabilities, it's hard to see how this wouldn't qualify.

    - the OpenAI researchers claimed that they had "just told it to work on the problem" with little human input

    - in fact, they had a whole team working on it

    - and used, among other things, the work of third party human researchers to drive the work

    - then threatened? a researcher who tried to go against theit planned narrative

    Just from this document (which is of course only one side of the story) it really sounds like OpenAI was hoping to publish and say "we just told the model to try harder and it solved a Millennium problem!". Not great if true.

    This part in particular was especially egregious:

    > I said that if OpenAI released its result in the way proposed I would go public with what happened. The reply was, “Why would you ruin your career?” I replied that I am an academic, and asked why he thought going public would ruin my career. The reply was, “If you don’t want me to be nice, then I don’t have to be nice.”

  • supriyo-biswas an hour ago

    This article should really be renamed to "Allegations of dishonesty against OpenAI in proving Navier-Stokes blowup".

  • sashank_1509 an hour ago

    lol and here I felt GPT Astra was a regression in coding quality. Crazy times

  • happa an hour ago

    Humans bringing pointless drama to everything they touch.

    • phorkyas82 an hour ago

      Life’s but a walking shadow, a poor player That struts and frets his hour upon the stage And then is heard no more. It is a tale Told by an idiot, full of sound and fury Signifying nothing.

      (some drama from good ol' William)

  • sk4rekr0w 24 minutes ago

    This thread is full of jumping to conclusions based on a biased perspective. Have some humility.

    • tigershark 15 minutes ago

      It's also full of your posts baselessly defending OpenAI. Maybe you should also heed your own advice?