How An AI math breakthrough ignited a controversy

(science.org)

55 points | by pseudolus 1 hour ago

13 comments

  • afavour 12 minutes ago
    The core section:

    > However, communications quickly became contentious. According to Buckmaster, OpenAI offered to give him sole authorship on the Navier-Stokes solution—but only if Alpöge’s name was removed from the work and if the write-up would acknowledge the problem had been resolved by an internal OpenAI model. Buckmaster refused, in part because he was troubled by the question of what OpenAI's system had actually seen. For example, Buckmaster said the company did not initially give him a clear answer about whether its agents had access to the pair's logs on Codex (which is an OpenAI product).

    > OpenAI executives have denied that any employee or AI agent saw the pair’s work before the researchers released it publicly on 7 September. But there still remains a separate question: Could the pair's work have reached OpenAI's models through its training data?

    > OpenAI’s blog announcing the Navier-Stokes solution does not dismiss the possibility: “While unlikely, we cannot rule out that de-identified data derived from [Buckmaster and Alpöge’s] usage of our products helped improve our models .”

    • huurtehoog 10 minutes ago
      So, this company wants everyone and every organization on Earth to use their software, and reserve the right to then tell anyone how the resulting work can be published and credited?

      Real solid business model there, how could it ever fail?

    • olmo23 2 minutes ago
      > While unlikely, we cannot rule out that de-identified data derived from [Buckmaster and Alpöge’s] usage of our products helped improve our models.

      Well yeah, if you use the free product they train on your data, ... I thought this was widely understood?

      • greggoB 0 minutes ago
        If you read Buckmasters statement, he specifically notes that they used the paid subscriptions, iirc.
  • rsfern 5 minutes ago
    Regardless of what you think of the priority dispute issue discussed on sibling threads, I’m highly skeptical of the closing quote that this Navier Stokes result means that the same approach of casually spending a few million on agentic computation is going to solve end to end materials design or drug development.

    Those problems can’t be formally verified with an automated theorem prover. We have a lot of physics based simulation tools, but they tend to focus on small subsets of the full design problem and they make limiting approximations because otherwise they’d be too computationally expensive, or we just don’t have the right data to parameterize them beyond describing qualitative behavior. Agents are helping accelerate research in these fields but I think it’s mostly a different class of problem that’s a lot harder to specify and verify

  • pseudolus 36 minutes ago
    Extensive discussion on OpenAI's blog post on Navier-Stokes: https://news.ycombinator.com/item?id=49613262 .

    Quanta Magazine article that also discusses some of the controversy: https://www.quantamagazine.org/ai-has-solved-one-of-maths-1-...

  • paxys 18 minutes ago
    > Navier-Stokes is one of six “Millennium Problems” on a list compiled by the Clay Mathematics Institute in 2000.

    Seven, not six. One is solved already, but is still a millennium problem.

  • timmg 3 minutes ago
    I think it was a pretty questionable thing to do by trying to front-run these researchers even if they didn’t make use of their techniques. The fact that they may have inadvertently “borrowed” their work via training data makes it much worse.

    OpenAI’s behavior here — even if you only consider there side of the story — was (at best) in bad taste.

  • lolakutty 2 minutes ago
    Who should get credit? All the humans who ever worked to create the data.

    The AI company for making the search program that searched through the data and found the solution.

  • rrhjm53270 10 minutes ago
    My impression: the re-aristocratization of scientific research seems inevitable.
  • elternal_love 3 minutes ago
    Hmm, is the formal verification through? Lean just asserts no errors in the proof, but like can prerequisites not be fullfilled?
  • snsr 13 minutes ago
    OpenAI apparently used Buckmaster and Alpöge‘s work w/ Codex to bootstrap “their” dis-proof. https://cims.nyu.edu/~tristanb/statement.pdf
    • cmiles8 9 minutes ago
      And according to that team OpenAI only started asking their own model these questions after those submissions had occurred. So OpenAI had these critical clues and info before they started. If OpenAI did or did not use that to produce their own “proof” is an open question, but OpenAI hasn’t definitely denied it.
  • zero-sharp 14 minutes ago
    Moving forward, I can't imagine other mathematicians wanting to have this kind of experience. So there has to be a shift away from these services.
    • mrngld 6 minutes ago
      If you mean move to products with ZDR policies, OK, very fair and I agree. If you're talking the equivalent of carpenters should give up pneumatic air guns because hammers are more authentic, then that's silly. These are tools, and like any tool how effective they are can come down to how well you use them.
      • ardacinar 0 minutes ago
        I don't think there are any pneumatic air gun "provider"s that come with an associated risk of them claiming your carpentry work.
    • serial_dev 10 minutes ago
      Like SaaS companies who don’t feed all their stuff to Anthropic, OpenAI and Cursor?
    • searls 11 minutes ago
      … which would only exacerbate the computational disparity they're operating under.
  • ltononro 7 minutes ago
    Does it matter who gets the credit at this point? Both used AI to do 99%+ of the work. So... do machines have ego?
    • mrbungie 1 minute ago
      OpenAI wants everyone to believe that it was done with a practically autonomous network of thousands of agents with little to no human intervention for 88 hours, while the Buckmaster/Alpöge were using AI in a more guided way for months. If OpenAI actually used anything from Buckmaster/Alpöge work they would be misleading the public.
    • usrnm 4 minutes ago
      I guess, you don't use AI for work? Otherwise, why should you be paid for it? Should you only be paid for the code you actually wrote yourself?
    • elternal_love 5 minutes ago
      Machine owners have stocks which valuation must increase to satisfy their debts.
  • jibal 15 minutes ago
    > OpenAI, meanwhile, says its experience with Navier-Stokes could open the door to solving puzzles with more practical relevance. “We are now able to spend millions of dollars on a problem that we really care about and that really matters: developing new materials, finding cures to diseases,” Bubeck said. “All of those things that we have been talking about for a long time—now they seem to be at our fingertips.”

    Eh? There's no connection at all between the Navier-Stokes work and those things.

    • vbezhenar 8 minutes ago
      He may be hinting that solving these complex mathematical problems will boost OpenAI clients' confidence and encourage them to spend millions of dollars on solving other complex problems.
    • bananaflag 8 minutes ago
      Yes it is they all require intelligence.

      Navier Stokes is a test of how high the intelligence is.

  • booster-rooster 10 minutes ago
    [dead]