‹ BackHN Continuity

Thread

Did OpenAI solve the wrong Navier-Stokes problem?

122 points · 70 comments · tomjakubowski

Loading the complete thread in the background. This saved snapshot is available now. Refresh

  1. bmacho · · focus · HN ↗
    Article says that there are 2 formulations of the NS problem and both are interesting: one is about fluid behaviour with no external forces, other is about fluid behaviour with external forces.

    For a counter-example the latter is easier since you can have a tricky external forcefield.

    1. tomjakubowski · · focus · HN ↗
      [delayed]
    2. HarHarVeryFunny · · focus · HN ↗
      Right - OpenAI proved the forced blow-up case rather than the harder unforced one, with the Millenium Prize problem statement saying it would be awarded for either one.

      The forced version is easier since you can custom design the force function to get the result (it doesn't have to be a realistic force like stirring), so getting the blow-up might be regarded just as much a function of your bespoke force function as of the fluid dynamics itself, which is apparently what OpenAI did, pushing the definition of the force function being "smooth" to it's limit.

      So, it appears OpenAI did legitimately meet the Millenium Prize solution criteria, but in the most unrealistic, and therefore least interesting, way possible.

      1. adriand · · focus · HN ↗
        My friend, who is a mathematician, sent this in our group chat:

        The title is a bit misleading. The variant with a smooth forcing was one of the four valid variants in the Clay formulation. It is interesting to solve it. It is still an interesting and impressive result. The no force version is also interesting and remains unsolved. It isn’t reasonable to just dismiss the proof on the grounds that 26 years later we claim it was never that interesting. This is the first time I’ve seen this attitude.

        1. pennomi · · focus · HN ↗
          It’s just people trying to cope in the most human way possible: “I don’t like AI, therefore AI is stupid, therefore its proof must be uninteresting.”

          Almost never do people judge situations entirely on merit.

          1. HarHarVeryFunny · · focus · HN ↗
            Not really - if you read beyond the headlines of "mathematicians don't like AI", and your own imagined (incorrect) reasons they may have for that "AI is stupid", then the truth seems a bit more interesting.

            If you want to abide by your own words and judge the situation on it's merit, then you need to look at the specifics, meaning the OpenAI proof itself (166 pages), and the analysis of it that is only just beginning. Assuming that the proof is wonderful and provides much insight into Navier-Stokes is just as dumb a take as saying that it doesn't. Judge it on its merit.

            The Scientific American article is sadly paywalled, but at least part of the discussion is based on the paper below, whose work the OpenAI proof appears to build upon.

            <a href="https:&#x2F;&#x2F;arxiv.org&#x2F;pdf&#x2F;2609.20803" rel="nofollow">https:&#x2F;&#x2F;arxiv.org&#x2F;pdf&#x2F;2609.20803

            When discussing the proof, below, with Sonnet, and asking it to explain the distinction between a function being smooth and analytic, one aspect that appears interesting is that the OpenAI forcing function is apparently constructed out of &quot;bump functions&quot;, meaning that it is not a uniform force acting upon the flow but rather a highly engineered pattern of pokes, localized in time and space, which as another commenter in this thread notes sounds similar to Maxwell&#x27;s Demon - another theoretical force, that neither tells us anything about Brownian motion nor the 2nd &quot;law&quot; of thermodynamics.

            So, we&#x27;ll have to wait for mathematicians to continue to analyze the proof, and determine to what extent is does deliver on providing insights into Navier Stokes, and any potential improvement to it, that was the goal of setting it as a Millenium Prize in the first place.

            <a href="https:&#x2F;&#x2F;cdn.openai.com&#x2F;pdf&#x2F;32d9f210-8b73-45e0-91bc-82a30aef8a9a&#x2F;navier-stokes.pdf" rel="nofollow">https:&#x2F;&#x2F;cdn.openai.com&#x2F;pdf&#x2F;32d9f210-8b73-45e0-91bc-82a30aef8...

  2. busssard · · focus · HN ↗
    &quot;AI solved it.&quot;

    &quot;No we almost solved it.&quot;

    &quot;No way, it was AI all alone&quot;

    &quot;no you used our data for train...&quot;

    &quot;Guys, Guys calm! You did not produce any useful results!&quot;

    1. rsfern · · focus · HN ↗
      I think this is dramatized to the point it’s talking past the article, the math community isn’t really making any of these arguments from what I can tell.

      The discourse is (1) models are capable of making really impressive mathematical advances, usefulness is not in dispute, (2) the frontier AI companies aren’t being super transparent about information sources so it’s hard to know exactly how to evaluate the level of capability that was demonstrated, and (3) there are lots of kinds of math that is interesting and there are open questions about how to get there.

      In particular this article highlights a particular open question I’ve seen discussed on HN before, which is that the particular proof strategy of finding a counterexample might be more amenable to RL than other strategies of proof that might be needed to resolve the other branches of the Navier Stokes problem (and probably other similar areas of math)

      1. busssard · · focus · HN ↗
        yeah i agree its dramatized, but the situation was quite dramatized by the parties involved as well. i just find it quite funny, that the perceived drama might play out like this now.
      2. fmbb · · focus · HN ↗
        Is it really impressive, or just kind of interesting?

        If they spent about 10 GWh solving the problem (was it solved?) then that is much much more than 500 lifetimes of a human brain working.

        1. dgellow · · focus · HN ↗
          A specific version of the problem has been solved. But not the broader, harder version. The article explains that fairly clearly.

          I’m very anti AI and OpenAI, and do think it’s a pretty interesting finding! Very likely not worth their spend, but interesting and novel nonetheless the less

          1. busssard · · focus · HN ↗
            yeah but the problem they solved is not really useful. the article is also quite clear on that
    2. guywithahat · · focus · HN ↗
      &gt; &quot;Guys, Guys calm! You did not produce any useful results!&quot;

      That is the best response I&#x27;ve heard to this argument. Assuming the solution is correct, the fact it is not the most interesting solution that could have been solved is besides the point. The team at OpenAI did an incredible job solving the problem.

  3. davidguda · · focus · HN ↗
    Does it matter if it did? They got the headlines.
  4. ChickeNES · · focus · HN ↗
    So what they are saying is that humans, in this case Charles Fefferman (a math prodigy, going by his history), failed to specify the problem correctly?
    1. dgellow · · focus · HN ↗
      No, that not the issue. If you look at <a href="https:&#x2F;&#x2F;www.claymath.org&#x2F;wp-content&#x2F;uploads&#x2F;2022&#x2F;06&#x2F;navierstokes.pdf" rel="nofollow">https:&#x2F;&#x2F;www.claymath.org&#x2F;wp-content&#x2F;uploads&#x2F;2022&#x2F;06&#x2F;navierst..., second page, you will see an option C is one of the four that is asked to be solved. And that option is the one that allows for an external force, which is what OpenAI solved.
      1. hatthew · · focus · HN ↗
        I think GP&#x27;s point is that if the larger math community doesn&#x27;t care about option C, Fefferman shouldn&#x27;t have given that option in the first place.
        1. ChickeNES · · focus · HN ↗
          yes!
        2. dragonwriter · · focus · HN ↗
          Fefferman and the larger math community are not the same actor; a divergence between their concerns is not surprising.
          1. ChickeNES · · focus · HN ↗
            But Clay chose Fefferman to represent &quot;the larger math community&quot;? I&#x27;m calling out Clay&#x2F;Fefferman here, not the community. Though the community also had over 25 years to contest the problem statement, and had no problem with it until now.
            1. dragonwriter · · focus · HN ↗
              &gt; But Clay chose Fefferman to represent &quot;the larger math community&quot;?

              Surprisingly enough, being chosen by another actor as proxy for some group doesn’t actually resolve the problem that an individual may not always be an accurate proxy for the concerns of the group (and especially for the same descriptive aggregate group a generation after the proxy acts on their behalf.)

          2. scheme271 · · focus · HN ↗
            I&#x27;m pretty sure Fefferman&#x27;s formulation of the problem got a lot of vetting before the Clay Institute accepted it. As such, it probably does reflect the math community&#x27;s understanding of the problem. The &quot;problem&quot; such as it is, is that it might be the least interesting case especially given the solution that was obtained.
        3. [deleted] · · focus · HN ↗

          [deleted]

        4. etdznots · · focus · HN ↗
          A solution to any of the four options made by an LLM armed with a theorem solver and millions of dollars of compute would still not be that interesting to the math community since it is not likely that any new insights oe questions can be derived from the result
      2. ChickeNES · · focus · HN ↗
        &gt; There is no question that OpenAI solved what the Clay institute is looking for. But that specific option C isn’t what the larger math community cares about, it’s a pretty niche case

        I think you are agreeing with me? My point is that &quot;the larger math community&quot; failed to set the bounds of the problem correctly.

        1. dgellow · · focus · HN ↗
          The clay institute doesn’t represent the broader math community
      3. aesthesia · · focus · HN ↗
        It is interesting that there&#x27;s a gap between the &quot;prove N–S existence&quot; conditions, which assume no forcing term, and the &quot;counterexample&quot; conditions, which allow a nonzero forcing term. In theory both (A) and (C) could be true.
    2. dragonwriter · · focus · HN ↗
      Yes. And more humans—in this case OpenAI researchers—similarly failed in choosing how to direct the AI tools.

      But its not news that computers are mere tools and that any error blamed on a computer involves at least two human errors, one of which is blaming the computer instead of the human(s) responsible.

    3. SpicyLemonZest · · focus · HN ↗
      [delayed]
    4. alkyon · · focus · HN ↗
      Fefferman was right to include option C as someone could have come up with less contrived counterexample accompanied by some interesting theorems that actually shed light on the general case
      1. perching_aix · · focus · HN ↗
        [delayed]
  5. IshKebab · · focus · HN ↗
    The goalposts are moving so fast you can barely see them!

    Next year&#x27;s headline: But did Skynet kill all humans?

    1. dgellow · · focus · HN ↗
      That has nothing with moving a goalpost, it’s about understanding the actual result and digging into the details
    2. gr_norm · · focus · HN ↗
      This is mathematicians using the solution, explicitly acknowledged to solve the original formulation of the problem, to pose interesting new questions. That is what mathematics is.

      Only someone who has never interacted with mathematics outside a rote-problem-solving capacity would describe it as you have.

      1. IshKebab · · focus · HN ↗
        Please. The headline is clickbait, trying to imply that OpenAI somehow screwed up.
        1. gr_norm · · focus · HN ↗
          What I&#x27;ve said is rather clearly spelled out in the article. Consider reading the it next time?

          &gt; It did, however, unambiguously solve the problem according to the Clay Institute’s original formulation.

          1. IshKebab · · focus · HN ↗
            I did read it. I was talking about the title.
      2. curt15 · · focus · HN ↗
        Or people with a vested interest in pushing the frontier labs&#x27; preferred narrative that they&#x27;ve achieved AGI...even while frontier labs continue to hire human &quot;Account Associates&quot;, &quot;Android Engineers&quot;, &quot;Applied AI Engineers&quot;...
  6. ActorNightly · · focus · HN ↗
    Betteridge&#x27;s law of headlines
    1. dgellow · · focus · HN ↗
      That would be wrong here…
  7. olliepro · · focus · HN ↗
    Classic example of moving the goal posts. “Exploiting a loophole” is how you solve many great problems in math.
    1. dgellow · · focus · HN ↗
      That’s not at all the topic of discussion. The article is talking about the fact that OpenAI solved a version of the problem that is niche and isn’t the one the math community cares about
      1. krackers · · focus · HN ↗
        &gt;and isn’t the one the math community cares about

        Then why was it allowed as an option in the millennium prize statement?

        1. SpicyLemonZest · · focus · HN ↗
          At the time that the Millenium Prize problems were formulated, the force term was understood to make the problem more realistic, since real fluids are always going to have external forces applied to them. A blowup that happens under constant gravity, for example, would probably be no less interesting than an entirely unforced blowup. The strategy of constructing impossibly complex external forces to induce a blowup was pioneered by Córdoba and Martínez-Zoroa only over the past few years.
          1. oliculipolicula · · focus · HN ↗
            Today that point is moot because.. the Mill problems that are least likely to have been solved are most likely to have real world impact..

            I&#x27;ll take on the biased hope that most of 100 solutions to be released will not have any applications for at least 30 years..

            In order to counter my own fear that the 9 big names will not be able to get openAI to act &quot;more responsibly&quot; (whatever that means)

            1. SpicyLemonZest · · focus · HN ↗
              [delayed]
              1. oliculipolicula · · focus · HN ↗
                The point of the committee is to benefit Ant.. (maybe only by striking fear in OpenAI)
        2. aleph_minus_one · · focus · HN ↗
          &gt; Then why was it allowed as an option in the millennium prize statement?

          Because quite some mathematicians only realized after this OpenAI Navier-Stokes announcement that this allowed variant is much less interesting than they originally thought. :-)

        3. xigoi · · focus · HN ↗
          Bocause they expected it to be solved by human mathematicians using techniques that could provide insight into the harder version.
      2. signatoremo · · focus · HN ↗
        Who is this math community? Since when did they form the consensus that this option is not at all what they care about, before or after they knew about OpenAI’s solution?

        What about Tristan Buckmaster and Levent Alpöge, did they also attempt to solve the same challenge? Didn’t they know it wasn’t interesting?

        1. dgellow · · focus · HN ↗
          &gt; What about Tristan Buckmaster and Levent Alpöge, did they also attempt to solve the same challenge? Didn’t they know it wasn’t interesting?

          Yes, they were working on what is considered a niche case, the option C from the millennium statement for Navier-Stokes. The article explains that clearly, you can just read it and get the details

          1. oliculipolicula · · focus · HN ↗
            I wish there were some better way to see, before the IPO, how openAI&#x27;s actions correlate with their bottomline.

            (I&#x27;m proAI (for the masses, but green) BUT antiopenAI and a bit less antiAnthropic)

            1. dgellow · · focus · HN ↗
              Dario is sexily dangerous until you read his essays in full then realize it’s just the same LessWrong sci-fi ideas at the core, all based on absurdly simplified thought experiments designed to justify their priors. Those people are deeply unserious and do not value existing humans. They already made their mind decades ago rereading the fact we need a machine god and that they should be the ones building it.

              Just the idea of making a conscious digital being, to then get it to process excel files for its whole existence is such an immoral concept

  8. jonlong · · focus · HN ↗
    No, OpenAI did not solve the &quot;wrong&quot; Navier-Stokes problem. OpenAI did not solve the hardest version of the problem (unforced blow-up), but did give a solution to the Clay Millennium Prize Problem as written and understood, choosing the explicitly allowed forced option.

    SciAm writes &quot;in a sense, the LLM found and exploited a loophole in the framing of the question&quot;. This is pure sensationalism. Choosing option (C) (out of an explicit list of four options) is neither a &quot;loophole&quot; nor something &quot;found by the LLM&quot;; everyone involved knew this was the option they were pursuing.

    With the grumbling out the way, there is some actual scientific content to the article: there&#x27;s a strong argument that OpenAI&#x27;s method will not extend to the unforced case, leaving our understanding of NS incomplete. This negative result is itself new and interesting (and predicated entirely on the solution found by OpenAI)!

    1. hn_throwaway_99 · · focus · HN ↗
      Totally agree. I wouldn&#x27;t quite call it clickbait, but the article does this thing I find annoying where it puts the &quot;sensationalist&quot; framing at the beginning (the &quot;loophole&quot; quote you put), but then closer to the end fully admits that it wasn&#x27;t really a loophole in any case:

      &gt; It did, however, unambiguously solve the problem according to the Clay Institute’s original formulation. The official problem statement, penned in 2000 by mathematician Charles Fefferman, offers an option called “C,” in which solutions are allowed to use an external force like OpenAI’s.

    2. HarHarVeryFunny · · focus · HN ↗
      As I understand it the &quot;loophole&quot;, if you want to call it that, is that OpenAI&#x27;s custom-designed forcing function was smooth, as the rules said it had to be, but was non-analytic, consisting of some construction of &quot;compactly supported bump functions&quot;, meaning a mass of tiny little pushes at precise points of space and time to push a vortex into blowing up the math.
      1. hotdog1492 · · focus · HN ↗
        Reminds me of Maxwell&#x27;s Demon.
    3. p-e-w · · focus · HN ↗
      &gt; SciAm writes &quot;in a sense, the LLM found and exploited a loophole in the framing of the question&quot;.

      God, it’s embarrassing to read stuff like this. They’re making it seem as if everyone involved was either stupid or dishonest just so they can pretend they have a scoop here.

      1. dmix · · focus · HN ↗
        There&#x27;s a big market in journalism to take down the popular thing in the news. &quot;Everyone is wrong&quot; and here is my article where I wildly exaggerates some minor details to justify the headline which made you click on the article.
    4. dist-epoch · · focus · HN ↗
      Or a better framing: Clay Institute chose the wrong Navier-Stokes problem for a Millennium Prize.
      1. glimshe · · focus · HN ↗
        Yet it had remained unsolved until OpenAI&#x27;s effort...
      2. seanhunter · · focus · HN ↗
        Not really. All of the Navier-Stokes options in the Clay Institute formulation of the prize are real problems of significant interest. The history of this particular problem is of finding specific conditions under which we can get something to work that turn out not to generalize in ways that people don’t expect eg iirc (It’s been a while since I read about it) there was a solution found early-ish in the 20th century for the 2-D case that turned out not to generalize to N-d, N&gt;2 case, there are special conditions under which the turbulent terms cancel out and you can get smooth flow, vortices etc.

        It seems to me that formulating problems at the boundary of human knowledge precisely is always going to be challenging and situations are bound to occur where you look back with the benefit of hindsight and wish that you had posed the question slightly differently based on some knowledge you didn’t have at the time.

    5. casey2 · · focus · HN ↗
      It&#x27;s just moving the goalposts, this happens every time an AI solves a problem, doesn&#x27;t matter if the goalposts were there for 26 years.

      What&#x27;s interesting is that there are a set of people who are &quot;in charge&quot; and can as they wish arbitrarily set the goalposts to the thing that they happen to be best at. While this might be satisfying for an Humanity vs AI narrative, it&#x27;s concerning for an us vs them one. Are these people really special? or do they just change the rules of the game so that outsiders (human or AI) can&#x27;t win.

      1. cyanydeez · · focus · HN ↗
        AI, the new minority!
  9. yieldcrv · · focus · HN ↗
    treating LLM&#x27;s under a separate apartness ruleset isn&#x27;t going to go well
  10. yababa_y · · focus · HN ↗
    I will admit that when I heard they only had forced blowup I went back to sleep (but I&#x27;ve never been more that cursorily interested in analysis).
  11. ex-aws-dude · · focus · HN ↗
    I dunno this sounds like some insane goalpost moving
    1. jrflo · · focus · HN ↗
      “Yes they solved one of the 7 most famous unsolved problems in math today but they only did the easiest version!”

      At the current rate (if they keep burning tokens on it, which maybe they won’t given the backlash) RH will be proven within a year and there will be some other thing that means it’s not actually that impressive…

  12. emmelaich · · focus · HN ↗
    <a href="https:&#x2F;&#x2F;archive.is&#x2F;52HlJ" rel="nofollow">https:&#x2F;&#x2F;archive.is&#x2F;52HlJ
  13. jongjong · · focus · HN ↗
    This situation illustrates exactly the limitation of AI and why we still need humans in the loop.

    It reminds me of a junior coding bootcamp lecture I once gave, one of the first slides said &quot;Computers will do what you say, not what you mean.&quot;

    1. dist-epoch · · focus · HN ↗
      Seems to me this illustrates more the limitation of humans, they are the ones which chose the wrong problem, not the one mathematicians cared about.
      1. jongjong · · focus · HN ↗
        Sure, but you gotta compare frontier AI with frontier human!
  14. cpard · · focus · HN ↗
    When the news spread about the solution of this problem by AI we started wondering what will happen when AI will start generating proofs we can’t comprehend.

    Today, we are discussing if AI cheated by picking the easy problem to solve which means that we at least still comprehend what’s going on.

    I wish mathematics and the rest of the human intellect wouldn’t turn into content marketing that is generated primarily to trigger strong human emotions.

    I feel that this is going to hurt both AI and the disciplines that can benefit the most from it

    1. irieir · · focus · HN ↗
      Cringe it’s not that deep.

      OAI should just get on with it and more importantly - produce more stuff that positively benefits the vast majority of the population.

      1. rafterydj · · focus · HN ↗
        That doesn&#x27;t help foster meaningful discussion. It&#x27;s very relevant to discuss what we understand.
  15. kittikitti · · focus · HN ↗
    The open question is whether the Clay institute will award OpenAI or others the Millenium prize for this solution to the Navier Stokes problem. I speculate that they won&#x27;t. OpenAI&#x27;s solution is undergoing peer-review and the counterargument presented in this article changed my mind. By default, if the Clay institute hasn&#x27;t officially recognized the solution as true or likely true, I can&#x27;t make an assumption that it is. I hope to go through the proof and perhaps AI can help better understand it and any potential weaknesses.
  16. chmod775 · · focus · HN ↗
    There might be more value in having AI systematically hunt for mistakes in existing and widely assumed correct math papers, or hunt for counter-examples to things thought proven. There&#x27;s likely to be a couple mistakes hiding in the less well scrutinized edges of mathematics.

    Maybe something interesting will fall over because of that, who knows?

  17. nikitamalyavin · · focus · HN ↗
    What will happen if we&#x27;ll build a real-life test if the proposed counterexample? The exactly same force, as defined, shape of the vortex, etc?

    We should either see the effect or get to understanding of how to correct the model.

    Note that the force is finite and doesn&#x27;t &quot;know&quot; the vortex config. The singularity should be achieved in a finite time.

Open on Hacker News to reply ↗

Unofficial Hacker News client; not affiliated with Y Combinator.