‹ BackHN Continuity

Thread

The LLMentalist Effect (2023)

235 points · 309 comments · jalev

  1. rahidz · · focus · HN ↗
    Man I remember back when a psychic conned me by solving the Navier-Stokes problem.

    Also >July 4th, 2023

    1. krupan · · focus · HN ↗
      The LLM did not solve it. It's not intelligent. Humans did, using a statistics-based computational tool (the LLM). We don't even know all the details of how the tool was used, we haven't been allowed to use the exact tool they used ourselves, we don't know much it really cost in dollars, energy, or time, etc. etc.
      1. claytongulick · · focus · HN ↗
        And it may have trained on a NYU professor's work.
        1. tripledry · · focus · HN ↗
          Why is this downvoted? Genuine question, I haven't followed up on the drama.
          1. pixl97 · · focus · HN ↗
            Because it's mostly not true. It is likely that it used the professors work, but the professor did not have a solution. It came up with new insights that solved the problem. Even the humans from the professors side said so.
            1. claytongulick · · focus · HN ↗
              > Because it's mostly not true. It is likely that it used the professors work

              These two statements appear contradictory.

              I simply said that it may have trained on a NYU professor's work.

              Work that the professor did not believe he was releasing for model training purposes. That feels worthy of mention.

          2. kingstnap · · focus · HN ↗
            According to OpenAI the cut off date for user data was too early for that (one sided evidence, so I'll give this partial consideration).

            The NYU professor was solving a different problem (no viscosity, aka the Euler equations). This is a big difference.

            The NYU professors' blowup construction was fundamentally not the same, it was a donut with a cascade of smaller and smaller vortexes driven by each other. OpenAI has that picture they made but its inwards spiraling and speeding up vortex.

            My overall opinion is that calling the work plagiarized is really underselling what the AI accomplished. It's like full on cope.

            In particular, Buckmaster's main claim to plagiarism is this:

            > “Almost nobody was seriously developing this particular constructive program for realizing C/D, and then OpenAI appeared in essentially the same general part of the landscape immediately after hearing about our progress.”

            What this fails to realize, is that this only points to plagiarism if the counterparty isn't AI. They had actually launched teams on all cases in parallel.

Open on Hacker News to reply ↗

Unofficial Hacker News client; not affiliated with Y Combinator.