‹ BackHN Continuity

Thread

Why I'm still bearish on LLMs after Navier-Stokes

496 points · 653 comments · jaykru

  1. pfdietz · · focus · HN ↗
    Specifically: bearish on LLMs generally, not bearish on LLMs for pure math.
    1. jaykru · · focus · HN ↗
      yes, huge for pure math and activities that look like it.
      1. danielmarkbruce · · focus · HN ↗
        Doesn't really even need to look like it. If you can verify rewards, RLVR will optimize really really well. If you can't... it's a struggle. There are probably fewer fields where you can verify rewards than one might hope.
        1. skydhash · · focus · HN ↗
          > There are probably fewer fields where you can verify rewards than one might hope.

          2 tasks I've done today that I believe robots are nowhere near being able to do: Cleaning my wardrobe and draining bad fuel out of my generator. As in generic use cases.

          1. danielmarkbruce · · focus · HN ↗
            Hard to verify that your wardrobe is clean. Also hard to verify that the bad fuel is out without physical sensors. Many, many tasks are quite difficult to verify beyond "you know it when you see it". That doesn't work so well for training a model.
Open on Hacker News to reply ↗

Unofficial Hacker News client; not affiliated with Y Combinator.