‹ BackHN Continuity

Thread

Why I'm still bearish on LLMs after Navier-Stokes

496 points · 653 comments · jaykru

  1. againstapples · · focus · HN ↗
    > the models generalize well only on tasks within a small neighborhood of the specific tasks they've been trained on, and even then with severe caveats. the frontier labs have developed a general recipe to teach models almost any specific task enjoying clearly defined levels of task performance; many tasks are covered in the training data

    Is this really any different to how humans learn, it takes a lot of training on one specific task to make a human expert as well?

    1. bananzamba · · focus · HN ↗
      Also doesn't the very good ARC AGI 2 score of GPT-6 Astra kinda contradict this, since each problem is its own game with very different rules
Open on Hacker News to reply ↗

Unofficial Hacker News client; not affiliated with Y Combinator.