‹ BackHN Continuity

Thread

The LLMentalist Effect (2023)

235 points · 309 comments · jalev

  1. vhantz · · focus · HN ↗
    Building up strawmen against LLMs will only make the hypers look more reasonable. A language model responds to inouts exactly as a model of anything else would. That is enough to explain all the "intelligence" without believing a model is somehow a "new kind of mind". Those who think LLMs are intelligent don't know enough about models. And those who think they are useless don't know enough about models.
    1. 10xDev · · focus · HN ↗
      So once we have online learning in LLMs, what do you think you will have that makes you intelligent but not LLMs? Better learning efficiency? That will be improved as well.

      I think we need to start moving on from the term LLMs because it clearly confuses people since they started modelling more than just language.

      1. vhantz · · focus · HN ↗
        What else are they modeling?
        1. 10xDev · · focus · HN ↗
          You realise tokens are just data and data can represent anything.
          1. vhantz · · focus · HN ↗
            Still, large language models model language.
            1. red75prime · · focus · HN ↗
              Almost all of the latest models are MLLMs (multimodal large language models). For exmaple, [1] evaluates GPT-6 Astra on vision tasks.

              [1] <a href="https:&#x2F;&#x2F;blog.roboflow.com&#x2F;gpt-6-astra-vision&#x2F;" rel="nofollow">https:&#x2F;&#x2F;blog.roboflow.com&#x2F;gpt-6-astra-vision&#x2F;

              1. [deleted] · · focus · HN ↗

                [deleted]

Open on Hacker News to reply ↗

Unofficial Hacker News client; not affiliated with Y Combinator.