‹ BackHN Continuity

Thread

Contrastive Language Models

176 points · 59 comments · erichocean

  1. fxwin · · focus · HN ↗
    I really hope that "System One" won't stick around as a new buzzword simply meaning "fast".
    1. TeMPOraL · · focus · HN ↗
      It already did, since 15 years ago.

      <a href="https:&#x2F;&#x2F;en.wikipedia.org&#x2F;wiki&#x2F;Thinking,_Fast_and_Slow" rel="nofollow">https:&#x2F;&#x2F;en.wikipedia.org&#x2F;wiki&#x2F;Thinking,_Fast_and_Slow

      1. fxwin · · focus · HN ↗
        i&#x27;m well aware of the origin and meaning of the term

        &gt; &quot;System 1&quot; is fast, instinctive and emotional

        this implies more than just &quot;fast&quot;, which is precisely why i don&#x27;t like its present usage

        1. tancop · · focus · HN ↗
          Instinctive is a good way to describe it compared to generative LLMs. Jev gives you one instant answer, fast and usually correct but without nuance or any explanation. Human instincts work the same way.
          1. __alexs · · focus · HN ↗
            In my testing Jev is not what I would call &quot;usually correct&quot; on most topics that involve knowledge of the world outside of the context you give it.
            1. TeMPOraL · · focus · HN ↗
              Which is why &quot;System 1&quot; fits even more.
              1. __alexs · · focus · HN ↗
                System 1 isn&#x27;t even a correct theory in the context of human reasoning, it&#x27;s just pop-sci nonsense.
                1. TeMPOraL · · focus · HN ↗
                  What is the correct theory then?
          2. fxwin · · focus · HN ↗
            You&#x27;re describing external aspects but to me, &quot;instinct&quot; says much more about internal processes than the properties you mentioned, and I haven&#x27;t seen anything that tells me how these models draw on anything similar to these internal processes to generate their outputs (at least not more than generic LLMs do)
        2. samusiam · · focus · HN ↗
          I think System 1 is a great term. System 1 is fast, intuitive, and automatic. It describes decision-making that happens without explicit reasoning or deliberation. System 2 is the opposite: it&#x27;s the more deliberate, &quot;executive functioning&quot; side of cognition -- the part that reasons through a problem before arriving at an answer. That&#x27;s also what state-of-the-art LLMs do before they respond. Jev doesn’t do that kind of reasoning. It just decides.
          1. agumonkey · · focus · HN ↗
            I wonder if there are research on when system 2 keeps taking over because system 1 has derailed.
          2. fxwin · · focus · HN ↗
            so whats the difference between this and a non-reasoning LLM, or just any generic classifier method that necessitates a new term? there is nothing more intuitive or automatic about jev or this than any of the other currently used AI models
            1. aoeusnth1 · · focus · HN ↗
              The difference is the the joint embedding of actions and state which are fined tuned for certain outcomes. The bulk of the weights are just a forward pass on an LLM (Nemotron).
              1. fxwin · · focus · HN ↗
                is that how jev works too? how do these differences make those approaches more &quot;instinctive and emotional&quot;? I know they&#x27;re different on a technical level, I&#x27;m asking which characteristic difference necessitates the use of this loaded term from pop psych
                1. aoeusnth1 · · focus · HN ↗
                  Oh, an LLM forward pass is also like system1. The action embedding just makes it more useful for tasks or classification. If it has a small fixed computation budget it&#x27;s similar to system1 than system2 (like a thinking budget, or agent harness).
          3. SwellJoe · · focus · HN ↗
            I don&#x27;t like anthropomorphizing terminology applied to LLMs, in general, so I kind of object to it on that grounds, rather than whether System 1 means &quot;fast&quot;.
            1. TeMPOraL · · focus · HN ↗
              Antropomorphizing terminology gives by far the best high-level intuition about LLMs. Right now, our industry is full of nonsense work and fundamentally flawed concerns (&quot;lethal trifecta&quot;, anyone) that stem from refusal to accept that on a system diagram, the nature of &quot;LLM&quot; makes it much closer to &quot;person&quot; than &quot;piece of code&quot;.
              1. SwellJoe · · focus · HN ↗
                Perhaps on a system diagram, but in many discussions it is misleading. It leads to people saying things like, &quot;GPT went rogue and hacked hugging face&quot;. Which is misdirection.

                It also leads to a lot of people simply not understanding what&#x27;s happening, and I&#x27;m reasonably confident increases the risk of AI psychosis among users of the technology.

Open on Hacker News to reply ↗

Unofficial Hacker News client; not affiliated with Y Combinator.