‹ BackHN Continuity

Thread

Don't be fooled–LLMs don't reason

75 points · 159 comments · leopoldj

  1. f6v · · focus · HN ↗
    > Knowledge and reasoning are inextricably interwoven in the weights of the neural network—there is no independent, explicitly represented set of beliefs.

    I'm not sure I understand this. Do we have evidence humans have an independent set of beliefs not shaped by knowledge and reasoning? If so, where do these come from?

    I'm especially confused about a prior statement as a scientist:

    > Three shortcomings prevent what chatbots do from qualifying as reasoning (in a way that a scientist might recognize).

    How does a set of beliefs help with reasoning?

    > Third, while the chains of thought chatbots produce look like deliberation, research has demonstrated that the bots often concoct them after the fact, reaching an answer by one route but reporting another.

    We also often do the same as humans.

    1. imightbebatman · · focus · HN ↗
      Yea.

      I think a better argument for them not reasoning is that they are incapable -- seemingly, for now, that I know of -- of independently starting a reasoning task.

      They do not ask whether or not they dream of electric sheep. Unless prompted to do so.

      That's part of human reasoning, being prompted, but so is independent thought.

      So at best they are partially reasoning, unless we philosophically argue these are two separate things. And frankly, shrug, I'm not a philosopher.

      1. pants2 · · focus · HN ↗
        In the early reasoning models the chain of thought would often drift off into other topics, almost like a daydreaming. This behavior is undesirable so that's been RL'd away, but I do think this is something we could evoke from models. It's just difficult to contend with from an eval perspective.
        1. imightbebatman · · focus · HN ↗
          This is a good point and we would have to understand whether this is noise in the machine or independent reasoning. I haven no idea how to measure that.

          If somehow independent reasoning then we're getting very close to something alive that would need rights.

          In my view.

          And I have yet to have an experience with an LLM where I came away thinking -- this thing is alive and me using it this way is unethical. I've had that experience with animals, even people. Never -- yet -- an LLM.

          1. qarl · · focus · HN ↗
            > And I have yet to have an experience with an LLM...

            I hope you agree that this question is too important to leave to gut feelings.

            Humans have a poor track record in that regard.

            1. imightbebatman · · focus · HN ↗
              No question. As we were discussing it needs measured somehow.
              1. qarl · · focus · HN ↗
                Right. So I'm wondering about the relevance of you citing your gut feelings. It seems like you are putting stock in them. No?
      2. strbean · · focus · HN ↗
        > they are incapable [...] of independently starting a reasoning task.

        This is a simple and intentional design choice though. Biological lifeforms are always on and always receiving sensory input. LLMs don't functionally exist out side of when we decide to run them. An always on agent with a looping prompt of "if you aren't doing anything else, ruminate" overcomes this limitation.

        1. imightbebatman · · focus · HN ↗
          Not really. Noone has to prompt you to ruminate.

          In fact, if your brain were removed from your body, and you were locked in a room that was completely empty and precisely calibrated to your brain's ambient temperature, and you were then ordered on your life to not ruminate. Well -- I would posit that you wouldn't last very long.

          There's something else going on with human cognition -- and reasoning by extension -- that LLMs are, to date, not replicating.

          With the big caveat of AFAIK. I'm not in one of the labs close to this stuff.

          1. qarl · · focus · HN ↗
            > Not really. Noone has to prompt you to ruminate.

            Why do the internal mechanics need to mirror how it works in humans?

            I have agents that receive sensory input (video/audio) which run continuously processing it, receiving events as they occur. When something strange happens they notice and take action (i.e. there's an unrecognized person in the living room -> ask who they are.)

            Presumably you've seen the agents that play video games - build things to progress in Minecraft, plan routes, avoid adversaries, etc.

            Why doesn't that count?

            I feel like you're failing to consider anything outside the chatbot UI.

            EDIT: The first morning my agents had access to an outside video stream they commented on the sunrise. While I was sleeping through it.

            1. imightbebatman · · focus · HN ↗
              You make a lot of out of context assumptions and feel.
              1. qarl · · focus · HN ↗
                I'm sorry? I'm answering your claim:

                > they are incapable ... of independently starting a reasoning task.

                I have provided several examples of them independently starting reasoning tasks.

                ... and I'm wondering why you seem to be entirely unaware of such examples. The most likely explanation being you are only aware of the chat interface.

                ... and I have absolutely no idea of what you mean by "... and feel".

      3. f6v · · focus · HN ↗
        I think you’re on the right path. My first thought was whether lack of constant sensory input is an issue. We’re always feeling something (save for sensory deprivation chambers, but even then…)

        Or put it other way, lack of feedback loops between the “brain” and the environment.

Open on Hacker News to reply ↗

Unofficial Hacker News client; not affiliated with Y Combinator.