‹ BackHN Continuity

Thread

The LLMentalist Effect (2023)

235 points · 309 comments · jalev

  1. sethev · · focus · HN ↗
    The challenge here is in trying to decide whether LLMs are intelligent or have a mind. Famously, the criteria for intelligence seem to slip with each advancement in technology. But going back to Turing, his test was actually more carefully phrased than we remember: he said that when machines could pass the test, the question of whether they are intelligent would become moot. That seems to be what we're actually seeing: if people can't tell the difference, it kind of won't matter whether they're "truly intelligent" or not.
    1. sublinear · · focus · HN ↗
      The Turing test was also never meant to be taken so seriously. It's not a rigorous statement of anything.

      Situations like this are precisely why academics tend to avoid the spotlight. You say one slightly off thing and your perceived authority echoes forever with the intellectually lazy.

      1. sethev · · focus · HN ↗
        Yes, the Turing test has been misunderstood for a long time. Turing published it, though - it wasn't some offhand comment he made and it wasn't intellectually lazy.
        1. sublinear · · focus · HN ↗
          Oh, I didn't say Turing was intellectually lazy. :-)
          1. sethev · · focus · HN ↗
            Fair - yes, it is ironic that his point was closer to "we can't possibly know/define whether a machine is intelligent" but somehow it got turned into "Turing's test will tell us when machines are intelligent".
            1. bonoboTP · · focus · HN ↗
              Turing's whole point was to show that it's an uninteresting question of definitions whether a machine can "think", like whether submarines can "swim" and airplanes can "fly". The only important part are observed outcomes and capabilities.
              1. sublinear · · focus · HN ↗
                Yes, but just because language fails to make a distinction doesn't mean there isn't one.

                > The only important part are observed outcomes and capabilities.

                That's wishful thinking. Not even an engineer would say that. The stability of a state is just as important as achieving it. This is trivially and more intuitively demonstrated with other more down-to-earth identity statements such as "I'm a billionaire" and "the building is standing".

                I think we can confidently say LLMs probabilistically achieve a perceived state that is remarkably similar to intelligence, but crumbles upon inspection and seeing it "in motion" so to speak. The same happens to AI-generated images.

                I'm not sure why this sparks so much debate every time. If we're looking for a fountain of "realism", you're not going to beat reality and nature itself. All else will eventually have tells that they are not real.

                1. bonoboTP · · focus · HN ↗
                  It may be of interest to philosophers, but it has little impact on economic job replacement and how people will earn their living and all the downstream upheaval from that. At some point maybe philosophers will find AI-generated philosophical musings about the nature of AI to be better than what comes from their peers (if blinded).

                  It also has little impact on the dangerous use cases.

                  1. sublinear · · focus · HN ↗
                    I think it matters a great deal that LLMs and other "AI" technology are stable within a tolerance that's acceptable.

                    You're jumping the gun talking about "job replacement". We have not thought about it enough from that engineering angle. It's still very early days. That engineering is going to require people. :-)

                    1. bonoboTP · · focus · HN ↗
                      Stability is certainly something you can investigate from behavior.
                      1. sublinear · · focus · HN ↗
                        Yes, and I suppose you believe the AI does this circularly? Perpetual motion with extra steps?
                2. famouswaffles · · focus · HN ↗
                  A property is either consequential or it isn't. If it is then it must be at the very least [1] observable. If it is not, then you just have an imaginary distinction. The instability of an unstable system is something that can be observed if important. Stability is part of "behaviours and outcomes".

                  It's baffling that people cannot see the obvious contradiction in simultaneously asserting a system is missing a crucial property whose effects cannot be observed.

                  [1] It can be invisible. The important part is that its effects are observed.

                  1. sublinear · · focus · HN ↗
                    Like an ancient roman, would you really be willing to stand beneath this arch?

                    The flaws are easily observable to everyone. I said as much in the comment you're replying to. Your argument here isn't going to change that.

                    1. famouswaffles · · focus · HN ↗
                      >Like an ancient roman, would you really be willing to stand beneath this arch?

                      Yeah. Do you think Roman arches were unstable? Guess you don't know much about history too.

                      >The flaws are easily observable to everyone. I said as much in the comment you're replying to. Your argument here isn't going to change that.

                      Then it should be pretty easy to enlighten us. What test of intelligence do LLMs fail that all humans pass ?

      2. NitpickLawyer · · focus · HN ↗
        > The Turing test was also never meant to be taken so seriously.

        citation needed. It has been used as a rubicon for a long time. Ever since Eliza, at least. And there were big headlines and lots of talk around the time LMs became "good enough". I specifically remember when someone had a test done around "a teenager talking in a different language" or somesuch, claiming it was the first time the test was passed.

        It is pretty normal that once it was unquestionably "passed", lots of people started claiming it wasn't even that big of a deal. Tesler's theorem and all that.

        And even if you think the specific formulation of Turing isn't that important (and I'd somewhat agree), you can still use the concept to look at other things. Imagine asking a mathematician 5 years ago the chances of a Erdos problem being solved by a computer end to end. Or a millennium prize. Or ask a swe if a repo could be generated by a computer from the input "write a mario style game", or any other examples of proven expertise.

        1. sublinear · · focus · HN ↗
          > citation needed

          Yes, if you insist on appeals to authority. Authority is a social construct and irrelevant to science.

          Thank you for proving my point.

          1. sethev · · focus · HN ↗
            I think the onus is on you to explain why Turing would publish something he didn’t intend people to take seriously. It seems like an odd claim. Perhaps you mean he didn’t intend it to be interpreted the way it was in popular culture?
            1. sublinear · · focus · HN ↗
              Turing was compelled to address an ongoing debate similar to the same one we're having right now. It continues to do its job as a thought experiment. It's meant to be taken about as seriously as we are right now.

              You either get it, or you don't. Whether machines think is a silly question that deserves its non-answer. We're at the end of what there is to explain, but it was good exposition for the reader.

              <a href="https:&#x2F;&#x2F;www.csee.umbc.edu&#x2F;courses&#x2F;471&#x2F;papers&#x2F;turing.pdf" rel="nofollow">https:&#x2F;&#x2F;www.csee.umbc.edu&#x2F;courses&#x2F;471&#x2F;papers&#x2F;turing.pdf

              Literally the very first opening sentences.

              &gt; I propose to consider the question, &quot;Can machines think?&quot; This should begin with definitions of the meaning of the terms &quot;machine&quot; and &quot;think.&quot; The definitions might be framed so as to reflect so far as possible the normal use of the words, but this attitude is dangerous, If the meaning of the words &quot;machine&quot; and &quot;think&quot; are to be found by examining how they are commonly used it is difficult to escape the conclusion that the meaning and the answer to the question, &quot;Can machines think?&quot; is to be sought in a statistical survey such as a Gallup poll. But this is absurd. Instead of attempting such a definition I shall replace the question by another, which is closely related to it and is expressed in relatively unambiguous words.&quot;

Open on Hacker News to reply ↗

Unofficial Hacker News client; not affiliated with Y Combinator.