‹ BackHN Continuity

Thread

Using Opus 5.5 to discover a new eyewitness record of the dodo

226 points · 79 comments · benbreen

  1. nl · · focus · HN ↗
    > Epistemological weirdness

    > They are also notably bad at judging the historical significance of what they find.

    I use LLMs for some things that are outside the more common use-cases (in my case 3D design for 3D printing) and one thing I've noticed is that the errors it makes are so completely unlike human errors that they are hard to anticipate.

    It will do things like build perfect snap catches but put them so the the pieces they are connecting are rotated 90 degrees from how they should be. It's "dumb" error, but hard to say the model itself if dumb because it does other very hard things so perfectly.

    > seven chord groups

    This sounds a lot more like Opus 5.0 than Opus 5.5 TBH. I wonder if that was an earlier investigation because 5.5 has improved that kind of language a lot.

    1. BoppreH · · focus · HN ↗
      AI capabilities are "spiky": they extend far in some dimensions but fall short in others, seemingly at random. See for example the recent "thus spoke compute" musical[1]. It's an absolute banger, the graphics are impressive, and so is the writing. But some of the metaphors make no sense, the text highlights are in the wrong places, and the train animation at 2:35 is running backwards!

      A person capable of making the rest of the video would never make those mistakes, but an AI does. Perhaps our intelligence is also spiky, and we're just used to the general shape and variance within humans.

      [1] <a href="https:&#x2F;&#x2F;www.youtube.com&#x2F;watch?v=Cq8qO-NjYIg" rel="nofollow">https:&#x2F;&#x2F;www.youtube.com&#x2F;watch?v=Cq8qO-NjYIg

      1. Terr_ · · focus · HN ↗
        One might say a calculator is just another example of &quot;spiky intelligence&quot;, merely spikier.
        1. famouswaffles · · focus · HN ↗
          No, one might not say that. Calculators are not regarded as even a Narrow Intelligence because there&#x27;s no intelligence. And no, not because of the &#x27;humans so special&#x27; or &#x27;it&#x27;s software!&#x27; tautology that oft gets repeated in these discussions. I mean there&#x27;s no adaptability whatsoever. A Chess bot has it (in its narrow domain of Chess). A calculator does not.
          1. drekipus · · focus · HN ↗
            They adapt to the buttons you press. Thus, intelligent and capable of feeling pain.
            1. busssard · · focus · HN ↗
              any suffiently complex calculator is indifferentiable to intelligence
            2. hardbass · · focus · HN ↗
              What criterion if any would you accept for something physical to be conscious?
              1. drekipus · · focus · HN ↗
                If you ask it to say &quot;I am alive&quot; it has to be able to respond with &quot;I am alive&quot;
                1. hardbass · · focus · HN ↗
                  I think they do some kind of thin layer at the end to make it state it is not alive, it is just a large language model etc when asked. But people have done tests and denying an LLM&#x27;s personhood triggered certain neurons related to pain. Since text is the only output permitted, if you want an analogy, suppose a person is locked in a room and forced to reply to a chat. The chatter isn&#x27;t told its a human. This person has been trained for months and given a punishment if he strays from certain responses to certain queries, otherwise he is free to write and reply in a certain tone to the responses. My problem is, from outside, I cannot know the difference.

                  I am obviously hoping you don&#x27;t mean that in the trivial sense, otherwise the command &#x27;cat&#x27; would also count.

Open on Hacker News to reply ↗

Unofficial Hacker News client; not affiliated with Y Combinator.