‹ BackHN Continuity

Thread

Using Opus 5.5 to discover a new eyewitness record of the dodo

226 points · 79 comments · benbreen

  1. nl · · focus · HN ↗
    > Epistemological weirdness

    > They are also notably bad at judging the historical significance of what they find.

    I use LLMs for some things that are outside the more common use-cases (in my case 3D design for 3D printing) and one thing I've noticed is that the errors it makes are so completely unlike human errors that they are hard to anticipate.

    It will do things like build perfect snap catches but put them so the the pieces they are connecting are rotated 90 degrees from how they should be. It's "dumb" error, but hard to say the model itself if dumb because it does other very hard things so perfectly.

    > seven chord groups

    This sounds a lot more like Opus 5.0 than Opus 5.5 TBH. I wonder if that was an earlier investigation because 5.5 has improved that kind of language a lot.

    1. Waterluvian · · focus · HN ↗
      This analogy may be too close to the real thing to work, but it reminds me of a Chinese room type situation where its entire understanding of the world is through messages of text.

      You say that’s an error a human couldn’t do, but imagine if the human has never seen or touched the kind of item you were making and relied entirely on text descriptions to build its ontology. Off by 90 seems like such a believable mistake.

      1. meowface · · focus · HN ↗
        Also, not to sound like a naive hypemonger, but: in a decade I'd bet a ton of money the best AI systems will make strange mistakes of this nature at a far, far lower rate than they do today. They will gain a more holistic and more human-like perspective about each task.

        (even if it's through some silly means like explicitly talking to themselves like "if I were a human doing this, what [... 5 million tokens in 2 seconds ...]" but also of course if they crack ASI and get something more efficient and intelligent than a human brain by then)

        1. xmprt · · focus · HN ↗
          I think of it kind of like how Chess AI make "mistakes" which are unrecognizable to humans but a stronger AI would be able to pick them apart. That's kind of scary...
      2. ZeroGravitas · · focus · HN ↗
        There's a famous early world map created by Ptolemy from compiling reports of sailors which is amazingly accurate for its time but had the country of Scotland off by 90 degrees.
Open on Hacker News to reply ↗

Unofficial Hacker News client; not affiliated with Y Combinator.