‹ BackHN Continuity

Thread

Claude Opus 5.5

1806 points · 1134 comments · km144

  1. sailingparrot · · focus · HN ↗
    > Claude Opus 5.5 is our first release since we called for pacing the frontier.

    Interesting how the very first line is used to remind the reader of their call to pace the frontier just last week, and everything else after that line is to demonstrate with very specific numbers how they absolutely are not pacing.

    1. BodyCulture · · focus · HN ↗
      It is much more important to discuss the fact that the models still produce wrong results.

      This type of product would not have been published in a world where we assumed that computers must produce reliably correct results.

      1. ordersofmag · · focus · HN ↗
        Spoken as someone who has never seen a weather forecast. Computers have long produced results that were, if interpreted naively, incorrect. These systems require humans to take into account the conditional, probabilistic nature of the output and the inherent limitations of the model the computer was executing, when they interpret the output. LLM's are no different. They give you next token prediction probabilities. That is all. There is no particular reason to think that comes (or even could come) with a promise that those tokens will express truths about the world and only truths. The fact that greed-driven hype, and our human inclination to anthropomorphize the token generator inclines us to imagine things should be different doesn't make it so. No model output is truth. Never has been.
Open on Hacker News to reply ↗

Unofficial Hacker News client; not affiliated with Y Combinator.