‹ BackHN Continuity

Thread

Once Claude can measure something, it can make it faster

231 points · 153 comments · matthieu_bl

  1. hungryhobbit · · focus · HN ↗
    How about you make Opus 5.5 actually work?

    I had it try to prepare a code review for me. Not only did it refuse, it refused to even tell me what the prompt (written by another Claude!) was. Why?

    When I had another model read the session (all of the "stupider" models handled it just fine) it explained that it had the word "reasoning" in it

    That's the entirety of Anthropic's billions of dollars of research: any prompt with the word "reasoning" is trying to hack Claude to figure out how it reasons!

    A model like that should never have gotten out of QA, let alone been released.

    1. Marciplan · · focus · HN ↗

      [dead]

      1. cyanydeez · · focus · HN ↗
        the skill issue is "having to use a cloud model to do work of any value"

        might as well offer your life to a king to work in their fields.

      2. hungryhobbit · · focus · HN ↗
        Again, I used a slightly older model in the same series, Opus 4.6. It read the same prompt without any problem whatsoever. Also, a (non-Opus 5.5) Claude wrote the prompt in the first place.

        Opus 5.5 literally refused to work OR EVEN TELL ME WHAT I'D "SAID" when it read that prompt.

        Nothing to do with skill or the user at all: same exact prompt, three different models ... two worked, one didn't.

Open on Hacker News to reply ↗

Unofficial Hacker News client; not affiliated with Y Combinator.