‹ BackHN Continuity

Thread

Prompting Claude Opus 5.5

207 points · 227 comments · Michelangelo11

  1. TheAceOfHearts · · focus · HN ↗
    One of my key complaints with Opus 5.5 so far has been that sometimes it'll execute long-running commands in a way that is blocking any further input or it starts doing stuff without providing much visibility. I've tried giving it instructions to stop doing that but it keeps falling into the same trap.

    I feel like hybrid AI-driver UIs are a bit underexplored and are probably a good way to increase visibility. Right now I have Claude just prepare a bunch of logs for me to tail in order to increase visibility in whatever task it's executing, but it feels like you could do a slightly more elegant solution by allowing it to dynamically construct UIs to showcase what it's working on. Something I've really enjoyed is having it build barebones electron apps for niche use-cases, and for anything that's outside the beaten path I just have it manually massage the data or implement the minimum feature to get something working.

    Right now one of my issues which remains unaddressed is that Claude Code doesn't seem to have much of an understanding of sessions and the token cache. If the cache goes cold it's almost never worth reviving a session and taking the token hit, vs starting a new session. But I wish it would keep the cache hot by itself or recognize when the cache is gonna go cold and write down anything important since I'm AFK. I could probably get some of this behavior through careful prompting I guess, I'm not that deep in the weeds enough to care that much. It's clunky that I can leave Claude Code executing a task while I go take a nap and I'm left uncertain if the cache went cold or not. I'd really like a gated "Are you sure?" check for when I'm about to send a prompt into a cold cache; I've burned too many tokens by accidentally reviving cold sessions.

    1. Lucasoato · · focus · HN ↗
      > One of my key complaints with Opus 5.5 so far has been that sometimes it'll execute long-running commands in a way that is blocking any further input or it starts doing stuff without providing much visibility.

      Is this a problem with the model or the harness in your opinion?

      1. TheAceOfHearts · · focus · HN ↗
        I have no idea how to evaluate this, but I've had it write down the same rule like 3 times and it still keeps managing to fall into this trap of blocking on commands. The fact that it refuses to adhere to my guidelines and rules is probably a model issue, but a better harness could probably overcome the issues.
        1. mnicky · · focus · HN ↗
          When it launches command in a blocking shell, just press something like ctrl+b and this sends the shell to the background and you can continue to use the agent..
Open on Hacker News to reply ↗

Unofficial Hacker News client; not affiliated with Y Combinator.