‹ BackHN Continuity

Thread

Getting the most out of Opus 5.5 in Claude and Claude Code

232 points · 156 comments · saikatsg

  1. epistasis · · focus · HN ↗
    I really hate long tasks. Claude never gets things right, at least for me, and wastes tons of time when a simple question would have gotten me to the right result rather than several turns of correcting bad decisions in addition to the long amounts of wasted thinking time.

    What sort of workloads do well with these long tasks? The big labs are optimizing for long run time on their own, but it seems like a terrible thing to optimize on unless you're trying to do something like prove a hard math theorem, which success is clearly defined and the route doesn't matter a ton.

    Plan mode has been made increasingly useless. I need to discuss to iterate to get the desired design, explore options, because Claude never gets it right first try and I don't have enough knowledge of options to specify everything up front.

    Ah well, the Chinese models will still work well, I guess.

    1. istjohn · · focus · HN ↗
      Look up the grilling skill[0]. To make it even better, tell Claude to modify it to use the AskUserQuestion functionality. It's so much better than plan mode.

      0. <a href="https:&#x2F;&#x2F;github.com&#x2F;mattpocock&#x2F;skills&#x2F;blob&#x2F;main&#x2F;skills&#x2F;productivity&#x2F;grilling&#x2F;SKILL.md" rel="nofollow">https:&#x2F;&#x2F;github.com&#x2F;mattpocock&#x2F;skills&#x2F;blob&#x2F;main&#x2F;skills&#x2F;produc...

    2. ricardobeat · · focus · HN ↗
      Plan mode has become pointless since Opus 5 came out, they know when to switch between planning and execution now. But that iteration&#x2F;discussion is still necessary unless you&#x27;re building completely blind - the model cannot read your mind.

      I&#x27;ve had it running 8h+ of non-stop optimizations, chasing a performance target, rewriting systems or building a series of prototypes for research. All it needs is a clear goal.

      1. epistasis · · focus · HN ↗
        My experience is that Claude has gotten absolutely terrible at switching between planning and execution, and never gets it right, and then I waste tons of turns fixing intent, and trying to make claude forget the bad shit it did.

        There needs to be a mode where it&#x27;s &quot;don&#x27;t change code, don&#x27;t lock in decisions, lets explore&quot;, and the problem with plan mode is that it all of a sudden presents a too-long, multi-screen plan that&#x27;s just completely off base with basically two option: &quot;go and do it all&quot; or &quot;tell me what&#x27;s wrong and then I&#x27;ll make a small modification on two screenfulls of text and not tell you what I changed.&quot;

        If it works for you, great, but Claude was far better for me before 5. it got slightly better with 5.5, but the harness deteriorates every days as programmers try to gain internal clout by shoving in poor features.

        Pi is such a relief in comparison to Claude, those who can use it really should.

    3. Merad · · focus · HN ↗
      Use the superpowers plugin. It&#x27;s a game changer. You start by building out a spec defining the work you want to be done, then claude writes a detailed implementation plan. It can easily work for hours without interruption, and the process integrates testing (TDD when possible) and review steps along the way. I even did one fairly complex POC for work that had claude implementing for almost 3 days straight, with me stopping it and handing off to a new session when context hit about 800k. The process is slower and more methodical than the way claude works by default but the output is much better.
Open on Hacker News to reply ↗

Unofficial Hacker News client; not affiliated with Y Combinator.