‹ BackHN Continuity

Thread

Sonnet 5.5

884 points · 613 comments · D2OQZG8l5BI1S06

  1. Sol- · · focus · HN ↗
    Probably a first world problem, but with Opus 5.5's efficiency, the limits on the 5x plan are simply sufficient for my everyday work, even when running 2-3 sessions at a time. So I wonder when I would use Sonnet 5.5.

    More concurrency than that isn't really practical for me if I want to retain some semblance of understanding. Perhaps it's different for purely web app or frontend tasks, where the outcome is more relevant than the process, I don't have much experience there (and also don't want to belittle these domains, I might be underestimating their complexity).

    So surprisingly, my own work is at least for the time being almost saturated by the model capabilities. I am not sure how I'd scale from here. Sure I could run all requests at max effort to burn tokens for the sake of it, but that can't be it. And for many tasks, I am not really able to define so clear cut success criteria or self-verification loops that I could benefit from letting an agent (or a fleet thereof) autonomously run for a day.

    So I realize it's a skill issue on my side, but I can't be the only one. I wonder if there is a limit to token demand, at least short term. Feels like either they accelerate to AGI and RSI, where the AI can find uses for token, or things might plateau at some point.

    Note I don't think this because I'm an AGI skeptic or think there's a ceiling to intelligence, but there might simply be a valley of economic hardship for the companies where the supply of tokens outpaces the demand, due to a lack of ideas of what to do with them. And this might slow down the funding enough that they never reach escape velocity with the training run scaling. But we'll see.

    1. miki123211 · · focus · HN ↗
      I find that "vibe coders" (that is, people who do not know anything about programming, but nevertheless produce useful tools for themselves and others) are using a lot more tokens than we do as programmers.

      I think this is partially because we're still attached to pre-LLM notions of architecture, good design and code quality (which are still important, but maybe less important than they once were and that we think they are), partially because their projects are in a messy state, so models have to work around the technical dept.

      They're essentially trading off programmer time for LLM time (which is a good trade financially speaking).

      1. meowface · · focus · HN ↗
        I am actually going to go out on a limb and guess the opposite of this is true, and that on average vibecoders burn tokens less readily than veteran software engineers. I could list several reasons why I think this would be likely. No idea which of us is empirically right, though.

        (With exceptions for what I can only call the "manic vibecoders" with like 10 simultaneous weird slopprojects they're spewing out at once. Generally with each project itself being something related to vibecoding. Steve Yegge being an example of a "manic vibecoder-actual programmer" hybrid.)

        1. vineyardmike · · focus · HN ↗
          I’d think this matches my hypothesis. I’d say that I spend more tokens rewriting and fixing things, so that contributes more.

          Also, I’d imagine the token-maxed user is a programmer that lives in chat. I’ll admit to having asked the LLM to move a method up/down in a file, and watched it burn tokens for a minute thinking and executing a menial task.

          1. meowface · · focus · HN ↗
            Same. I spend tokens on so much more than just the initial implementation of a feature I have an idea for.
          2. gbalduzzi · · focus · HN ↗
            This I don't understand. I'm faster at moving the method then at prompting the LLM to do so
            1. vineyardmike · · focus · HN ↗
              Sometimes I don’t have the text editor/IDE open, and I’m just looking at the code in a PR or similar UI.
              1. pmg101 · · focus · HN ↗
                "Sometimes"? I and I think many other people have been doing exactly this for most of 2026, assuming they look at the diff/PR at all and aren't all-in on dark factories.
            2. hasbot · · focus · HN ↗
              Sure, if it's right there in front of you in the editor. But if have to locate the file and the location within the file, it's easier to just type out what I want and have the LLM do it.
            3. avadodin · · focus · HN ↗
              I would be faster than Claude or Gemma if I did it.

              That's a big if though and the blank page syndrome was already getting worse long before AI.

              With age, it becomes easier and easier to get angry at someone or something until they work as expected than it is to actually do it.

              I think this is why we've been seeing the genius coders from two generations ago embracing vibe coding even before it was cool or any good.

            4. VMG · · focus · HN ↗
              LLMs are sometimes stupid and get confused when you change the files without them knowing. So asking them to do even simple things keeps the context in sync.

              Plus you get a bonus random line "methods are all on the top" in the commit message that makes no sense to anybody.

Open on Hacker News to reply ↗

Unofficial Hacker News client; not affiliated with Y Combinator.