‹ BackHN Continuity

Thread

GPT 6.1 Sol: Near-Astra intelligence for a fifth of the price

1066 points · 953 comments · crorella

  1. minimaxir · · focus · HN ↗
    > Cached input costs just $0.10 per million tokens—95% less than standard input pricing and 50% less than GPT‑6 Sol’s cached input pricing

    This is the actual big announcement. 50% cheaper cache than GPT-6 Sol will get you far more mileage on Codex.

    1. joshstrange · · focus · HN ↗
      > 50% cheaper cache than GPT-6 Sol will get you far more mileage on Codex.

      Cache doesn't help you much when you are compacting every 5 minutes...

      I was shocked at how quickly I ran out my $100/mo subscription with a single agent (sol medium).

      1. redox99 · · focus · HN ↗
        If you run out of sol medium with $100 you're doing something wrong. Astra destroys your usage, I get 1 day of usage with Astra, but 6 sol is almost unlimited and I only use xhigh.
        1. shimman · · focus · HN ↗
          "You're holding it wrong." Is hardly a retort from a real paying customer having problems with their paid services.

          This is why these companies are struggling to make money, they're chastising their customers just like they've been chastising the human race.

          1. trio8453 · · focus · HN ↗
            > "You're holding it wrong." Is hardly a retort from a real paying customer having problems with their paid services.

            It's very appropriate in the cases when you're holding it wrong. The fact that you're paying doesn't mean that you can't make mistakes or waste resources.

            1. shimman · · focus · HN ↗
              I don't find it appropriate at all, especially regarding technology that workers deeply hate and are skeptical of.

              If this is how you want to get people on your side, I can understand why the entire country/human race are against these companies.

              1. trio8453 · · focus · HN ↗
                Sides? Hate? This is all very emotional. Try to put the facts down plainly and see how ridiculous it is --

                It's a product and if you're using it incorrectly, we can either

                1. say so

                2. pretend that you don't so to get/keep you on "our side"? or not say is because you're skeptical or hate it? (how does that last bit even follow logically?)

                How is 2 better in any way for anyone involved?

            2. crossroadsguy · · focus · HN ↗
              [delayed]
          2. Anonasty · · focus · HN ↗
            Literally the prompting and task definition is main variable how LLM's performs. There are literally millions of examples of vibe coders and new AI adopters who run out of tokens since they don't know how the LLM's work.
Open on Hacker News to reply ↗

Unofficial Hacker News client; not affiliated with Y Combinator.