‹ BackHN Continuity

Thread

GPT-6 Sol and Luna

1779 points · 855 comments · OfficialTurkey

  1. pookieinc · · focus · HN ↗
    I don't see how anyone can be using Claude with prices like this, it's pretty incredible what the OpenAI team is doing, w.r.t model quality and pricing.

      Prices per 1M tokens     Claude Opus 5.5    Claude Opus 5
       Cache reads              $0.20              $0.50
       Input tokens             $4                 $5
       Output tokens            $20                $25
       Cache writes             $5                 $6.25
    
    
    
    Model

    Input

    Output

    Price reduction

    GPT‑6 Sol vs. GPT‑5.6 Sol

    $4 → $2

    $20 → $10

    50% cheaper

    GPT‑6 Luna vs. GPT‑5.6 Luna

    $0.20 → $0.10

    $1.20 → $0.50

    50% cheaper

    1. an0malous · · focus · HN ↗
      These are the pre rug pull prices. They'll increase prices 10x and nerf the models after they IPO.
      1. solenoid0937 · · focus · HN ↗
        Before IPO. This is why Anthropic isn't playing the same games
      2. blovescoffee · · focus · HN ↗
        there are still competitive market forces for co's post IPO
      3. selectodude · · focus · HN ↗
        Okay? I didn’t sign a 10 year contract. We’re month to month and I use my own harness.

        If they’re subsidizing my usage, that’s great.

        1. infinitezest · · focus · HN ↗
          You're building your livelihood/workflows on a set of inputs that you have no idea what they actually cost or how reliable they'll be when the VC cash stops flowing. If you're OK with that, do your thing but it seems a little foolish to me.
          1. derac · · focus · HN ↗
            If the market crashes they will be much cheaper to run actually, no? Hardware would flood the market.
            1. ssl-3 · · focus · HN ↗
              [delayed]
            2. foepys · · focus · HN ↗
              I wouldn't bet on hardware flooding the market. I bet the machines running in the data centers don't use traditional PCIe connectors and cards. Maybe somebody could pull the chips and put them on standardized PCIe cards, but that is not a given.
              1. Leynos · · focus · HN ↗
                It happens already. These are plenty of cheap V100s on eBay, and PCIE to SXM2 adapters

                External example: <a href="https:&#x2F;&#x2F;ebay.io&#x2F;m&#x2F;lV8UsD" rel="nofollow">https:&#x2F;&#x2F;ebay.io&#x2F;m&#x2F;lV8UsD

                Internal example: <a href="https:&#x2F;&#x2F;ebay.io&#x2F;m&#x2F;z1ygRU" rel="nofollow">https:&#x2F;&#x2F;ebay.io&#x2F;m&#x2F;z1ygRU

                V100s are three generations behind current and missing many of the features that modern inference benefits from, but they are the cheapest way to get a 32GB gpu.

          2. fragmede · · focus · HN ↗
            It seems silly to say we have no idea when we actually do, though. We know how much hardware costs, we know how to reliably run a webservice that hits an API hosted on a machine with a GPU, we know how to operate these things at scale outside of OpenAI and Anthropic (not Nvidia). VC money can be patient, Uber&#x27;s profitable, yeah $1 Uber rides got us hooked and they&#x27;re running the same playbook. Unfortunately the convenience is worth paying for, so it seems dumb to think we can control the beast or ignore it, or get everyone to agree to hold back.

            Is there a world where OpenAI starts charging $2,000&#x2F;month for what we previously were paying $20 for? What are we going to do? AWS could totally jack up the prices for EC2 instances as well, but we&#x27;ve come to rely on that as well.

          3. selectodude · · focus · HN ↗
            Push comes to shove, OpenAI could go out of business tomorrow and I could pick up roughly where I left off for $25k, which is the cost to serve GLM 5.3 Flash on four Nvidia GB10s. Granted, if OpenAI et al go kaput all at the same time, I could probably get a whole lot more compute for a whole lot less money.
          4. slopinthebag · · focus · HN ↗
            huh? i use the plans because they&#x27;re cheap and i get strong models, but i could go back to deepseek flash on commodity api pricing and be just fine
          5. goosejuice · · focus · HN ↗
            [delayed]
          6. andybak · · focus · HN ↗
            I&#x27;m fairly sure most open weight model providers are serving them at a sustainable price - and I&#x27;ve used them enough to know that I could live with them if the big boys did a rug pull.
      4. minimaxir · · focus · HN ↗
        That would only work if OpenAI were a monopoly, which they are not.
Open on Hacker News to reply ↗

Unofficial Hacker News client; not affiliated with Y Combinator.