‹ BackHN Continuity

Thread

GPT-6 Sol and Luna

1779 points · 855 comments · OfficialTurkey

  1. pookieinc · · focus · HN ↗
    I don't see how anyone can be using Claude with prices like this, it's pretty incredible what the OpenAI team is doing, w.r.t model quality and pricing.

      Prices per 1M tokens     Claude Opus 5.5    Claude Opus 5
       Cache reads              $0.20              $0.50
       Input tokens             $4                 $5
       Output tokens            $20                $25
       Cache writes             $5                 $6.25
    
    
    
    Model

    Input

    Output

    Price reduction

    GPT‑6 Sol vs. GPT‑5.6 Sol

    $4 → $2

    $20 → $10

    50% cheaper

    GPT‑6 Luna vs. GPT‑5.6 Luna

    $0.20 → $0.10

    $1.20 → $0.50

    50% cheaper

    1. minimaxir · · focus · HN ↗
      I legit question if these prices are still inference-profitable for OpenAI. They likely didn't have 100% profit margin.
      1. user43928 · · focus · HN ↗
        If they had a 100% margin the cost would be 0.

        Let&#x27;s look at open-weights models with 3T size: <a href="https:&#x2F;&#x2F;inferencex.semianalysis.com&#x2F;run&#x2F;kimi-k3-on-b200" rel="nofollow">https:&#x2F;&#x2F;inferencex.semianalysis.com&#x2F;run&#x2F;kimi-k3-on-b200

        This suggests inference margins in the ballpark of 98% if we assume 5.6 Sol is about as efficient to serve as Kimi K3.

        We also do not know what efficiency improvements have been made with GPT 6 Sol and Luna.

        There is some speculation that 6 Sol could be a smaller model comparable in size to 5.6 Terra, and that this is why the improvement in intelligence is modest over 5.6 Sol.

        This would line up with a faster serving speed and benchmarks that show a small improvement in coding tasks with regressions in knowledge tasks.

Open on Hacker News to reply ↗

Unofficial Hacker News client; not affiliated with Y Combinator.