I don't see how anyone can be using Claude with prices like this, it's pretty incredible what the OpenAI team is doing, w.r.t model quality and pricing.
Prices per 1M tokens Claude Opus 5.5 Claude Opus 5
Cache reads $0.20 $0.50
Input tokens $4 $5
Output tokens $20 $25
Cache writes $5 $6.25
Let's look at open-weights models with 3T size: <a href="https://inferencex.semianalysis.com/run/kimi-k3-on-b200" rel="nofollow">https://inferencex.semianalysis.com/run/kimi-k3-on-b200
This suggests inference margins in the ballpark of 98% if we assume 5.6 Sol is about as efficient to serve as Kimi K3.
We also do not know what efficiency improvements have been made with GPT 6 Sol and Luna.
There is some speculation that 6 Sol could be a smaller model comparable in size to 5.6 Terra, and that this is why the improvement in intelligence is modest over 5.6 Sol.
This would line up with a faster serving speed and benchmarks that show a small improvement in coding tasks with regressions in knowledge tasks.
pookieinc · · focus · HN ↗
Input
Output
Price reduction
GPT‑6 Sol vs. GPT‑5.6 Sol
$4 → $2
$20 → $10
50% cheaper
GPT‑6 Luna vs. GPT‑5.6 Luna
$0.20 → $0.10
$1.20 → $0.50
50% cheaper
minimaxir · · focus · HN ↗
user43928 · · focus · HN ↗
Let's look at open-weights models with 3T size: <a href="https://inferencex.semianalysis.com/run/kimi-k3-on-b200" rel="nofollow">https://inferencex.semianalysis.com/run/kimi-k3-on-b200
This suggests inference margins in the ballpark of 98% if we assume 5.6 Sol is about as efficient to serve as Kimi K3.
We also do not know what efficiency improvements have been made with GPT 6 Sol and Luna.
There is some speculation that 6 Sol could be a smaller model comparable in size to 5.6 Terra, and that this is why the improvement in intelligence is modest over 5.6 Sol.
This would line up with a faster serving speed and benchmarks that show a small improvement in coding tasks with regressions in knowledge tasks.