Prices per 1M tokens Claude Opus 5.5 Claude Opus 5
Cache reads $0.20 $0.50
Input tokens $4 $5
Output tokens $20 $25
Cache writes $5 $6.25
Opus 5 is the model with highest spend on openrouter (<a href="https://openrouter.ai/rankings#task-spend" rel="nofollow">https://openrouter.ai/rankings#task-spend) and it seems plausible that Opus 5 is/was the highest spend model in the world, and certainly Anthropic's biggest moneymaker.
If you are forced to reduce price despite raising capabilities, that certainly tells something about the market, and potentially about Anthropic future profitability too, since this model is their biggest topline contributor
Unfortunately, they're full of it <a href="https://artificialanalysis.ai/models/claude-opus-5-5#token-use" rel="nofollow">https://artificialanalysis.ai/models/claude-opus-5-5#token-u...
It does work out to be a similar cost per task though
so don't use it at max? The benchmarks suggest that high/xhigh are more than sufficient to be ahead and a whole magnitude below max with regards to token usage. I'd treat that as an outlier and not how verbose the model is in general (QED I know)
5.5 is higher for max effort, slightly higher for xhigh and lower for high, medium and low effort.
The biggest proportional difference seems to be at max (5.5 is 38% more) and at high (5.5 is 21% less).
I think most people run at high and xhigh. At xhigh it is close enough to be task dependent and I don't think most people will notice. At high effort I think it looks like it will be an improvement for most people.
5.5 Max should probably be compared to Fable - it performs a lot better than 5 Max.
parent means that they could get more client / a larger part of the market, which would lead to more income (more tokens) despite lower marginal prices
GodelNumbering · · focus · HN ↗
If you are forced to reduce price despite raising capabilities, that certainly tells something about the market, and potentially about Anthropic future profitability too, since this model is their biggest topline contributor
AJ007 · · focus · HN ↗
mcintyre1994 · · focus · HN ↗
drbscl · · focus · HN ↗
It does work out to be a similar cost per task though
jsnell · · focus · HN ↗
<a href="https://artificialanalysis.ai/models/claude-opus-5-5#intelligence-comparisons" rel="nofollow">https://artificialanalysis.ai/models/claude-opus-5-5#intelli...
It is most of the pareto frontier.
drbscl · · focus · HN ↗
93po · · focus · HN ↗
persedes · · focus · HN ↗
drbscl · · focus · HN ↗
persedes · · focus · HN ↗
naasking · · focus · HN ↗
<a href="https://artificialanalysis.ai/models/claude-opus-5-5?models=claude-opus-5-5-high%2Cclaude-opus-5-high#token-use" rel="nofollow">https://artificialanalysis.ai/models/claude-opus-5-5?models=...
piotrdz · · focus · HN ↗
johnbellone · · focus · HN ↗
epolanski · · focus · HN ↗
piotrdz · · focus · HN ↗
piotrdz · · focus · HN ↗
nl · · focus · HN ↗
The biggest proportional difference seems to be at max (5.5 is 38% more) and at high (5.5 is 21% less).
I think most people run at high and xhigh. At xhigh it is close enough to be task dependent and I don't think most people will notice. At high effort I think it looks like it will be an improvement for most people.
5.5 Max should probably be compared to Fable - it performs a lot better than 5 Max.
<a href="https://artificialanalysis.ai/models/claude-opus-5-5?models=claude-opus-5-5%2Cclaude-opus-5-5-xhigh%2Cclaude-opus-5-5-high%2Cclaude-opus-5%2Cclaude-opus-5-xhigh%2Cclaude-opus-5-high%2Cclaude-opus-5-5-medium%2Cclaude-opus-5-5-low%2Cclaude-opus-5-medium%2Cclaude-opus-5-low#intelligence-index-token-use-tabs" rel="nofollow">https://artificialanalysis.ai/models/claude-opus-5-5?models=...
make3 · · focus · HN ↗
meerita · · focus · HN ↗