I wonder if API is affected by this issue, especially Claude on public clouds? Would that means the subsidized rate just means they use cheaper quantized models and it's not comparable to API spending.
I've always used Enterprise per-token billing for Claude Code and I've never understood these nerf complaints. I've never noticed any slow downs at certain times of day, or a gradual decline in quality.
There’s probably contractual guarantees in the enterprise plans. My understanding of the subscriptions is they can swap the models out if any of them is getting too heavily loaded for a period of time
Open AI aims to have a stable API and admits to meddling with effort levels and such for subscriptions here: <a href="https://news.ycombinator.com/item?id=49804316#49809266">https://news.ycombinator.com/item?id=49804316#49809266
whs · · focus · HN ↗
madeofpalk · · focus · HN ↗
ENGNR · · focus · HN ↗
Computer0 · · focus · HN ↗