If the performances are comparable, and there is no evidence it's not.in/out ($) Gemini : 1.5 / 9.0 | Qwen 3.8: 0.15 / 0.47That is a massive cost reduction.Refs: <a href="https://www.alibabacloud.com/help/en/model-studio/model-pricing#china-beijing-h4" rel="nofollow">https://www.alibabacloud.com/help/en/model-studio/model-pric... <a href="https://runware.ai/gemini-omni" rel="nofollow">https://runware.ai/gemini-omni
You can't just look at the per token cost, but how many tokens it takes on average to do a task. The difference can be massive.
On some benchmarks models like Qwen 3.8 Max which cost < $6/m out cost more than Astra 6 to run at $50/m out. That’s a huge price gap and yet Astra would be cheaper if your work looks like the benchmark.
_ache_ · · focus · HN ↗
in/out ($) Gemini : 1.5 / 9.0 | Qwen 3.8: 0.15 / 0.47
That is a massive cost reduction.
Refs: <a href="https://www.alibabacloud.com/help/en/model-studio/model-pricing#china-beijing-h4" rel="nofollow">https://www.alibabacloud.com/help/en/model-studio/model-pric... <a href="https://runware.ai/gemini-omni" rel="nofollow">https://runware.ai/gemini-omni
killingtime74 · · focus · HN ↗
vntok · · focus · HN ↗
alphabettsy · · focus · HN ↗