Gemini 4 Argon (High): Intelligence, Performance and Price Analysis
Thread
Loading the complete thread in the background. This saved snapshot is available now. Refresh
Unofficial Hacker News client; not affiliated with Y Combinator.
Gemini 4 Argon (High): Intelligence, Performance and Price Analysis
Loading the complete thread in the background. This saved snapshot is available now. Refresh
Unofficial Hacker News client; not affiliated with Y Combinator.
anuragdaram · · focus · HN ↗
krat0sprakhar · · focus · HN ↗
<a href="https://blog.google/innovation-and-ai/models-and-research/gemini-models/gemini-4-argon/" rel="nofollow">https://blog.google/innovation-and-ai/models-and-research/ge...
1/5th the price of Astra and Fable
sroussey · · focus · HN ↗
anuragdaram · · focus · HN ↗
A_D_E_P_T · · focus · HN ↗
thereitgoes456 · · focus · HN ↗
asdfasgasdgasdg · · focus · HN ↗
godbox · · focus · HN ↗
yipinwong · · focus · HN ↗
The price is enticing for cost per tasks, but let's see how it goes.
I have montly (cheapy) sub to gemini models and has been underwelming and lowered the tier.
dom96 · · focus · HN ↗
godbox · · focus · HN ↗
enraged_camel · · focus · HN ↗
losvedir · · focus · HN ↗
augment_me · · focus · HN ↗
sourweasel · · focus · HN ↗
piyh · · focus · HN ↗
$10 per million output tokens isn't improving on frontier price?
godbox · · focus · HN ↗
netdur · · focus · HN ↗
cindyllm · · focus · HN ↗
[dead]
dzhiurgis · · focus · HN ↗
ai-x · · focus · HN ↗
So, like any optimal game theory move, they are better off not starting a price war
pooper · · focus · HN ↗
ai-x · · focus · HN ↗
a) internal use b) embedding in their products c) stay abreast with capabilities
they can sit it out.
miohtama · · focus · HN ↗
blinding-streak · · focus · HN ↗
<a href="https://firstpagesage.com/reports/top-generative-ai-chatbots/" rel="nofollow">https://firstpagesage.com/reports/top-generative-ai-chatbots...
<a href="https://techcrunch.com/2026/06/16/chatgpts-market-share-slips-below-50-for-first-time/" rel="nofollow">https://techcrunch.com/2026/06/16/chatgpts-market-share-slip...
gradus_ad · · focus · HN ↗
OpenAI and Anthropic are existential threats to Google and it will operate accordingly.
notatoad · · focus · HN ↗
but that's not a battle between astra and gemini pro, it's a battle between luna and gemini flash-lite.
aliljet · · focus · HN ↗
tomrod · · focus · HN ↗
mkotlikov · · focus · HN ↗
tomrod · · focus · HN ↗
notatoad · · focus · HN ↗
the deal has a 30-day cancellation policy, and they raised a bunch of debt around the same time to fund their own datacenter expansion.
fragmede · · focus · HN ↗
genxy · · focus · HN ↗
largbae · · focus · HN ↗
zozbot234 · · focus · HN ↗
mpyne · · focus · HN ↗
If it's smart enough to do the job then it won't matter that Opus is smarter. At the right price and performance, at least.
petesergeant · · focus · HN ↗
I pay for lots of models because they’re good at different things: $20 a month each for Grok and GLM have easily paid for themselves by finding bugs that my main work models didn’t, but I’m yet to have any Gemini model find a real bug, and Gemini’s results for general work will sometimes border malicious compliance, when it’s not having a hissy fit about some imagined issue.
jug · · focus · HN ↗
Honestly, I think this will be a trend in 2027 when all those models become "good enough" for elite coding and whatever. I predict they'll have to branch out more and build their platform to differentiate themselves from each other, maybe even in terms of branding and trust.
aleqs · · focus · HN ↗
Anthropic has some of the lowest usage per $ in general, not sure what you're taking about.
jjice · · focus · HN ↗
aleqs · · focus · HN ↗
8n4vidtmkvmk · · focus · HN ↗
It's all relative. Some people just can't use up their quotas with their normal usage.
aleqs · · focus · HN ↗
UltraSane · · focus · HN ↗
nl · · focus · HN ↗
Opus and Sol usage levels vs the API are currently roughly the same, but Opus 5.5 outperforms at low and medium effort levels.
sroussey · · focus · HN ↗
Melatonic · · focus · HN ↗
nl · · focus · HN ↗
Sol 6.1 scores one point less than Gemini 4 on intelligence AND costs less than half ($0.72 vs $1.99) per task.
Additionally, if you are using OpenAI you have the option to pay a bit more and get Astra which - despite the benchmarks - does outperform Sol on some things.
Also, people are - rightly - very wary of Google's benchmaxxing tendencies. I think lots of people remember Gemini 3.0 (I think?) which benchmarked amazingly, but as soon as you used it would go off-track and needed constant babysitting if you wanted to use it for agentic work.
re-thc · · focus · HN ↗
Via the API. The $200 OpenAI plan just got cut and most say general quotas got cut before that so for users on a plan the numbers might be different.
nl · · focus · HN ↗
qrify_app · · focus · HN ↗
[dead]
scrollop · · focus · HN ↗
<a href="https://www.bridgebench.ai/nerf-bench" rel="nofollow">https://www.bridgebench.ai/nerf-bench
<a href="https://marginlab.ai/trackers/claude-code/" rel="nofollow">https://marginlab.ai/trackers/claude-code/ <a href="https://marginlab.ai/trackers/codex/" rel="nofollow">https://marginlab.ai/trackers/codex/
<a href="https://github.com/ninjahawk/livenerf" rel="nofollow">https://github.com/ninjahawk/livenerf
<a href="https://isitnerfed.org/" rel="nofollow">https://isitnerfed.org/
ozgung · · focus · HN ↗
Methodology section in some of these benchmarks doesn’t say if they use subscription or API. API usage may not be nerfed as much as subscription.
They must be using the same lever to “pace the frontier”. All of the best effort models from different companies have similar scores. There is no standard definition of “max” effort level.
alvarolucero · · focus · HN ↗
[dead]
algoth1 · · focus · HN ↗
dang · · focus · HN ↗
Gemini 4 Argon - <a href="https://news.ycombinator.com/item?id=49913571">https://news.ycombinator.com/item?id=49913571
jwpapi · · focus · HN ↗
small_model · · focus · HN ↗
mlmonkey · · focus · HN ↗
<a href="https://imgur.com/a/h96yg5t" rel="nofollow">https://imgur.com/a/h96yg5t
This is on a $20/mo paid plan :cry:
radicality · · focus · HN ↗
brainwad · · focus · HN ↗
lhk931122 · · focus · HN ↗
epolanski · · focus · HN ↗
Companies out there are on google cloud or microsoft offerings and getting Gemini in their bundle, they aren't going through lawyers, etc, to provision from Anthropic or OpenAI just because they look a bit better on nerd benchmarks.
mchusma · · focus · HN ↗
After being nowhere near the frontier for a long time, they are pre announcing a model that ranks 3rd, roughly on par with models today that are cheaper.
Good for them to think about releasing to stay in the frontier game.
(I do think 3.7 flash was a solid release, so they are around the conversation. And their image and audio and live models are good)