It's also the only model that generates accurate translation and localization. No other frontier model comes close. Although Gemini's coding capabilities are subpar, its natural language processing is top-tier.
I’m curious how you guys keep track of each model’s coding capabilities. The landscape keeps changing. I don’t suppose you benchmark all frontier models every other month, right?
I've used Antigravity as my main coding agent on one of my biggest projects for about a year. It's been great for me. (and I use Claude, Codex, Grok and Muse for all the other projects)
Zsfe510asG · · focus · HN ↗
phenomen · · focus · HN ↗
thisgoodlife · · focus · HN ↗
riddlemethat · · focus · HN ↗
safog · · focus · HN ↗
Most of the time I don't need what the bench tests and I'm not really giving them completely ambiguous tasks without any refinement.
I only find marginal differences between models at this point and it almost feels like personality quirks in each model than anything.
BenzeneDream · · focus · HN ↗
qingcharles · · focus · HN ↗