‹ BackHN Continuity

Thread

Ember-1

589 points · 249 comments · gmays

  1. netvarun · · focus · HN ↗
    Off topic:With sol pricing drop tbh kimi k3’s value prop has not been that great. For our internal use case/testing/benchmarks sol come out with way better quality and much cheaper costs. Kimi really needs to drop their pricing (I heard it’s set by them across all the neoclouds) Sol is at 2/10 vs kimi’s 3/15
    1. 7777777phil · · focus · HN ↗
      I was surprised by that. I run my benchmark [1] every couple of days and was sure this model will be ath the pareto frontier, if not THE pareto frontier. But no:

      Ember isn't picked yet. In planning, Opus 5.5 wins under the planning weights. In code, GPT-6 Sol dominates it: also 10/10, but with a higher quality score and a lower estimated cost. Ember has no intelligence index, so its starting score is only 0.73, which holds its 10/10 down to 0.954 against Sol's 0.975.

      [1] <a href="https:&#x2F;&#x2F;philippdubach.com&#x2F;posts&#x2F;jev-model-router-for-pi&#x2F;" rel="nofollow">https:&#x2F;&#x2F;philippdubach.com&#x2F;posts&#x2F;jev-model-router-for-pi&#x2F;

Open on Hacker News to reply ↗

Unofficial Hacker News client; not affiliated with Y Combinator.