‹ BackHN Continuity

Thread

Sonnet 5.5

884 points · 613 comments · D2OQZG8l5BI1S06

  1. wkcheng · · focus · HN ↗
    The cost / performance chart shows that in almost all configurations, it looks worse than Opus. Why would you use Sonnet 5.5 on xhigh if you would get better results (higher score, cheaper cost) on Opus 5.5 high?

    Is there a good use case? This isn't like Luna where it's much cheaper/effective just to use Luna in certain situations.

    1. usaar333 · · focus · HN ↗
      Per the charts, there is largely no point to using Sonnet 5.5 at high+ as opus low generally will give similar performance at similar or lower cost.

      But Sonnet 5.5 at medium and below gives you a cheaper option at a performance worse than the lowest thinking Opus (low), which may be viable for "low intelligence" use cases.

      1. verdverm · · focus · HN ↗
        at that point, you can switch to dirt cheap open models
        1. NotSuspicious · · focus · HN ↗
          Only if your company (and government) lets you!
          1. verdverm · · focus · HN ↗
            yup, we use Fireworks.ai, and American company with ZDR

            Their new Ember-1 model is pretty good, fine-tune of Kimi3 with way less thinking

            Does this make it an American model or is it still Chinese?

Open on Hacker News to reply ↗

Unofficial Hacker News client; not affiliated with Y Combinator.