‹ BackHN Continuity

Thread

Ember-1

589 points · 249 comments · gmays

  1. themgt · · focus · HN ↗
    The result? Ember-1 set a new Pareto frontier for Bedside Bench across both open and closed models including GPT-5.6 Sol, GPT-6 Astra, and Claude Opus 5 on cost/task.

    "Pareto": 8 hits

    "Opus 5.5": zero hits

    1. wmf · · focus · HN ↗
      Obviously this research was done before 6.0 Sol and Opus 5.5 came out. Your point stands that the frontier moves quickly and small gains can be eclipsed quickly.
Open on Hacker News to reply ↗

Unofficial Hacker News client; not affiliated with Y Combinator.