‹ BackHN Continuity

Thread

Beating GPT-5.6 Sol on retrieval with 100x cheaper open models

414 points · 114 comments · moonikakiss

  1. andai · · focus · HN ↗
    Nice, but there's no mention of how Luna or DSFlash perform on the same task? (Being 25x and 50x cheaper respectively.)

    Nor of how much faster their custom model performs?

    1. [deleted] · · focus · HN ↗

      [deleted]

    2. seahyinghang8 · · focus · HN ↗
      we actually have the test benchmark against luna but no deepseek flash (we haven't added DSFlash into our benchmarking model pipeline)

      you can check out the full comparison against all the other models here: <a href="https:&#x2F;&#x2F;app.castform.com&#x2F;train&#x2F;a7a898f6-d802-4908-b044-acb812f14a48?tab=comp" rel="nofollow">https:&#x2F;&#x2F;app.castform.com&#x2F;train&#x2F;a7a898f6-d802-4908-b044-acb81...

      - founder of castform

Open on Hacker News to reply ↗

Unofficial Hacker News client; not affiliated with Y Combinator.