‹ BackHN Continuity

Thread

Clef: Open-weight decision models, and new RL fine-tuning platform

637 points · 217 comments · jasondavies

  1. bityard · · focus · HN ↗
    Clef is based on Qwen3.8-27B and Clef-flash is based on Qwen3.8-9B (edit: actually Qwen3.5-9B). So, similar in spirit to Kev by my understanding, but based on a newer model.
    1. okpatil · · focus · HN ↗
      Atom is 60M Param (around 133x to 400x smaller).

      16ms latency. And locally run.

      <a href="https:&#x2F;&#x2F;at0m.pienomial.com&#x2F;" rel="nofollow">https:&#x2F;&#x2F;at0m.pienomial.com&#x2F;

      Why go big when you can go small ?

      1. ricardobeat · · focus · HN ↗
        Because with a tiny model you&#x27;re skipping all the intelligence and world knowledge that makes it useful without fine tuning. `typed-decisions` is almost entirely text classification tasks.
        1. okpatil · · focus · HN ↗
          Don&#x27;t assume. It has generic world knowledge. It is performing well on the benchmarks we didn&#x27;t even train it on.
Open on Hacker News to reply ↗

Unofficial Hacker News client; not affiliated with Y Combinator.