‹ BackHN Continuity

Thread

Clef: Open-weight decision models, and new RL fine-tuning platform

637 points · 217 comments · jasondavies

  1. bityard · · focus · HN ↗
    Clef is based on Qwen3.8-27B and Clef-flash is based on Qwen3.8-9B (edit: actually Qwen3.5-9B). So, similar in spirit to Kev by my understanding, but based on a newer model.
    1. NitpickLawyer · · focus · HN ↗
      > and Clef-flash is based on Qwen3.8-9B

      There is no official qwen 3.8 9b

      From the model card:

      > Clef-Flash is post-trained from Qwen/Qwen3.5-9B. See Clef for the larger variant.

      1. ddarolfi · · focus · HN ↗
        It's based on Qwen3.5-9B, maybe a typo
        1. verdverm · · focus · HN ↗
          I've been guilty of the same wishful projection
      2. bityard · · focus · HN ↗
        Thanks, I missed that. Fixed my comment.
    2. okpatil · · focus · HN ↗
      Atom is 60M Param (around 133x to 400x smaller).

      16ms latency. And locally run.

      <a href="https:&#x2F;&#x2F;at0m.pienomial.com&#x2F;" rel="nofollow">https:&#x2F;&#x2F;at0m.pienomial.com&#x2F;

      Why go big when you can go small ?

      1. kamranjon · · focus · HN ↗
        Cause it&#x27;s not open?
        1. okpatil · · focus · HN ↗
          Good point.

          To counter, most of the AI is not open. So is none of Microsoft Products. As long as they work, we keep using them.

          1. verdverm · · focus · HN ↗
            counter point, I&#x27;ve stopped using all closed models and harnesses as a life choice

            Ai is too important and transformational to let Big Ai dominate in a closed ecosystem, thankfully the Chinese have a different mindset and approach

            1. okpatil · · focus · HN ↗
              That&#x27;s an absolutely fair way. More power to you !
      2. ricardobeat · · focus · HN ↗
        Because with a tiny model you&#x27;re skipping all the intelligence and world knowledge that makes it useful without fine tuning. `typed-decisions` is almost entirely text classification tasks.
        1. okpatil · · focus · HN ↗
          Don&#x27;t assume. It has generic world knowledge. It is performing well on the benchmarks we didn&#x27;t even train it on.
Open on Hacker News to reply ↗

Unofficial Hacker News client; not affiliated with Y Combinator.