‹ BackHN Continuity

Thread

Clef: Open-weight decision models, and new RL fine-tuning platform

637 points · 217 comments · jasondavies

  1. manlymuppet · · focus · HN ↗
    Am I hearing this right, that they made a decision model based on Typesafe's new paradigm, and actually made a model better than Jev based on Typesafe's own ranking?

    And it's only been a few weeks.

    1. heliosAtwork · · focus · HN ↗
      There was some parallel independent work from Sep 28. They mention it in the blog post. But I am sure Jev has opened a lot of eyes on the possibilities.

      "In the same week that Jev came out, we posted about some experiments [1] we had with our own homegrown decision model."

      [1] <a href="https:&#x2F;&#x2F;x.com&#x2F;michellechen&#x2F;status&#x2F;2101091012559151480" rel="nofollow">https:&#x2F;&#x2F;x.com&#x2F;michellechen&#x2F;status&#x2F;2101091012559151480

Open on Hacker News to reply ↗

Unofficial Hacker News client; not affiliated with Y Combinator.