‹ BackHN Continuity

Thread

Clef: Open-weight decision models, and new RL fine-tuning platform

637 points · 217 comments · jasondavies

  1. AmazingTurtle · · focus · HN ↗
    quick reminder: this is a one shot pydantic token guess. if you try to use a model that was trained on reasoning? you're not getting any of the reasoning.
Open on Hacker News to reply ↗

Unofficial Hacker News client; not affiliated with Y Combinator.