‹ BackHN Continuity

Thread

Clef: Open-weight decision models, and new RL fine-tuning platform

637 points · 217 comments · jasondavies

  1. cakoose · · focus · HN ↗
    > This means that a human does not necessarily need to be in the loop for agentic decisions anymore — agents can programmatically gather context, make decisions, and take actions on tasks, or defer to a human when needed.

    1. Humans are already not in the loop for lots of LLM agent actions. Isn't that just a function of how much you trust it and not some completely new paradigm? Am I missing something?

    2. How can it gather context if it just outputs a single decision?

    One guess: Maybe it's decision can be "gather more context and re-run me"? But an LLM can be much more expressive about what context it needs.

    1. lmc · · focus · HN ↗
      Yeah reading that made me question the credibility of the whole thing.
Open on Hacker News to reply ↗

Unofficial Hacker News client; not affiliated with Y Combinator.