Clef: Open-weight decision models, and new RL fine-tuning platform
Thread
Unofficial Hacker News client; not affiliated with Y Combinator.
Clef: Open-weight decision models, and new RL fine-tuning platform
Unofficial Hacker News client; not affiliated with Y Combinator.
cakoose · · focus · HN ↗
1. Humans are already not in the loop for lots of LLM agent actions. Isn't that just a function of how much you trust it and not some completely new paradigm? Am I missing something?
2. How can it gather context if it just outputs a single decision?
One guess: Maybe it's decision can be "gather more context and re-run me"? But an LLM can be much more expressive about what context it needs.
lmc · · focus · HN ↗