Clef: Open-weight decision models, and new RL fine-tuning platform
Thread
Unofficial Hacker News client; not affiliated with Y Combinator.
Clef: Open-weight decision models, and new RL fine-tuning platform
Unofficial Hacker News client; not affiliated with Y Combinator.
agrippanux · · focus · HN ↗
Clef was 2-3x slower and worse (it caught less hate speech) than Jev. Overall disappointing.
teleforce · · focus · HN ↗
Although they included the benchmark against DJev and Clef is better, perhaps if you can test it to see the real-world performance.
nostrebored · · focus · HN ↗
wongarsu · · focus · HN ↗
For small inputs Jev is slow, but its latency curve is very flat. Fine-tuning a decently-sized llm (like this 27B model) gives you something that's faster on small input sizes, but even with moderate contexts quickly becomes much slower than Jev. Characterizing it as "faster than Jev" is very misleading, unless you know all your questions have tiny context (less than 1k tokens or so)
tomrod · · focus · HN ↗
wongarsu · · focus · HN ↗