Clef: Open-weight decision models, and new RL fine-tuning platform
Thread
Unofficial Hacker News client; not affiliated with Y Combinator.
Clef: Open-weight decision models, and new RL fine-tuning platform
Unofficial Hacker News client; not affiliated with Y Combinator.
manlymuppet · · focus · HN ↗
And it's only been a few weeks.
TeMPOraL · · focus · HN ↗
There are many, many of those left around, because AI frontier is moving forward so fast, everyone is racing ahead. Which is why I laugh when people say AI is not transformative and LLMs are a dead end (and my favorite, "what are we going to do with all those GPUs when the bubble pops?"). Even if SOTA LLMs hit a hard capability limit tomorrow and never advanced again, there's a good decade of growth and advancement to be extracted just from all the low-hanging fruits that were left unpicked along the way.
tomrod · · focus · HN ↗
Your comment here made me laugh, because I think we will finally be able to play Crysis at 10fps.
Just kidding, of course. GPU half lives are quite a bit less than standard compute half lives, no?[0] That's what I've been trying to understand regarding data centers focusing as GPU clusters -- seems like the ROI window would have to be very short for the capitalization.
[0] <a href="https://www.tomshardware.com/pc-components/gpus/datacenter-gpu-service-life-can-be-surprisingly-short-only-one-to-three-years-is-expected-according-to-unnamed-google-architect" rel="nofollow">https://www.tomshardware.com/pc-components/gpus/datacenter-g...