7-8 figures annual spend will buy a hell of a lot of capable local inference hardware you can own, though it won't be at the absurd token/s rate, you'll be able to run almost anything on it... And it'll still have a good residual resale value after 4 years the way things are going now.
Feel like you could spend 6 figures building out a team and the rest renting compute for a whole year, and get the team to create a local inference solution with that kind of budget…
bearjaws · · focus · HN ↗
I've used it on a few for fun projects and its decent but the speed is crazy to watch.
ford · · focus · HN ↗
[0] <a href="https://www.cerebras.ai/blog/cerebras-kimi-k2-Enterprise" rel="nofollow">https://www.cerebras.ai/blog/cerebras-kimi-k2-Enterprise
walrus01 · · focus · HN ↗
podocarp · · focus · HN ↗