Samsung is expected to more than double output of its HBM4 and HBM4E DRAM
Thread
Unofficial Hacker News client; not affiliated with Y Combinator.
Samsung is expected to more than double output of its HBM4 and HBM4E DRAM
Unofficial Hacker News client; not affiliated with Y Combinator.
amelius · · focus · HN ↗
GoToRO · · focus · HN ↗
IshKebab · · focus · HN ↗
hypfer · · focus · HN ↗
We're not seeing the progress in those "frontier models" that we have previously seen. There's certainly still gas left in tank tank, but we're way into the diminishing returns by now.
Cloud inference still beats hardware investments by orders of magnitude of course, but that's only if your data doesn't really matter to you.
airspresso · · focus · HN ↗
bunderbunder · · focus · HN ↗
But I suspect returns may have already diminished into negative territory for at least some other use cases. One of my least favorite job responsibilities in this brave new era is figuring out how to avoid performance and behavior regressions when an older model were using for some application reaches end of life. It’s getting uncommon for me to look at our benchmark results and say, “Oh, good, it does better on one of the newer models!”
3eb7988a1663 · · focus · HN ↗
Is it just that the providers are generating tons of synthetic datasets on coding tasks so that the models get more exposure to the right thing to do? Every time someone points out an LLM stupidity they add some training data to patch over the weakness (trivial to generate "there are two 'l's in llama")?