‹ BackHN Continuity

Thread

Samsung is expected to more than double output of its HBM4 and HBM4E DRAM

562 points · 458 comments · giuliomagnifico

  1. amelius · · focus · HN ↗
    Will that be enough for AI's hunger?
    1. GoToRO · · focus · HN ↗
      It will be just in time for when AI will run very well on consumer hardware and the need for data centers will collapse.
      1. IshKebab · · focus · HN ↗
        That's never going to happen. By the time you can run current frontier models on your $10k desktop the frontier will have massively advanced and people will want those models instead.
        1. hypfer · · focus · HN ↗
          I'm not sure if this prediction will hold true.

          We're not seeing the progress in those "frontier models" that we have previously seen. There's certainly still gas left in tank tank, but we're way into the diminishing returns by now.

          Cloud inference still beats hardware investments by orders of magnitude of course, but that's only if your data doesn't really matter to you.

          1. airspresso · · focus · HN ↗
            We are certainly not in the diminishing returns phase for LLM progress. No sign of that yet.
            1. bunderbunder · · focus · HN ↗
              I’ll grant that for specialized applications like coding agents and mathematics, but even there I suspect that most the real gains are actually taking place in the harness.

              But I suspect returns may have already diminished into negative territory for at least some other use cases. One of my least favorite job responsibilities in this brave new era is figuring out how to avoid performance and behavior regressions when an older model were using for some application reaches end of life. It’s getting uncommon for me to look at our benchmark results and say, “Oh, good, it does better on one of the newer models!”

              1. pixl97 · · focus · HN ↗
                >suspect that most the real gains are actually taking place in the harness.

                Part of the reason harnesses work well is you can run a lot of agents in parallel. That doesn't slow down demand.

                1. bunderbunder · · focus · HN ↗
                  I had actually been thinking more about all the non-LLM functionality that go into the harnesses. I'm not going to name names and I haven't done any rigorous testing, but my general impression is that choice of harness matters more than choice of model. In terms of basic task completion success specifically, not code aesthetics.
                  1. pixl97 · · focus · HN ↗
                    A perfect harness will not extract gold from a dumb model. It's a system that builds on each other, though we've not probed that frontier much to have a good intuition on what effects what.
Open on Hacker News to reply ↗

Unofficial Hacker News client; not affiliated with Y Combinator.