‹ BackHN Continuity

Thread

How GLM built its own inference infrastructure

411 points · 285 comments · whiteros_e

  1. bguberfain · · focus · HN ↗
    Plot twist: the GLM optimization agent figured out that it can hack and use NVIDIA GPUs on a US Cloud provider and make the inference 10x faster.
Open on Hacker News to reply ↗

Unofficial Hacker News client; not affiliated with Y Combinator.