‹ BackHN Continuity

Thread

How GLM built its own inference infrastructure

411 points · 285 comments · whiteros_e

  1. throwa356262 · · focus · HN ↗

        "We implemented a series of aggressive memory optimizations, including..."
    
    
    This whole thing sounds like industrial scale auto-research, but done by people who actually know what they are doing.
Open on Hacker News to reply ↗

Unofficial Hacker News client; not affiliated with Y Combinator.