‹ BackHN Continuity

Thread

How GLM built its own inference infrastructure

411 points · 285 comments · whiteros_e

  1. 9cb14c1ec0 · · focus · HN ↗
    Given the huge amount of money being spent on AI chips in the US, what prevents US AI labs from doing the same level of software optimization? It could be a solve for some of the capacity constraints.
    1. vblanco · · focus · HN ↗
      They have already been doing it for months <a href="https:&#x2F;&#x2F;openai.com&#x2F;index&#x2F;openai-broadcom-jalapeno-inference-chip&#x2F;" rel="nofollow">https:&#x2F;&#x2F;openai.com&#x2F;index&#x2F;openai-broadcom-jalapeno-inference-... . OpenAI on their custom chip brought up lightspeed deepseek as experiment by using AI in the exact same way as this zAI blogpost. And the kernel optimization contests&#x2F;etc have all been havily done through AI based optimization loops for half a year+.
Open on Hacker News to reply ↗

Unofficial Hacker News client; not affiliated with Y Combinator.