‹ BackHN Continuity

Thread

M5 Ultra Mac Studio Review

269 points · 262 comments · piotrgrabowski

  1. simonw · · focus · HN ↗
    The numbers I was most interested in are tucked away in a chart towards the bottom - the speed comparison of the Mac Studios v.s. a RTX 5090:

      Qwen3.8 27B tokens/sec generation speed
    
      Prompt size    8K    64K   128K   256K
      RTX 5090 PC    59    51    44     n/a
      M5 Ultra       48    39    32     24
      M3 Ultra       31    23.5  20     15
    
    A whole bunch more comparison numbers in this section: <a href="https:&#x2F;&#x2F;www.macstories.net&#x2F;stories&#x2F;m5-ultra-mac-studio-review-the-dream-mac-for-local-ai-agents&#x2F;#mx-pc" rel="nofollow">https:&#x2F;&#x2F;www.macstories.net&#x2F;stories&#x2F;m5-ultra-mac-studio-revie...
    1. peri-cl · · focus · HN ↗
      Those are some incredible graphs, that leap in prompt processing going from M3 to M5.

      Also: ~30 token&#x2F;s on GLM 5.3-flash, locally. (That&#x27;s roughly Opus 4.8-tier. I think).

      &#x2F;meta Here&#x27;s a CSS filter that stops those nuisance chart animations,

          macstories.net##*:style(animation: none !important; transition: none !important)
Open on Hacker News to reply ↗

Unofficial Hacker News client; not affiliated with Y Combinator.