‹ BackHN Continuity

Thread

M5 Ultra Mac Studio Review

269 points · 262 comments · piotrgrabowski

  1. simonw · · focus · HN ↗
    The numbers I was most interested in are tucked away in a chart towards the bottom - the speed comparison of the Mac Studios v.s. a RTX 5090:

      Qwen3.8 27B tokens/sec generation speed
    
      Prompt size    8K    64K   128K   256K
      RTX 5090 PC    59    51    44     n/a
      M5 Ultra       48    39    32     24
      M3 Ultra       31    23.5  20     15
    
    A whole bunch more comparison numbers in this section: <a href="https:&#x2F;&#x2F;www.macstories.net&#x2F;stories&#x2F;m5-ultra-mac-studio-review-the-dream-mac-for-local-ai-agents&#x2F;#mx-pc" rel="nofollow">https:&#x2F;&#x2F;www.macstories.net&#x2F;stories&#x2F;m5-ultra-mac-studio-revie...
    1. gpugreg · · focus · HN ↗
      Those RTX 5090 numbers are bad. You can get over 200 tps with ninfer using NVFP4 and MTP.
      1. beastman82 · · focus · HN ↗
        can confirm.

        I dont&#x27; know why people spend huge money on these and Spark. The 5090 is running qwen 3.8 at 200+ tps!! That&#x27;s 1-2 orders of magnitude faster.

        1. mathisfun123 · · focus · HN ↗
          same reason they spend huge amounts of money on rolexes when seikos work better (the tech crowd isn&#x27;t immune from vanity).
          1. throwaway27448 · · focus · HN ↗
            If you seriously think apple products are nothing but a status item, you&#x27;re deluding yourself and probably have been for decades.
            1. _hugerobots_ · · focus · HN ↗
              This 1000%. Data centres don&#x27;t equate to medium sized labs and businesses. A stack of Macs is up and running without digging trenches, an electrician on staff and a department of PhDs to justify the spend.
              1. bigyabai · · focus · HN ↗
                It&#x27;s likely that a stack of Macs will draw more power for slower prefill&#x2F;decode than equivalently priced Nvidia GPUs. If power efficient inference is the goal, Macs are a non-starter.
                1. _hugerobots_ · · focus · HN ↗
                  So if it isn&#x27;t a comparative ability, now it&#x27;s a power cost issue? This reads like goal post moving.
                  1. bigyabai · · focus · HN ↗
                    Oh, it&#x27;s absolutely both. The power you waste waiting for TFTT on prefill will absolutely compound at the &quot;medium sized labs and businesses&quot; scale.
Open on Hacker News to reply ↗

Unofficial Hacker News client; not affiliated with Y Combinator.