‹ BackHN Continuity

Thread

Apple M6 Pro achieves the highest single-core CPU score in Geekbench 7

118 points · 145 comments · gainsurier

  1. GeekyBear · · focus · HN ↗
    There are also leaks for the M5 Ultra & Base M6.

    M5 Ultra CPU:

    <a href="https:&#x2F;&#x2F;browser.geekbench.com&#x2F;search?k=parkdale_cpu&amp;q=mac17%2C15" rel="nofollow">https:&#x2F;&#x2F;browser.geekbench.com&#x2F;search?k=parkdale_cpu&amp;q=mac17%...

    M5 Ultra GPU:

    <a href="https:&#x2F;&#x2F;browser.geekbench.com&#x2F;search?k=grand_gpu&amp;q=mac17%2C15" rel="nofollow">https:&#x2F;&#x2F;browser.geekbench.com&#x2F;search?k=grand_gpu&amp;q=mac17%2C1...

    Base M6 CPU:

    <a href="https:&#x2F;&#x2F;browser.geekbench.com&#x2F;search?k=parkdale_cpu&amp;q=mac18%2C5" rel="nofollow">https:&#x2F;&#x2F;browser.geekbench.com&#x2F;search?k=parkdale_cpu&amp;q=mac18%...

    Base M6 GPU:

    <a href="https:&#x2F;&#x2F;browser.geekbench.com&#x2F;search?k=grand_gpu&amp;q=mac18%2C5" rel="nofollow">https:&#x2F;&#x2F;browser.geekbench.com&#x2F;search?k=grand_gpu&amp;q=mac18%2C5

    The Base M6 CPU single core is averaging a bit over 4000 on Geekbench 7.

    For comparison, the AMD Ryzen 9 9950X3D2 averages 3161 on the same test.

    <a href="https:&#x2F;&#x2F;browser.geekbench.com&#x2F;processors&#x2F;amd-ryzen-9-9950x3d2" rel="nofollow">https:&#x2F;&#x2F;browser.geekbench.com&#x2F;processors&#x2F;amd-ryzen-9-9950x3d...

    The Intel Core Ultra 7 270K Plus averages 2940 on the same test.

    <a href="https:&#x2F;&#x2F;browser.geekbench.com&#x2F;processors&#x2F;intel-core-ultra-7-270k-plus" rel="nofollow">https:&#x2F;&#x2F;browser.geekbench.com&#x2F;processors&#x2F;intel-core-ultra-7-...

    1. revolvingthrow · · focus · HN ↗
      I don’t understand those benchmarks.

      According to geekbench 7800x3d is 2400 single core, 15500 multicore. M4 pro is 3350 single core and 24750 multicore. Yet when I convert video using libsvtav1 with ffmpeg I’m getting noticeably faster performance on the desktop. And that’s with mbp, which doesn’t thermally throttle within 15 seconds.

      Is it the 96mb cache? Avx-512? Are benchmarks bullshit when comparing different architectures?

      1. kakacik · · focus · HN ↗
        &gt; Are benchmarks bullshit when comparing different architectures?

        Bingo, each CPU is too unique with its own strengths and weaknesses to make broad statements, marketing picks up what they like and ignore rest

      2. galad87 · · focus · HN ↗
        Which version of svt-av1 are you using? Neon optimisations were added quite recently in svt-av1, and are still a bit incomplete, so it might make a difference.
      3. snek_case · · focus · HN ↗
        Probably AVX-512. AFAIK Apple CPUs are still using ARM NEON instructions with 128-wide SIMD registers. I guess they are banking on you using the GPU if you want to parallelize those kinds of workload.
      4. GeekyBear · · focus · HN ↗
        Remember all those years when people claimed Apple chips were &quot;cheating&quot; on benchmarks because they had crypto instructions that x86 lacked?

        Apple chips don&#x27;t have AVX-512.

        They do have media engines (that don&#x27;t support AV1 encode), so if you switch to H.265, it will pull way ahead.

        1. ajross · · focus · HN ↗
          &gt; Apple chips don&#x27;t have AVX-512.

          Zen 4 isn&#x27;t 512 bits wide though, it&#x27;s a split cycle 256 bit wide SIMD engine otherwise very similar (except in register size) to M6&#x27;s.

          The answer is more that Geekbench is at this point[1] heavily tuned to exactly the code Apple silicon does well: implicitly parallel wide-issue scalar code that you typically get out of modern compilers and JIT engines when throwing mostly-unoptimized &quot;regular source code&quot; at them. Apple has an enormous amount of instruction issue parallelism compared with x86.

          The grandparent is looking at transcoding tasks where the limit isn&#x27;t instruction issue but actual compute hardware on the core. And Apple doesn&#x27;t actually win by much there.

          It&#x27;s just hard to know what to measure. Geekbench tends to be a metric for &quot;feels fast doing boring interactive user stuff&quot;, which probably matches Apple&#x27;s marketing imperatives well.

          [1] Really they keep moving harder in that direction with every release. The &quot;cooling pauses&quot; in v6 likewise seemed very much like an attempt to boost the score on fanless Apple devices. If one were the type to allege a dark conspiracy, this is a tempting spot.

          1. thejazzman · · focus · HN ↗
            I would expect cooling pauses to have the opposite effect. The

            MacBook Air&#x2F;neo are the only ones without fans. Even the Studio Display has a fan

            1. ajross · · focus · HN ↗
              The cooling pause thing was just a dig at Geekbench, not specific to this discussion (thus putting it in a footnote). They tuned (the last version of their) benchmark specifically to give better numbers on hardware that tends to throttle under performance load.

              Now... you can make a reasonable case that this matches real world interactive load better. But it also happens to have the practical effect of making Apple&#x27;s numbers better, and no one else&#x27;s. The new benchmarks are measuring subtly different things, and the new thing they measure happens to be what Apple wants to sell. At best, that&#x27;s backwards.

          2. krunkcoin · · focus · HN ↗
            Cooling pauses have been a part of Geekbench forever, they&#x27;re not just a GB6 thing.

            GB&#x27;s author, John Poole, has stated the intent of Geekbench CPU is to measure the CPU and CPU alone, as in he doesn&#x27;t want to measure limits caused by the form factor of the device the CPU under test is embedded in. Since GB runs on phones, that means he had to implement the cooldown pauses.

            Intel&#x27;s turbo boost means this GB feature almost certainly inflates lots of x86 scores too, even with fans involved. Not sure why you&#x27;re being so conspiratorial about it.

        2. [deleted] · · focus · HN ↗

          [deleted]

      5. luxuryballs · · focus · HN ↗
        sounds like it&#x27;s the software
      6. BoingBoomTschak · · focus · HN ↗
        I think software with a strong focus on SIMD really favours x86 over Apple, since AVX2 support was pretty uniform and AVX-512 (except Intel&#x27;s fiasco) has been supported for a while, while the post-NEON landscape is strange in ARM country.

        From what I&#x27;ve been able to gather, Apple only supports SME and the small &quot;streaming&quot; part of SVE2 required by SME since the M4. Which seems to be mostly useless for video encoding (I only see NEON&#x2F;SVE in <a href="https:&#x2F;&#x2F;gitlab.com&#x2F;AOMediaCodec&#x2F;SVT-AV1&#x2F;-&#x2F;tree&#x2F;master&#x2F;Source&#x2F;Lib" rel="nofollow">https:&#x2F;&#x2F;gitlab.com&#x2F;AOMediaCodec&#x2F;SVT-AV1&#x2F;-&#x2F;tree&#x2F;master&#x2F;Source... or <a href="https:&#x2F;&#x2F;github.com&#x2F;Multicorewareinc&#x2F;x265&#x2F;tree&#x2F;master&#x2F;source&#x2F;common" rel="nofollow">https:&#x2F;&#x2F;github.com&#x2F;Multicorewareinc&#x2F;x265&#x2F;tree&#x2F;master&#x2F;source&#x2F;...) it&#x27;s basically a GEMM engine.

        So basically, video encoding is a bad show for Apple who seem to be saying &quot;use crappy hardware encoding and buy amd64 if you need more&quot;.

        Very hard to find benchmark data. Found <a href="https:&#x2F;&#x2F;openbenchmarking.org&#x2F;vs&#x2F;Processor&#x2F;Apple+M4+Pro,AMD+Ryzen+9+9900X+12-Core" rel="nofollow">https:&#x2F;&#x2F;openbenchmarking.org&#x2F;vs&#x2F;Processor&#x2F;Apple+M4+Pro,AMD+R... that shows a good x265 performance at 1080p but not 4K. Guess the wider SIMD units matter more there.

        1. GeekyBear · · focus · HN ↗
          &gt; So basically, video encoding is bad show for Apple

          Apple has dedicated hardware for video decode and encode for several video formats.

          That&#x27;s why the comparisons between PCs and Macs running video editing software favor the Macs so heavily.

          1. BoingBoomTschak · · focus · HN ↗
            HW and SW encoding are not comparable with regard to quality (for the same size, obviously) though. That&#x27;s kinda what I meant with my quip.
            1. GeekyBear · · focus · HN ↗
              Nothing prevents a vendor from implementing a high quality hardware encoder, just as nothing prevents the creation of a low quality software encoder.

              Video editing is an area that Apple dominates because PCs become a laggy mess on complicated high resolution edits.

              AMD has started to copy the same strategy with the recent Ryzen AI chips. They include hardware video encode&#x2F;decode units as well as unified memory.

              1. BoingBoomTschak · · focus · HN ↗
                &gt; Nothing prevents a vendor from implementing a high quality hardware encoder

                Uh yes, it&#x27;s called reality. Advanced video coding techniques are just too complex, full of serial and branch heavy algos to work in decently sized and priced ASIC or even in GPGPU. Which is the reason none exists outside of very expensive special stuff for professional render&#x2F;streaming farms.

                Not even mentioning that few coding tools are shared between codecs, so while you only need 1 CPU to handle all of them, you&#x27;d need almost an ASIC per codec...

                1. GeekyBear · · focus · HN ↗
                  If your claim is true, you won&#x27;t have any trouble finding reputable tech publications saying that the video output from Apple M series hardware is substandard in some way.

                  I&#x27;ll wait.

                  I would recommend you find a floor plan for Apple&#x27;s M series chips and take a look at how big Apple&#x27;s media engine is.

                  1. BoingBoomTschak · · focus · HN ↗
                    I don&#x27;t know how to say nicely that you don&#x27;t realize how ignorant about the field you must seem to anyone who isn&#x27;t. You&#x27;ll have trouble finding such comparison for the same reason you won&#x27;t find a &quot;caterpillar vs motorbike&quot; speed test easily. Everyone in the target market including Apple knows hardware encoding will never be about output quality.

                    Anyway, here, I&#x27;m feeling nice. Some guy comparing the M1 Pro&#x27;s HEVC encoder vs x265 (amongst others): <a href="https:&#x2F;&#x2F;colinmckellar.com&#x2F;2024&#x2F;01&#x2F;11&#x2F;video-encoder-comparison&#x2F;" rel="nofollow">https:&#x2F;&#x2F;colinmckellar.com&#x2F;2024&#x2F;01&#x2F;11&#x2F;video-encoder-compariso...

                    Here&#x27;s the only graph you need to look at if you don&#x27;t want to bother (encoding time vs file size at fixed perceptual quality, log scale axes): <a href="https:&#x2F;&#x2F;colinmckellar.com&#x2F;wp-content&#x2F;uploads&#x2F;2024&#x2F;01&#x2F;VMAF_90.png" rel="nofollow">https:&#x2F;&#x2F;colinmckellar.com&#x2F;wp-content&#x2F;uploads&#x2F;2024&#x2F;01&#x2F;VMAF_90...

                    1. GeekyBear · · focus · HN ↗
                      &quot;Some Guy&quot; just doesn&#x27;t cut it.

                      We&#x27;ve got five generations of Apple&#x27;s Media Engine shipping in M series chips.

                      If things are as dire as you claim, one of the many hardware reviews in reputable publications over the years would have mentioned this unacceptable quality at some point.

                      1. timschmidt · · focus · HN ↗
                        I&#x27;m not BoingBoomTschak who you&#x27;re replying to, but he&#x27;s right. My experience comes mostly from decades of reading xiph.org and ffmpeg mailing lists where such things are discussed, rather than implementing them myself. But there is constant discussion of encoder performance &#x2F; quality trade-offs in software and hardware encoders. Hardware encoders, especially ones attempting to meet strict performance targets, simply cannot take advantage of some of the most complex quality improvements as they depend on information only present in past&#x2F;future frames. Sometimes as many as 30 frames away.

                        It seems like you are interpreting this as a slight against the quality of Apple&#x27;s hardware encoders, which may legitimately be very good. As are Nvidia&#x27;s, Intel&#x27;s and AMD&#x27;s. But all of them will produce larger file sizes and lower quality than equivalently optimized non-realtime software encoders, which simply have more information and more time, memory, and flexibility to compute over it.

                        We&#x27;re talking about fundamental properties of compression and computational time&#x2F;space trade-offs. Even Apple can&#x27;t design around them.

                        That doesn&#x27;t mean Apple&#x27;s hardware encoder is in any way bad or unusable. All lossy compression will be imperfect, yet much of it is useful. And most modern codecs and encoders seem to be capable of high quality results. The implications of the differences under discussion are percentages of a bitrate or tiny nearly imperceptible artifacts or breadth of available resolutions, refresh rates, and color modes or codec choice. Software encoders are always at the bleeding edge of what&#x27;s possible. Hardware encoders are necessarily a snapshot frozen in silicon with limitations imposed by the implementation. The middle ground is largely already occupied by SIMD and other transform-specific ISA extensions already present in most CPUs.

        2. api · · focus · HN ↗
          X64 does have a stronger SIMD story, but this is partly compensated for by the fact that ARM64 is much easier to decode wide. M series has a very wide decoder and a lot of instruction level parallelism. It can turn chunks of those 128-bit SIMD operations into what amounts to wider operations.

          AVX still wins though.

          M series still wins on performance per watt and now apparently leads on general purpose code.

          All these leading edge chips are very good. We have an embarrassment of riches when it comes to blistering fast chips here.

          1. Archit3ch · · focus · HN ↗
            &gt; AVX still wins though.

            On what, microbenchmarks? I have realtime audio workloads where the hot loop is essentially linear algebra. Exactly the kind of work that suits AVX2&#x2F;AVX512. Guess what, Apple Silicon still pulls ahead because real workloads are branchy, cache-hungry, full of dependencies and do not line up in 8 neat f64 operations per cycle.

      7. tom_ · · focus · HN ↗
        Judging by my PCs (M4 Max Mac Studio; 2990WX desktop PC), Geekbench might flatter the Apple chips for multicore a bit I think:

        Geekbench 7 results:

        * AMD 2990WX: 1384 (single), 13052 (multi) (<a href="https:&#x2F;&#x2F;browser.geekbench.com&#x2F;v7&#x2F;cpu&#x2F;181239" rel="nofollow">https:&#x2F;&#x2F;browser.geekbench.com&#x2F;v7&#x2F;cpu&#x2F;181239)

        * Apple M4 Max: 3552 (single), 29863 (multi) (<a href="https:&#x2F;&#x2F;browser.geekbench.com&#x2F;v7&#x2F;cpu&#x2F;390256" rel="nofollow">https:&#x2F;&#x2F;browser.geekbench.com&#x2F;v7&#x2F;cpu&#x2F;390256)

        For parallelisable stuff that can occupy all cores for an extended period, the 2990WX typically takes about ~1.2x as long to do the same work&#x2F;does ~0.83x the work per unit time, assuming code compiled with clang or gcc. Which isn&#x27;t really coming across in the numbers here.

        CPUMark is a bit better:

        * AMD 2990WX: 2282 (single), 32040 (multi) (<a href="https:&#x2F;&#x2F;www.cpubenchmark.net&#x2F;cpu.php?cpu=AMD+Ryzen+Threadripper+2990WX&amp;id=3309" rel="nofollow">https:&#x2F;&#x2F;www.cpubenchmark.net&#x2F;cpu.php?cpu=AMD+Ryzen+Threadrip...)

        * Apple M4 Max: 4590 (single), 43911 (multi) (<a href="https:&#x2F;&#x2F;www.cpubenchmark.net&#x2F;cpu.php?cpu=Apple+M4+Max+16+Core&amp;id=6348" rel="nofollow">https:&#x2F;&#x2F;www.cpubenchmark.net&#x2F;cpu.php?cpu=Apple+M4+Max+16+Cor...)

        I haven&#x27;t spent much time timing single core stuff, except - regarding clang, which looks like it contributes to the Geekbench 7 score, I did some measurements a few months ago suggesting that clang compiles for x64 more slowly than for ARM, all else being as equal as I could be bothered to try to make it: <a href="https:&#x2F;&#x2F;news.ycombinator.com&#x2F;item?id=46938682">https:&#x2F;&#x2F;news.ycombinator.com&#x2F;item?id=46938682 - and the single threaded test runs I did of my code suggest that the Geekbench 7 single core might be about right?

        (Whether the clang timing discrepancy is actually relevant to Geekbench, I&#x27;ve no idea, but I thought it interesting anyway.)

        If you need a benchmark that makes the PC look massively faster than the Mac, I&#x27;m sure those are available too.

        1. bob1029 · · focus · HN ↗
          &gt; If you need a benchmark that makes the PC look massively faster than the Mac, I&#x27;m sure those are available too.

          Go check out blender CPU scores if you need any reassurance that AMD still makes some kind of sense.

          <a href="https:&#x2F;&#x2F;opendata.blender.org&#x2F;benchmarks&#x2F;query&#x2F;?compute_type=CPU&amp;blender_version=3.1.0&amp;group_by=device_name" rel="nofollow">https:&#x2F;&#x2F;opendata.blender.org&#x2F;benchmarks&#x2F;query&#x2F;?compute_type=...

        2. jltsiren · · focus · HN ↗
          Geekbench multi-core has been a single-task benchmark since version 6. The multi-core &#x2F; single-core score ratio is supposed to tell how much you will benefit on the average when you use all available CPU cores. Some tasks parallelize better, while others have bottlenecks that prevent effective parallelization. For every CPU, the ratio is well below the nominal speedup you can get under ideal circumstances.
          1. wtallis · · focus · HN ↗
            Note that Geekbench 7 no longer runs the same subtests in multi-core mode as in single-core mode, so any ratio calculated from the overall scores instead of from the individual subtests is misleading. There are a lot of subtests that it only runs in single-core mode and omits from the multi-core tests, seemingly to placate the critics that didn&#x27;t like the inclusion of poorly-scaling tests in the multi-core mode. It&#x27;s not as dumb as the Geekbench 5 strategy of just running N independent copies of the test, but it does seem like a dumb change to me.
      8. ksec · · focus · HN ↗
        Just want to add apart from all the other comments. SVT was brought by Intel and spent years to make it extremely well tuned for x86. While not the same for ARM and not even for Apple.

        Another point is that the multicore part uses all core including E-Core. On AMD the multicore are all the same. Meaning for some benchmarks this will flavour Apple more.

        Again there is nothing that stop people from optimising it for ARM Mac. The problem is the usage of it is so small it probably doesn&#x27;t make sense to focus on it. SVT took a really long time for it to reach quality parity with AOM&#x27;s AV1 encoder and later exceed it.

      9. TiredOfLife · · focus · HN ↗
        Fast video encode was how AMD fans were coping during bulldozer era
    2. esperent · · focus · HN ↗
      There&#x27;s also the Snapdragon X2 Elite Extreme, 3,438 points in single-core and 27,519 points in multi-core, which I think was the record when it was released a year ago.

      I guess they&#x27;ll release an X3 in six months. So it&#x27;s interesting to see the race is now between Apple and Qualcomm, both on Arm, with Intel and AMD getting left behind, at least on this benchmark.

      <a href="https:&#x2F;&#x2F;signal65.com&#x2F;wp-content&#x2F;uploads&#x2F;2026&#x2F;08&#x2F;Signal65-Insights_Qualcomm-Snapdragon-X2-Elite-Extreme-Performance-1.pdf" rel="nofollow">https:&#x2F;&#x2F;signal65.com&#x2F;wp-content&#x2F;uploads&#x2F;2026&#x2F;08&#x2F;Signal65-Ins...

      1. GeekyBear · · focus · HN ↗
        Samsung&#x27;s chips use high end ARM stock cores, and are also ahead of Intel and AMD.
        1. TiredOfLife · · focus · HN ↗
          Which ones?
    3. ksec · · focus · HN ↗
      I know these are all Mac CPU, but for context people may want to look at A20 Pro [1], which in 2 years time will be in MacBook Neo. Pretty damn impressive if you ask me.

      A20 Pro is around 4K on Geekbench 7. And the P-Core is not the same as M6. The A20 Pro is actually only 9 wide compared to 10-wide on M6. Supposedly more power efficient and smaller die size.

      [1] <a href="https:&#x2F;&#x2F;browser.geekbench.com&#x2F;search?q=iPhone19%2C2" rel="nofollow">https:&#x2F;&#x2F;browser.geekbench.com&#x2F;search?q=iPhone19%2C2

      1. [deleted] · · focus · HN ↗

        [deleted]

Open on Hacker News to reply ↗

Unofficial Hacker News client; not affiliated with Y Combinator.