‹ BackHN Continuity

Thread

Show HN: Maple-Preview – Ternary 20B MoE running at 120 tok/s on a iPhone

169 points · 52 comments · edwardbzhang

  1. beautiful_apple · · focus · HN ↗
    A benchmark table comparing to Qwen 3.5 35B-A3B seems strange when Qwen 3.6 35B-A3B has been out for some time and is significantly better.

    I didn't notice the version difference when first reading the article! So this is a heads up to people like me.

    1. ricardobeat · · focus · HN ↗
      Their main comparison is 1-bit Bonsai 27B (Qwen3.6 27B) which beats A3B anyway.
      1. beautiful_apple · · focus · HN ↗
        I'm not sure what you mean.

        Looking at the chart on this website, Bonsai Qwen 3.6 27B has a lower average benchmark score than Qwen 3.5 35B-A3B (77.1 vs 82.9)

        1. ricardobeat · · focus · HN ↗
          The 27B model beats 35B A3B. The quantized 1-bit version obviously doesn't, but it would likely beat a 1-bit quantized A3B variant. Comparing 3.5 35B or 3.6 35B to this model (Maple) is not meaningful.
Open on Hacker News to reply ↗

Unofficial Hacker News client; not affiliated with Y Combinator.