‹ BackHN Continuity

Thread

Mercury 2.5 LLM hits 770 tokens per second

151 points · 92 comments · Retro_Dev

  1. nylonstrung · · focus · HN ↗
    I honestly think the diffusion LLM approach is a dead end

    It's telling that frontier labs like Google toyed around with it but didn't invest further even for their most speed and cost sensitive small models

    Still unclear for what, if any use cases this is pareto frontier

    1. clhodapp · · focus · HN ↗
      Personally, I think it's more that text diffusion is not the ideal driver of an agentic work loop than that text diffusion is a total dead end. I am still hoping to see how it does on authoring and editing with further scaling and optimization. I think the push for AGI has put a bit too much focus on the idea of one general model doing everything.
Open on Hacker News to reply ↗

Unofficial Hacker News client; not affiliated with Y Combinator.