‹ BackHN Continuity

Thread

How OpenAI Used Its Own LLMs to Design Its Jalapeño Chip

206 points · 139 comments · maxall4

  1. globnomulous · · focus · HN ↗
    > Jalapeño can reduce end-to-end latency (the time between prompt to last token) by up to 3.6 times

    I'm never sure what on earth this kind of impressionistic math is supposed to tell me. Is the comparison between 4.6 and 1.0? 3.6 and 1.0? Clearly the comparison isn't supposed to be 1.0 and -2.6, even though that's what the words literally mean. I can't be the only person who finds this infuriating and distracting. These numbers shouldn't be impressionistic. They should be precise. That this is an article on spectrum.ieee.org makes the imprecision all the stranger. I'd expect their readershipt to care, for instance, about what's even being measured. Is this the geometric mean of something? The arithmetic mean? And what latency has improved?

    1. caidan · · focus · HN ↗
      That odor you are detecting is just good old fashioned bullshit, my friend. It’s just that nowadays everything and everyone is covered in it, and we are not supposed to notice. The emperor has no clothes… and is covered in shit.
Open on Hacker News to reply ↗

Unofficial Hacker News client; not affiliated with Y Combinator.