‹ BackHN Continuity

Thread

OpenAI is well positioned to fast-follow Jev

328 points · 233 comments · JohnBerryman

  1. tolugenius · · focus · HN ↗
    I'm not exactly following through with the claim, can someone explain how the built-in classification would not necessitate more tokens used, or be much different from turning on reasoning? Not that I don't see the difference, I just doing see how OpenAI would do it well.
    1. deepsquirrelnet · · focus · HN ↗
      It's hard to say without knowing their architecture, but I'd guess something like block attention. You can process the prompt separately from the classifications into a latent space and then do some kind of late interaction with the encodings from the classifications.

      There are plenty of other ways to do zero shot classification that would result in more "token usage" (really just having to reprocess everything for each class), but the pricing and the way they describe it narrows it down somewhat.

Open on Hacker News to reply ↗

Unofficial Hacker News client; not affiliated with Y Combinator.