‹ BackHN Continuity

Thread

Jev in 25 Lines of Python

691 points · 212 comments · bashbjorn

  1. bruhhhhhh · · focus · HN ↗
    I am hearing about Jev for the first time here so no idea about the hype. So their(Jev) is that the thing is faster at classification than a frontier model? Because the whole type safe aspect is already fully solvable with structured output. But their example is classification but that would also be possible and faster with a classic BERT model. So their pitch is a task specific smaller model or am I completely misunderstanding the whole thing?
    1. killerstorm · · focus · HN ↗
      You need to train data for a BERT-based classifier, and then there's a risk that it will pick up specific biases from the data instead of what you want.

      As far as I understand, the idea of Jev is zero-shot or few-shot classifier: it learns a lot of stuff at pre-training, but unlike a classic LLM it doesn't need to learn how to chat, so it can be much smarter at a particular size

Open on Hacker News to reply ↗

Unofficial Hacker News client; not affiliated with Y Combinator.