‹ BackHN Continuity

Thread

Decision models like Jev don't beat LLM-as-a-judge or traditional classifiers

96 points · 34 comments · tomncooper

  1. petesergeant · · focus · HN ↗
    There are plenty of benchmarks that show they do, too, though, so this is a single data point.
Open on Hacker News to reply ↗

Unofficial Hacker News client; not affiliated with Y Combinator.