‹ BackHN Continuity

Thread

Decision models like Jev don't beat LLM-as-a-judge or traditional classifiers

96 points · 34 comments · tomncooper

  1. dominotw · · focus · HN ↗
    prompts that these evaluations were done are too trivial
Open on Hacker News to reply ↗

Unofficial Hacker News client; not affiliated with Y Combinator.