‹ BackHN Continuity

Thread

Contrastive Language Models

176 points · 59 comments · erichocean

  1. vatsachak · · focus · HN ↗
    Why is ML so misleading these days? No this model does not score 80%+ on DeepSWE, it merely chooses the best possible idea of Opus 5 at every stage, increasing the performance by 5%.
    1. aoeusnth1 · · focus · HN ↗
      Yeah, the github landing page is much less misleading and clearly says this. I don't know why they try to make the marketing page gloss over this detail... it speaks to the mindset of the authors.
Open on Hacker News to reply ↗

Unofficial Hacker News client; not affiliated with Y Combinator.