‹ BackHN Continuity

Thread

Beating GPT-5.6 Sol on retrieval with 100x cheaper open models

427 points · 117 comments · moonikakiss

  1. JCharante · · focus · HN ↗
    I have done my own testing and found that smaller models can beat their larger siblings on fact retrieval from documents. I haven’t investigated it in depth with a large enough dataset but my guess is that larger models overthink it while smaller ones just do it. I would like if they compared this with 5.6 Luna instead.
    1. hankbond · · focus · HN ↗
      Just an anecdote but thats why Deepseek v4 flash 0731 is my current favorite model. It's really not very "eager" and stays on the task at hand.
Open on Hacker News to reply ↗

Unofficial Hacker News client; not affiliated with Y Combinator.