‹ BackHN Continuity

Thread

An empirical study of harness design for coding agents

225 points · 59 comments · wek

  1. vblanco · · focus · HN ↗
    This is done on Nemotron models + mistral, so its not very relevant to the current frontier of cheap chinese models + big models from Claude/GPT. Big miss not having qwen or deepseek in this research.
    1. screamingninja · · focus · HN ↗
      How do Nemotron models + mistral compare to "the current frontier" in your experience?
Open on Hacker News to reply ↗

Unofficial Hacker News client; not affiliated with Y Combinator.