‹ BackHN Continuity

Thread

Claude's Load-Bearing Seams

120 points · 50 comments · rzk

  1. jldugger · · focus · HN ↗
    So like, who writes like this and how do we delete them from the training corpus?
    1. mohamedkoubaa · · focus · HN ↗
      Do we know if it in the training data or if the humans giving rewards in the labs have such awful taste in prose?
      1. capnrefsmmat · · focus · HN ↗
        It&#x27;s not the training data. When we&#x27;ve tested the writing style of base models that have not gone through instruction tuning, they&#x27;re much more human-like (<a href="https:&#x2F;&#x2F;arxiv.org&#x2F;abs&#x2F;2410.16107" rel="nofollow">https:&#x2F;&#x2F;arxiv.org&#x2F;abs&#x2F;2410.16107). The style shift seems to come from something in the instruction-tuning process, so our current research problem is figuring out what in the process is doing it.
Open on Hacker News to reply ↗

Unofficial Hacker News client; not affiliated with Y Combinator.