It's not the training data. When we've tested the writing style of base models that have not gone through instruction tuning, they're much more human-like (<a href="https://arxiv.org/abs/2410.16107" rel="nofollow">https://arxiv.org/abs/2410.16107). The style shift seems to come from something in the instruction-tuning process, so our current research problem is figuring out what in the process is doing it.
jldugger · · focus · HN ↗
mohamedkoubaa · · focus · HN ↗
capnrefsmmat · · focus · HN ↗