It's not the training data. When we've tested the writing style of base models that have not gone through instruction tuning, they're much more human-like (
https://arxiv.org/abs/2410.16107). The style shift seems to come from something in the instruction-tuning process, so our current research problem is figuring out
what in the process is doing it.