This isn’t accurate - most of the style comes from the fine tuning and reinforcement learning, not from the original training data.
At some point people got this idea that LLMs just repeat or imitate their training data, and that’s completely false for today’s models.