>This isn’t accurate - most of the style comes from the fine tuning and reinforcement learning, not from the original training data.
Fine tuning, reinforcement, etc are all 'training' in my books. Perhaps this is your confusion over 'people got this idea'