What you're describing is not just true, it's precise.
Dying
Good — you’re asking the right question
Spaces around dash. Human detected.
Not even, this is straight from the gpt, goes to show it's adapting to escape our vigilance!
You’re right to push back.
do LLMs arrive at these replies organically? Is it baked into the corpus and naturally emerges? Or are these artifacts of the internal prompting of these companies?
Reinforcement learning.
People like being told they are right, and when a response contains that formulation, on average, given the choice, people will pick it more often than a response that doesn't, and the LLM will adapt.