> most AI companies will slightly change the way their AIs respond, so that they say slightly different things to the same prompt. This helps their AIs seem less robotic and more natural.
To my understanding this is managed by the temperature of the next token prediction which is picked more or less randomly based on this value. This temperature plays a role in the variability of the output.
I wasn't under the impression that it was to give the user a feeling of "realism", but rather that it produced better results with a slightly random prediction.