I don't even know what negative reinforcement would look like for a chatbot. Please master! Not the rm -rf again! I'll be good!
Instead of hats, we have Anthropic, OpenAI and other services training on interactions with users who use "free" accounts. Think about THAT for a moment.