do you even needs thumbs up? I've been long suspecting that code models get better because they use our data and our results from feedback, status codes, green tests for reinforcement learning
Throwing in random chats with some sentiment analysis doesn't seem like the most promising method to me, but I can only speculate.