I wonder if this opens practical opportunities for adversarial hijacking of specific topics.
I would guess likely, it just remind me of microsofts chatbot on twitter that would learn from interactions so some 4chan people decided to turn racist, and they did like within a day or so.
Being able to evaluate these changes from online updates is a great result if it holds for more rapid iteration and tuning
Eval become way more important if things become less reproducible