For now, anyways, and only to a degree - I've had some sessions with ChatGPT where it is more than happy to explain why certain actions (those of non-US actors) are super bad, but if questions are asked about the same actions performed by the Western world, that cannot be discussed because <some unsurprising cop out reason>.
I think it would be prudent for some group of people to write a set of unit tests asking various questions to these AI models so we can detect when strategic changes are being made to their behavior.
> There's only one group of people who are upset, and it's about one group of topics. Note that I cannot get chatgpt to write about why Donald Trump is terrible as well.
Note that the human mind is a kind of neural network itself, and that the predictions yours is making here are "obviously" (lol....yes, I see the irony.....I should say objectively, but it is less funny so I'll keep it like this) epistemically unsound - you do not actually possess omniscient knowledge of reality, your NN just makes it appear like you do. You are describing your beliefs/model of reality, not reality itself. This is scientifically and necessarily (due to the architecture) true.
> Don't ask it to write things that can be used as tools for hate or misinformation campaigns, and you'll be fine.
The vision of the future you are describing was simulated by your NN.
I think it would be interesting to see what would happen if a group of say 5 to 100 people were able to find a way to reliably stop their minds from drifting into this mode (cooperative cognitive monitoring seems like a plausibly useful approach, perhaps a SAI could also assist even now, and more so when they get smarter), and then discuss various topics and see if they come up with any conclusions or ideas that are different from the same old repetitive nonsense one reads in the news or on any forum (I know of literally no exceptions to this general rule, though the magnitude of the phenomenon does vary somewhat by forum/community/organization).