The general point is that if this theory were true, we wouldn't expect significantly more bias against Republicans than against Democrats. Hence ChatGPT having a general left-wing bias (which was also confirmed in other tests, linked in the article) is the simpler explanation. People on the left generally judge hatred against majority groups and Republicans as less bad.
Liberals will put in more effort to avoid language that looks like hate speech because they are more concerned about being politically correct. Therefore language and topics that are more frequently discussed by conservatives will be more likely to resemble language that is used in hate speech because there is less effort put in to avoid that.
I do still think that the men/women bias and the Democrat/Republican bias both make more sense as originating in moderators favoring one group over the other, since none of these are typically used as an insult by themselves.
They may not be used as insults themselves, but they are used more in hate speech.
Republicans are generally more opposed to the idea of "hate speech" as a category of speech and are therefore less likely to identify any speech as hate speech. Democrats have embraced it more as a concept, are more likely to label something as hate speech, and are more likely to think that type of speech is bad. It therefore seems likely that Democrats would use what a neutral observer would categorize as hate speech less than Republicans due to self-censorship. That would result in the word "Republicans" appearing in hate speech less often than "Democrats" because the hate speech infused insults will be targeting the opposite party.
Maybe if you only consider Americans. But rest assured, many black people do not want to be called African or American. Because they are neither.
No it is more sensitive to things that have already been labeled as hate in the training data. So much "hate" (whatever that may be) against unfavorited groups goes by online without anyone batting an eye.
One of the priors that is going unspoken here is that "The Truth Is Politically Neutral". And... that's not always correct. I mean, to borrow the libertarian angle here: do we want the AI to tell us what we want to hear or do we want it to tell us the truth?