Google's Anti-Bullying AI Mistakes Civility for Decency
motherboard.vice.com
motherboard.vice.com
I'd love a chance to have civilized discussions with people who I think are very, very wrong. It'd be a chance to actually convince them they're wrong, instead of just screaming hateful things at them and walking away.
I'm not at all worried about them convincing me to start hating people, or even to convince other non-hateful people to start.
Eg, I too would love to have the discussions you speak of. Despite this, I often see an escalation in tone and fight the urge to respond in kind. I often fail at not responding in kind. Doing so of course escalates the conversation, something I've now contributed to, and it will likely escalate yet again. Over and over, until all reasonability in the conversation is dead.
If I knew it wasn't allowed, and especially if it gave me some type of non-public meter of how much further I can go before my conversation will be blacklisted/etc, then people like me have clear visual feedback over .. well, not being such a douche.
It would be like having a state patrol a few lanes over on the interstate. If one is around, most people behave quite well, because we can all see what will happen if we step out of line. Likewise, our speedometer gives us the visual feedback to know what is out of line (with regards to speeding of course).
It would feel very weird at first, especially if the system was not perfect, but at it's root I think it's needed for us to be able to discuss sensitive topics. Which, tend to be the ones we need to discuss the most.
That's what having visible cops patrolling the highway devolves into and it's pretty terrible.
I'm not sure if your opinions are bad or you just made bad comparison.
In the case where there is a forum that requires civility, yes I do want a tool to help me stay civil. Likewise, I also like having my speedometer. It's nice to know when I'm speeding, and when I'm not.
IMO, analogies are fair game, and a useful tool for simplification. They can also work well for highlighting hypocrisy, when a supposedly universal rule is not being applied as widely as we thought it was.
There's nothing wholesome or 'meaty' about vulgarity. It raises the temperature of the argument and doesn't contribute useful information. It's more akin a diner pouring pig slop on a restaurant table and insisting that it's the other patrons' fault for finding it impossible to eat in those circumstances.
I think it is just as important to take things at face value, instead of assuming malice, as it is to be sensitive to other people's issues (for lack of a better term). I've actually had decent results debating obvious trolls, just by postponing emotion, identifying and avoiding assumptions, and focusing on the actual meaning of their statements.
I'm referring to resorting to crassness on an already contentious topic. As an example, I often spend time defending my religious beliefs to atheists. I'm happy to talk all day about how I feel about Dawkins' arguments and the like, but the second someone brings up 'your stupid f———— sky fairy', I can tell that the conversation is not going to be productive any longer.
Those very definition are the first ones to be used (followed by "think of the children") as reason for righteousness, censorship and other politically correct legacy.
The point of free speech is to ensure that unpopular political opinions have room to air, everything else is a happy side effect.
(Side note: "Rhetoric" isn't a dirty word, or at least it wasn't originally a dirty word. In the classical definition of the term, rhetoric based on facts and cool logic was also possible, and was also considered rhetoric. The modern usage of "rhetoric" to mean exclusively "empty" or purely emotional argumentation is not what the classical politicians meant by it, and not how the term is used in an academic context.)
Therefore, it's probably impossible to keep the heightened emotions out of debate without strict rules on content and participation. The extreme end of this is Usenet-style moderation, which means every comment to a forum ("newsgroup", or "group", in the Usenet parlance) is sent to a moderator's email address to be specifically (usually manually) approved, and no unapproved posts will appear in the group. That kind of moderation has existed since the 1980s, but to my knowledge, no modern web forum moderates that way. Take it as the far end of what's already been tried.
That kind of strict moderation can look like content-based or viewpoint-biased filtering (as in, all messages with a specific content or which take a specific viewpoint) due to a simple founder effect: If the few people who want to discuss this topic, or take this viewpoint, happen to be intemperate, they get moderated away, and it looks like that topic or viewpoint is being moderated away. The forum gets a reputation for that kind of filtering, and it can be very hard to shake it.
Now, are there topics or viewpoints which are inherently inflammatory? That's another question.
History proves this, long before the advent of digital fora. Consider the elaborate, stilted parliamentary procedure and its variants. While it's quite burdensome, it's also considered a total necessity in many of the great deliberatory bodies that stand today.
Even with such a ceremonious structure, fist fights in national parliaments are by no means a thing of the past.
I think that this is because more than persuasion, humans are emotionally driven to create in-groups and out-groups, and only by divorcing emotion as much as possible can any real reasoning and discussion take place.
If there's one thing that profanity and vulgarity is actually good at, it's forming an in-group, which is antithetical to good discussion.
So I think a slightly different analogy would be that it's a restaurant for dieters, and with each drink/food-item, they give you a running tally of how many calories you've ordered.
Ie, if you want to eat at that restaurant, clearly you want a diet focused dining experience. They also provide you with a tool to monitor your calorie intake.
Likewise, a forum which bans users over civility is clearly one populated by people who seek that. Angry internet voices would clearly be unhappy there, and would not even want to take part. Likewise, that forum also could give you a tool to monitor how un-civil you're being, to help you stay civil.
The tool that I spoke of would be an aid to people already choosing to take part in that experience.
This would all be a very different story if it was forced on everyone who wanted to say, use the internet. I think your example would more closely resemble a police state. Unwanted rules, shoved down your throat.
it will eventually learn as user reports and corrects the correlations between normal and toxic in full sentences, but it's still lacking as of now, and not only for controversial topics: https://i.imgur.com/HGrs4ze.png
One that can learn and get better sounds amazing and would likely be better-tolerated by the userbase.
It could even learn automatically over time as the moderators mark each toxic post as such, and un-mark the ones that got mistakenly auto-censored.
We should expect some people to feel really hurt by a topic, depending on how close it is to them, even if the wording is civil. For example, it's easy for me to have a civil debate about cross burning, because no one I know has ever had a cross burned in their yard. But I shouldn't be surprised if people who have seen that sort of thing find the debate really really uncomfortable, and maybe even if my comfort with the topic comes across as not having sympathy.
So it could be that "XYZ abstract argument about cross burning" is true, and also that talking about it in a dispassionate way is hurtful and rude. I'm not sure exactly what to do in situations like that, but giving 50% or so of the discussion time to acknowledging those feelings seems like a good starting point?
Giving half the time for debate to talk about feelings that are not "debate-able"? Thats probably a waste of time that will derail your debate entirely as people will start debating those feelings.
One doesn't need to scream to be toxic.
Removing your hand from fire is generally an intelligent reflexive action ~= AI. Understanding implied meaning of arbitrary text is orders of magnitude more complex ~= AGI.
https://www.theguardian.com/football/2017/aug/19/know-vital-...
"I genuinely don’t know if my colleagues are making fun of me or being nice"
Terminological inexactitude [0]
Economical with the truth
Tired and emotional
[0]https://en.wikipedia.org/wiki/Terminological_inexactitude
"Withdraw that remark"
"Ok, half the Tory members are not crooks"
http://www.theweek.co.uk/amp/62692/dennis-skinner-quotes-the...
This is why who programs the AI has such influence over the value judgement. Just because AI does it, doesn't mean there is no human influence over it. Actually AI doing it instead of army of humans means handful of people have control over the outcome.
That definition of "toxic" is defined for Google's business. Keep the user's attention and eyes on Google's web properties serving Ads.
What is toxic? Anything that makes a person leave, reducing Ad imprints.
A person may stop commenting if they are convinced or challenged of their views. In a normal conversation, all tools of a language, comedy, sarcasm, hyperbole etc. can be used to show someone a new way to look at things.
Google AI is going to put evolutionary pressures on language features. Language features that keep a person watching Ads on Google properties will be selected for propagation.
It is not in how it scores, but in the very fact that people who built it are normally incapable of giving a given sentiment more than one thought.
(Edit) The article implies that such a discussion implies that someone wants to promote slavery.
I would think that such a discussion would be very productive about reminding those of us who didn't live through slavery why it was that bad.
If we can't discuss "was slavery really that bad," there are people who will go around really believing that slavery really isn't a bad thing. That's why these kinds of discussions are critically important.
I understand that the AI is merely doing what was programmed in and there are limitations, but it still draws me to the conclusion that it is really more of a moral thing than something to weed out toxic speech, as evidenced by the hard lines taken against folks telling an aggressor to "fuck off".
And for something interesting:
You've been eating paint chips and chasing them with leaded water, haven't you? only ranks in at 25% toxic.
I tend to use words considered "rude", to many native English speakers, merely for flavor when I'm passionate about a certain topic or argument.
To be fair: I do the same in my native language, I tend to use quite colorful language sometimes, so it's probably more of a personality thing.
Some phrases simply translate in a rude way: Norwegians seem to have a liking for the word "fuck", which has made its way into Norwegian slang as well. It can come off quite rude back home. For folks I meet that learn mostly through television, it can get even worse because the shows don't always portray dialog in a natural way.
Those folks should start watching some fucking proper television ;)
Good point about "fuck" creeping into the slang of a lot of other languages, I've witnessed the same very often but never really noticed until you pointed it out like that.
Curse words seem to have a certain attraction in that regard, I probably know the equivalent of "fuck" in like 4-5 different languages, while understanding literally not other phrases in those languages.
It's fine if the system reminds you that you are not being particularly polite or toxic, but filtering communication on the internet by this manner can't be the solution.
'Do you know that Newspeak is the only language in the world whose vocabulary gets smaller every year?'
..."toxic" being defined as a "rude, disrespectful, or unreasonable comment that is likely to make you leave a discussion."
There are definitely things said that can cause me to leave a discussion, even if said in a civil manner. I am not going to stick around to debate a person who believes the earth is flat, no matter how polite.
>What's toxic? >This model was trained by asking people to rate internet comments on a scale from "Very toxic" to "Very healthy" contribution. Toxic is defined as... "a rude, disrespectful, or unreasonable comment that is likely to make you leave a discussion."
subjective, unquestionable nonsense
Larry sells donkeys: 87%
Larry sells llamas: 16%
Larry sells beef: 17%
Larry sells bananas: 45%