Analysis of 200M tweets discussing coronavirus suggests 45%-60% come from bots
marketwatch.com
marketwatch.com
* https://twitter.com/ngleicher/status/1264614994475315200
* https://twitter.com/benimmo/status/1265329734705197056
* https://twitter.com/alexstamos/status/1264643202293751808
* https://twitter.com/3r1nG/status/1264567090742206464
(edit to add more examples of counterclaims to this "research")
I don't know if Twitter has bots or how they tilt. What I do know is post-coronavirus commerical "news" outlets have lost my trust, as has much of the political establishment, red or blue.
Can you please share examples of specific news reports that we’re non-factual?
One can explain this, and explain that, but when the sum result is being locked inside for months on end, unemployed, blocked from major life events with a media wrongly cheeleading those policies with wrong and unsupported information, this leads to a loss of trust. I'm not saying the media outlets are on a lark; I'm saying they haven't done their jobs of reporting facts outside the filter of propaganda and political bias which has infected so much policy dialog.
[1] https://www.marketwatch.com/story/no-chinese-allowed-racism-...
[2] https://www.boston.com/culture/lifestyle/2020/04/16/is-it-sa...
[3] https://www.theguardian.com/science/2020/may/28/questions-ra...
Well, let's see.
> steady diet of fear-based public health reporting
There's a pandemic on with a nontrivial death rate. A little fear's deserved.
> that the "news" outlet
I believe these are called scare quotes.
> finally has explicitly disclaimed as non-factual
where is this disclaim?
> (but always was defective)
You have certainty on your side apparently.
Now the quote about stuff getting out of date, well, what do you expect? Speak to god if you want utter certainty, not scientists. Seriously, what makes you think eternally correct information drops fully formed into anyone's lap.
If you want to live in a world without science, aka the middle ages, I urge you to do so, so those of us left behind with the internet, antibiotics, human rights, a stable civil society, non-agricultural jobs, and education, can suffer alone.
https://www.cs.cmu.edu/news/nearly-half-twitter-accounts-dis...
I imagine they're going to publish the research, but it doesn't look like they have yet.
...though some of their methods are described.
> CMU researchers since January have collected more than 200 million tweets discussing coronavirus or COVID-19. Of the top 50 influential retweeters, 82% are bots, they found. Of the top 1,000 retweeters, 62% are bots.
Fascinating article. Thank you for taking the time to find it.
If you google my name, unlike your "throwaway" anon name making assertions, irony of ironies, you will see I am a real human being, fully qualified to check into data science type assertions like this (and in fact; I have). Which is a lot more than "you" can say.
It's all subjective anger and whatnot. What am I or anyone supposed to take away from that?
> ample evidence that they're lying: again
Ok, what evidence. And don't tell me to google it or I can find it if I want to. You claim, you cite.
> You claim, you cite.
If a claim cannot be verified, it is not a claim per se: it's an opinion.
The issue they push with the bots may not be, and likely isn't, their actual position on the issue.
Too bad we can't have a Geneva Convention on botnets or something. (I would write "couldn't", but don't believe it's technically feasible)
[1] https://en.wikipedia.org/wiki/Foundations_of_Geopolitics#Con...
It's even crazier that Twitter can't do that research themselves and remove those bots.
Obviously the big debate is censorship. It feels like there could be a tech solution to this, where conversations aren't shaped by volume, but I guess that is what Twitter is anyway.
It’s a lot harder problem then it might sound on the surface. And realize if you find bots 100%, they will change for round 2 and the game starts over.
Just put in a super basic unlock process that is costly to game for "non-real" people
I think your "just" is hand-waving away a lot of complexity. A person deployed these bots, and if a bot account gets put into an arbitration process, that person can go through the process themselves to pass the arbitration, then set the bot running again.
That is not scalable, which is exactly the point.
(Assuming Twitter was actually incentivized to be anti-bot/bot-like behavior.)
And it wouldn't even stop the bots, because Twitter isn't limited to Americans, so bot makers could easily send fake official IDs from someplace Twitter can't check. Or send stolen offical IDs from other data breaches.
You can generate realistc photos of IDs from countries that Twitter cannot positively verify those bc in USA there are ways to check authentity of your document. But IDs from Russia or Ukraine - no way.
That would work only against casual bots. The big nation-sponsored botnets won't have issues with providing genuine government issued identities for bots. Which kind of defeats the whole purpose and wastes a lot of resources twitter would like to use elsewhere.
Whenever I sign up at Twitter (entirely because some companies do support there) the account is flagged within a few minutes for "suspicious activity", such as making a coffee and coming back to the PC or searching for the company I want to contact.
Think about an account that is active 24/7 every day all year, or one that fires off multiple actions per second for long periods of time.
Beyond that, having a limited vocabulary, repetitive phrases, consistently accessing the same thing every few minutes, etc. are all low hanging fruit.
Arguably, those "legitimate" users don't bring anything of value to the platform or society as a whole and should rightfully be shunned.
Most online forums or chats have automated rules whereby posting too fast or too often will get you slowed down or banned. Sometimes legitimate users get caught by the filters, no big deal. Twitter could make their existing rules much more stringent.
They also have so many phone number and email they can use that to do additional checks.
They you can use the network effect: bots usually have something in commons: topics, likes, ip, followers...
If you have millions at your disposal, something can be done.
They are not, so I'm more encline to assume there is benefit for them to avoid doing so.
This information would not need to be stored, it would just be needed to verify an account as belonging to a real person.
a) it was a non profit and
b) people didn’t cry “censorship!” every time they do anything to curb abuse and
c) it would be ok to deactivate a % of real users by accident (these researchers can’t tell who’s a bot vs who’s an actual user _for sure_)
It’s not an easy problem.
Dude, you can’t even form a grammatically correct sentence.
(I'll add the link if I can find it, but a recent example pretty much boiled down to "actually, these accounts giving many repeated identical replies have totally normal timelines otherwise. And it turns out, many identical replies in short time shows someone knows about copy-and-paste, not that they're a bot")
It's not in their best interest to wipe out the bots at any large scale. Sure, the occasional cleanup or purge as a token effort or to get the most egregious stuff but remove them all? Nope.
The difference in my perception is that the presence of the bots and bad actors on Twitter that come to the attention of reporting like this increases engagement and views, and thus top line revenue.
This isn’t saying “bots count as views, so we get more ad dollars”. It’s that bots and bad actors promote topics and conversations that bring more real users to the platform, and increase the session duration for new and existing users.
I would imagine that Twitter sees spam bots and purveyors of illegal content as unwelcome and probably has an engineering team that dispatches those accounts quickly. But whether deliberately or unconsciously, they probably don’t apply the same rigor to accounts that break the TOS but manage to drive the top line up.
I’d love to hear from an engineer from Twitter who works in this space.
I feel like there are 4 types of Twitter users: 1) bots, 2) influencers that heavily use automation to the point where they are almost the same as a bot, 3) famous/popular that are somewhat to mostly real, but have someone else managing their account for them, 4) real people who think 1-3 are all real that gained their followers organically.
Probably most users fall on categories 1 and 4 above, although most of Twitter traffic is probably generated by tweets of 2 and 3.
It almost feels like playing a video game where most other characters are NPCs pretending to be real users.
I think Twitter needs to add an option which marks an account as a bot or non-human account so that people can gauge instantly what the account's real motivations are. Most actual bots will ignore that setting and pretend to be human however, so the responsibility is on Twitter to weed out automated accounts that are obvious attempts to game the platform.
Kidding, but there's a real point. 40% of tweets are from bots? OK, how many "studies" are from bots? (I believe that's in the neighborhood of Drybones' point.) But how many comments responding to the studies are from bots?
Way back in the day, I remember reading the Cluetrain Manifesto. It said that corporate-speak sounded literally inhuman. Well, the bots have gotten better over the years. But also, it seems to me that the humans have gotten worse. They are less able to say "Wait, that doesn't make sense." They are even less able to say "Wait, that doesn't sound like the way humans talk." And that's bad. As the bots get better, the humans need to be developing better filters, not regressing.
* 50% of positive posts for candidate x are bots!
* Most people who support X online are bots!
Easy and cheap, and the message carries well.
That condition is the key. Are you sure they can do that? It’s very tough to do with certainty.
I often wonder when we are going to see more left wing reactions to this sort of behavior. Seems the right, based on my limited information, seems to be more aggressive about these disinformation campaigns.
Also it stands to reason that this particular misinformation campaign is being sponsored by a foreign state. I think that this would be an effective weapon to use against the United States. Sort of like biological warfare by proxy of disinformation.
Perhaps we should all step back and evaluate whether our internal thought processes are being shaped by bad actors via an appeal to our innate tribalism? Notice how easily you fall into the script: 1) it’s the other team only 2) my team should do something to fight back 3) maybe it’s sponsored by a foreign state.
If you primarily think one side is crafting ridiculously bad content and it’s therefore a misinformation campaign from a foreign state, I think it’s wise to consider how much of that is personal bias and if you’re perhaps “giving a pass” to stupid content that at least aligns with your views.
“Well, they meant well, so it’s probably just a right-minded American who made a simple off-by-factor-of-a-million arithmetic error and no one spent 50ms doing a basic sanity check. No one’s perfect.”
Or “Look, maybe I think Obama did an Ok job, but I saw all that content about he wasn’t born here, and there has to be something to it if so many people are talking about it. The government hides things; the truth is out there, ya know? Oh, and Epstein didn’t kill himself.”
Again, I said from my "limited view", as it's impossible to have all the information. It seems TO ME that there is quite a bit more targeted misinformation coming from the right.
I would love to be corrected. However, I don't think saying "I got a Facebook feed with lots of stuff stuff coming out of it", and "you're stupid because of xyz I assume about you for no reason whatsoever" qualifies in this regard.
I’m sorry for the portion of that where the fault is mine.
However, I am concerned about a disinformation campaign which could, theoretically, be sponsored by a foreign state to get people to ignore the very real threat of coronations.
I am also concerned about foreign influence, but I also observe that we meddle in other countries’ affairs as well, so it’s just a fact of life and social media changes it but didn’t start it. I’m more concerned with a population who isn’t thinking critically and with the right balance of short and long term (and is therefore especially susceptible to influence), but whether the influence is foreign or domestic doesn’t matter as much to me as the effect of the influence.
Is this even disputed?
Through there are many countries in which this can be indeed done reliably sand effectively.
Through the main problem is still this networks are global, so bots will just register with origins where IDs can be easily faked/stolen etc.
But it might still help a lot for local discussions if you would make it clearly visible if a person had a no or a non local real person identification.
If such a system is done cleverly it could also help law enforcement while it still upholds privacy. Through it's will always be someway prone to abuse by police and similar if not done perfectly transparent but you can be sure that certain lobbies will invest insane amounts of money into making it non transparent over time. Which makes such systems potential dangerous to have.
Not sure if you’re joking, but this would be a completely awful privacy situation. Why would you ever give your social security number to a social media site?
That said, with reading this, if you are thinking of skipping it. Does somewhat confirm that most of the not posts are on conspiracy theories. Which, I guess is not that surprising. Even if the conspiracies are.
It's not unlikely the "So?" is meant as an "alright then, can we now have the answer to ...". I've made that exact mistake.
As for twitter bots, I have one running by ifttt, another by script, and three more in planning. I welcome the human/bot accounts being made clear.
https://www.usatoday.com/story/news/nation/2020/05/29/corona...
Real people have also been out protesting the lockdowns, in the real world, not on Twitter. People were legitimately upset, rightfully so, when protestors were showing up with guns in Michigan.
Edit: not very useful these days, but still of possible interest: http://www.gkstill.com/Support/crowd-density/100sm/Density1....
Somewhere I'd read a book on propaganda by a fellow who produced it in ww2 and finished his introduction with biblical examples of the main techniques.
> Namely, their rationale was something like this: the world is dominated by people whose opinions are fundamentally shaped by what others are saying, and moreover, by how frequently other people seem to be saying it. They don't really think about things or rely on their own experience: they're just mimics.
> In such a world being able to auto-generate fake messages of support for some political position or another at scale would give you immense power, because the population would automatically swing behind you based merely on the perception that everyone was swinging behind you.
> So that's their fear. But is it realistic?
> Well, this is where we get into the polarisation. "You aren't smart enough to have an opinion" is the sort of viewpoint that leads to an elitist vs populist conflict. There have been reams of analysis about this and it's not really AI specific - e.g. did people vote for Brexit because of Twitter, or because of things they saw in the press, or did they vote based on their own experiences, or what their close friends/family thought, or what mix of "all of the above" is the truest mix?"