How many bots are on Twitter? Question is tough to answer and misses the point
niemanlab.org
niemanlab.org
When you sit down and read these papers you find the whole field of academic bot research is just ideologically motivated quackery. It isn't taken seriously by anyone who actually works in the spam fighting industry, but does harm the overall credibility of science. If these people actually want to work on bot fighting (like I did) they need to quit academia and go apply for jobs at tech firms. Anything they try produce externally will be noise, because as they acknowledge:
"External researchers do not have access to the same data as Twitter, such as IP addresses and phone numbers. This hinders the public’s ability to identify inauthentic accounts."
It doesn't hinder it, the lack of such basic signals makes it impossible, which is one of several reasons their studies keep yielding junk-quality results.
Further reading:
https://blog.plan99.net/fake-science-part-ii-bots-that-are-n...
https://blog.plan99.net/did-russian-bots-impact-brexit-ad66f...
https://papers.ssrn.com/sol3/papers.cfm?abstract_id=3814191
https://archive.org/details/hopeconf2020/20200726_2000_Peopl...
To be honest he does have a point (I hate this) as I could easily create and app with 500 bot users and 50 real users. Selling it for if I had 550 is really scummy.
It can help to solve problems with bots, account theft, as well as accountability for hate speech/intimidation (the system should allow law enforcement to reveal identities over which they have jurisdiction).
If you include Youtube under the label Google (i.e., Alphabet), it's even worse, there is tons of spam and deliberate political misinformation in the comment sections, although arguably most of these are probably generated by humans. For example, any TV report about Ukraine on Youtube is spammed with pro-Russian troll posts from St. Petersburg and likely hired posters from India. There is also lots of commercial spam in comments.
So no, I don't think they've solved the problem. They're using captchas against bots and that's about it.
Spam is also not the same thing as posts you disagree with. I went to YouTube, searched for [ukraine war report] and picked the first result that came from TV news to try and check your claim [1]. The top comments are all clearly human, some are claiming to be Ukrainians, and are all anti-Russian.
Regardless, your post is a good illustration of the problems with academic bot research. Although researchers claim to be researching spam bots, their methodologies are often defining any post they don't agree with politically as "spam" or from "bots".
I agree that spam by bots differs from spam by humans. However, they need to be dealt with in union because they usually go hand in hand, i.e., bot nets often support and amplify select human posts and the human origins of these networks can be mapped and traced back to particular actors and sources.
I doubt anybody would disagree with your general claim that there can be methodological problems with detecting any kind of spam, whether by bots or by humans. Of course, that's difficult, especially for researchers with only limited or no access to sensitive account information such as creation date, post history, IP numbers. If it was easy, then the spam problem would have been solved already, but I have argued that Google hasn't solved it.
But how do you know they're misinformed/intending to misinform, unless you disagree with the post?
Last but not least, in the case of Russia's current aggressive war against Ukraine, it's not just about misinformation. Not everyone is aware of that, but in many jurisdictions supporting Putin's war can constitute a crime punishable under penal law, just like using the word "war" can constitute a crime in Russia. In my opinion, all posts in favor of Russia's aggression should at least be removed from social media. This shouldn't even be controversial, 141 countries have condemned this aggression which violates all international laws. It was not common during WW2 for US companies to give the German Nazi party free airtime on all broadcast channels. People could have their personal, pro-Nazi opinions in allied countries, but expressing them publicly had - and should have had - negative consequences. There is no reason to think this should be otherwise in the current conflict, which is essentially a large proxy war between NATO countries + Ukraine as defenders against the sole aggressor Russia. Not just misinformation should be stopped in this particular case, but also open expressions of support for Russia, displaying the swastika-like "Z" sign, and so forth. And, of course, the NSA and other intelligence agencies should hack back to destroy Russian propaganda outlets and take them off the net.
However, I admit that this last point kind of deviates from the original discussion. In a nutshell, there are many indicators for spam and many indicators for misinformation, and you judge these in a way similar to how a judge would evaluate cases that primarily involve circumstantial evidence. Humans can do this reasonably well, algorithms are still bad at it.
That said, the question is certainly worth asking beyond the realm of advertising. Media has a significant impact on people and society beyond its profitability.
If those impressions are off by X% due to bots, that's a problem. Especially if as a public company you've declared that issue at 5%, and it's more.
If you pay an advertising company a rate per view and the advertising company charges you knowing that a large percentage of those "views" were fake then you're being ripped off.
If you are advertising just to get an idea out, or for something you have no capacity to measure, sure. You need to know accurate numbers. But if you can track conversions on your end, and conversions are good, does the exact number of real views really matter?
Note. I’m not American and have no political affiliation to have selected NYT.
EDIT: this was before the shooting that's just happened.
It was mostly “influencers”.
One thing I have noticed is that Twitter REALLY wants me to follow Elon Musk. He’s at the top of pretty much every recommendation I’ve seen on the site.
Lots of people follow the news / latest most followed celebrity account and dont engage with that person at all.
check out the views to comments ratio on this https://www.youtube.com/watch?v=K4II9cYrQvE
More important, nothing has changed about the bot problem since Musk signed the merger agreement. Twitter has published the same qualified estimate — that fewer than 5% of monetizable accounts are fake — for the last eight years. Musk knew those estimates, and declined to do any nonpublic due diligence before signing the merger agreement. He knew about the spam bot problem before signing the merger agreement, as we know because he talked about it constantly, including while announcing the merger agreement. If he didn’t want to buy Twitter because there are spam bots, he should not have signed a contract to buy Twitter. No new information has come to light about spam bots in the last three weeks.
What has happened in the last three weeks? Well, the prices of tech stocks have gone down, making the $54.20 price that Musk agreed to look a bit rich. (Snap Inc., a social-media competitor to Twitter, is down more than 30% since Musk made his offer on April 13.) And the price of Tesla Inc. stock, which he is relying on to finance part of the purchase price, has also gone down, making him poorer and making the $54.20 price look even more expensive. (Tesla is down almost 30% since he made his offer.) So he is angling to reprice the deal for straightforward market reasons. But that is very clearly not allowed by the merger agreement that he signed: Public-company merger agreements allocate broad market risk to the buyer, and he can’t get out just because stocks went down.
So he is pretending that he wants to reprice the deal for other reasons. He is not pretending very hard — the poop emoji is not going to hold up in court! — but he’s doing enough to confuse the public and give his fans a pretext to believe that he is really the victim here."
Twitter's Board are a clown car.
[1] https://www.bloomberg.com/opinion/articles/2022-05-17/elon-m...
bit of clickbait
Musk wants to know so he doesn't have to pay the price he agreed to pay given the entire tech stock market just took like a 50% hit. It's nothing to do with bots and everything to do with getting out of the deal/getting a more sane price and honestly fair enough I wouldn't want to massively overpay for something like that either.
For users I see nothing useful. For third party using users tweets to get live coverage of something a little use.
For human network and mud analysis by Twitter owners and those who buy/commission such studies a significant use.
So the point is: WHY THE HELL so many people want to work for free mostly against their own interests spending time on the platform? That's an interesting point because having flock of humans (sorry for being so rude, but that's is) willing to work for free for someone else interests is VERY interesting... I have many things I'd like to outsource for free to third parties willing to work gratis for me. Just basic stuff eh! Who want to came at my home to clean it up regularly for free or even pay for being allowed to clean up more regularly?