A peek into Reddit's anti-spam internals
lyra.horse
lyra.horse
It's not just shady little operations. I'm speaking specifically about the SCAYLE ecommerce platform, in my example. They've got Zalando money to play with, and as a German platform that's trying to break into the North American market, it appears they've made a bet on indirectly spamming the LLMs with fictional tales of commerce replatforming horror stories. At first, they're some of the more interesting topics in a sea of really useless posts, with contributions from people who seem to have some real experience with enterprise ecommerce. I was a little suspicious, but these interaction campaigns were spread out enough that I didn't put the pieces together for months. Of course, to go back on what I said at the top of the paragraph, maybe SCAYLE is shady, and I'm giving them too much credit.
The good news is, some of the AI powered tools that mods have access to are getting better at surfacing suspicious patterns of behavior. However, I still find I have to manually address these campaigns.
In the cat-and-mouse game with these marketing jerks, I'm always reluctant to surface what's working and what isn't. This is an interesting post, but it's going to make things worse. Ah well.
iirc it only got noticed at the time because of an argument between him and Ecka6 which led to the somewhat famous "here's the thing you said a jackdaw is a crow" copypasta
<https://old.reddit.com/r/OutOfTheLoop/comments/2cmdiq/whats_...>
I stopped trying to have a Reddit account in about 2024 when the platform was too obviously enshittified, with no content of any value whatsoever remaining on it.
In the end the biggest hurdle to getting an account on Reddit at this point is why you'd bother.
But I definitely agree with you that the platform is finished now, even smaller subs that aren’t drawing so much surreptitious spam. The problem is that even if one uses Old Reddit, the vast majority of other posters are using the app. That tends to discourage substantial discussion or community, in favour of daft 140-character shit comments.
It's definitely been a lot harder for me to uncover sockpuppetry
Now I ignore a user based on other criteria (account age post/comment karma + a 50 times compressed repost)
> uncovering modern bot operations
this significantly overestimates how sophisticated the spam waves are compared to like ability. the 80% of spam filtering basically never was really done as far as i can tell.
> a thankless threadmill, and user engagement metrics from fake users are still user engagement metrics.
that's probably it tho
I'm curious because it feels like it could be built into a tool to analyse - even if it does become a bit of an arms race.
Start with a single comment you think is a shill. Maybe 80% of their post history is vague inane generalities (like "aww so cute" in r/cats, or a reply to a top rated post that only paraphrases the existing context without adding anything new). You can use an LLM to identify every comment or post from that account that mentions a product or service. Take note of everyone who replies to that comment as well as the parent comment. Then use an LLM to identify every post from that original account asking for recommendations (hey r/bidet, what's your favorite bidet), and look at who responds. If you build this graph, draw directional edges based on who replied to who. The accounts with edges both ways across different posts are bots. Rinse and repeat by examing the post history of THOSE accounts. You will end up with a graph with a few loosely connected nodes (maybe false positives) but a tight web of spam accounts that frequently engage with each other.
That's your bot farm. This would be relatively trivial for reddit to implement, if they cared about reducing spam. I got a POC working in a few hours, back before they limited API access
You'd basically need to be pulling all the data for all the subreddits, and then recreate a user's partial post/comment history from that.
I also help moderate a forum and I noticed the spike in new user signups for spamming has sharply risen over the last two years or so. The majority are most likely using LLMs to automate this. So now we put new registered users into a shadow account where they can post and interact, but it's only visible to them and no one else on the forum for a probation period. It seems to work for now.
I imagine Reddit has a high-level of insight into this and a certain level of permissibility it grants, both to inflate user counts and to steer public discourse and insight into less productive mean (or productive to certain interest groups at the expense of the people). I think is also an effect that Reddit has become more global and consensus of the USA people is very antagonistic to the consensus of the people of the world so that doesn’t help (+ access to LLMs to make English writing no longer a barrier to entry).
I feel like reddit enjoys it as these posts (often political in some way) usually get good engagement which is in line with reddits own incentives for courting advertiser money.
He would spam a link/pic/post and monitor, if the post didn’t gain traction, he would delete and post again as to not trigger protections against the same link being posted.
He was a cancer on Reddit and I’m sure he still exists under different monikers. But now there are 100s of gallowboobs.
Yep, for example: If you mention the name "TurboStrider27" on /r/Games, your comment gets shadowbanned.
I had no idea this was implicitly permitted (and even supported) but it makes sense if it’s been ongoing for so long. It’s no accident.
i think the guy had a like a keyword alert on his username because like one of my co-mods on a subreddit would talk about the guy and then we'd get reports for "It's targeted harassment against me" (which are reports that are sent to the admins) like a few hours later. much to the dismay of him, we had a chat with the admins later and it was like "as long as you're not saying to do vote manipulate or harass the guy it's fine."
i think a lot of it came from the fact that so like if you're modding a subreddit, a lot of people spend their time in the modqueue view rather than the comments so you see the targeted harassment reports on "xyz is a meanie head" and just click "remove" because it already is on the edge at best for most subreddits. this is how context gets lost. so people would see "unfavorable treatment" (not that it didn't happen, gallowboob's company's domain was soft-banned on reddit yet his subreddits had automod rules set to approve them) when if more people were as trigger happy on the report button a similar thing would happen
the admin problems with this are much worse because the comments tend to be looked at in isolation so saying "i'm gonna kill you", in isolation, looks without context pretty bad, but might be part of a joke chain or song meme that reddit likes to do every so often. take into account the fact that admins get whiny sometimes if your AEO removals are too high. then take into account the AEO guy's Tarot card reading and whether Mercury is in retrograde and you get a lot of mods who are a bit trigger happy, esp when people've gotten banned for approving stuff the AEO removed for dumb reasons
this somewhat led to a bit of an inflated ego with regards to reddit but eventually from what i see he left... at least under that username anyway.
I have screenshots somewhere, but it basically said if I continued to abuse the report feature my account would be banned.
Reddit is a publicly traded company and I sincerely doubt the company is taking some organized racket money on the side. But there is some serious conduct issues with admins, and I won't speculate about their motivations.
First and foremost is the destruction of organization of people. If you’ll recall, Reddit used to be an oasis of the internet. Growing up, I was always in awe of the intelligent and impactful discussions that would occur organically (something Hacker News can’t even rival). Now, it’s slop.
There is real value there that Reddit is offering the government. I’m sure other foreign governments (the globalists) also benefit from a weaker people. So I’d imagine it’s the kind of racket that is so high-up, secretive, and decentralized that there is no real culpability and everyone is aware (except the people).
Of course, nobody can view my profile anymore anyway (I'm waiting on appeal), but on my account, only posts from the last 6 years have the "Sorry this post was removed by reddit filters" message.
> My test account (5 years old!) got banned immediately, and all of its post history got wiped too. RIP
I want to know the real string of the event I ever want to delete my account and content. This would be much faster than using a browser script to manually delete.
In case you didn't realise, a massive portion of content on reddit now is LLMs.
I'm not a massive fan of how reddit has played certain hands in the last 5 years or so, but I do hope they win the war on dead internet theory.
In the past, those post removals didn't even exist in the moderation log, so perhaps a reason could give me a clue... On the other hand, I'm taking a kind of emotional damage just remembering.
The appeals process exists to fill a checkbox that says there must be an appeals process, not to actually unban anyone.
However everything remained in the nether-realm and the appeals page claimed that my account was normal, and that therefore couldn't be used
No response from any support ticket either. They still advertised me IPO opportunities though...
banned_by true is more accurate to say "admin or automatic". in "admin mode," you can see these although not sure the UX for these nowadays now that it is spewing a gazillion lines of text into them).
Anti-Evil Operations removals (nee Trust & Safety) are generally human(-assisted) actions (although these actions can be applied en masse). there's some more information nowadays in the API which was really nice. it also helped because people stopped blaming "the mods" for removals when the spam filter slopped all over the place. this was also annoying because previously you had to previously guess from the API how it was removed even if you were a mod.
the 3 ways to remove a post/comment (i.e. in reply to: train_spam):
- remove not spam: removes it but doesn't train the spam filter, obvious
- spam: removes it and trains the spam filter, obvious
- confirm spam: only happens when you remove after removing for any reason, *does not* train the spam filter
- reinforce spam: trains the spam filter even if the spam filter already caught it. *does* train the spam filter. you can do this by doing `action: spam` in automod. not sure if there have been any more in the last few years
also you can tell the legacy of "removals", back in the day stories were "banned" instead of "removed" by moderators and administrators.
also also also... you can see a lot of the stuff from this article in the `approved_by` side of it as well. if you hover over a checkmark of someone who has been unshadowbanned, you'll see it says "approved by Reddit (shadowban removed)"
if an admin manually unspams someones stuff (say someone who got accidentally shadowbanned and got hit with an overzealous spam filter multiple times >.>), it'll say "approved by <username> (all)". there are some consequences to this. it approves stuff that has been "filtered" (as AutoMod filtering is a weird hack where it removes something but keeps in the modqueue).
> spammit
i believe this is the thing that is "pretty similar to a naive Bayesian classifier"[1][2] that reddit used. /u/Deimorz iirc was a reddit dev at the time and it was somewhat public info. i say somewhat because you kinda had to be both interested in the this and probably be around the metasphere
iirc from some other comments i pieced together there are also per-subreddit spam filters. in the olden days sometimes they'd get way out of whack and you could ask an admin to reset it for you... or something idk
> em
guessing em in this case btw refers to /u/hueypriest, who was reddit's GM at the time
> would’ve been catastrophic for Reddit’s spam issues
the thing that surprised me at the time was just how bad reddit's spam filtering is. i did a small little thing at the time where i'd just look at stuff following some basic spam filtering rules (like stuff you'd probably get out of an artisinal spamassassin ruleset) and even that deluge was amazing to see.
like the ML stuff is cool and all but seriously 90% of this could probably still be solved with some basic rules. the profile hiding stuff didn't help either but that was way after my time.
[1]: https://reddit.com/r/TheoryOfReddit/comments/10ko5h/comment/... (2012)
[2]: https://www.reddit.com/r/modnews/comments/6bj5de/state_of_sp...
The fact that they change the timestamp is also very stupid (yes you can hover and still return the datestamp, but this is by definition a dark pattern). These posts should preserve the timestamp vs masking it and even should be flagged as [Second Chance] in the title imo.
How does it work?
I'm open to suggestions of how to do it better! But you also need to consider the cost of adding explicit details to the UI. If we did that every time something like this came up, HN would have become an unreadable mess a long time ago.