AI Slop Is Flooding Medium
wired.com
wired.com
I have noticed that someone immediately calls out pretty much every bit of content as AI generated, whether it obviously is or not. I guess this is a best case scenario.
I do wonder what will spring up next to allow high quality content to be found.
You know which travel advice I prefer? The one from a small, local blogger that loves their city. You know where I find my music? At house parties. More and more, I also crawl the profiles of individual users for content that I might like.
The cosy web is my refuge from scale and its consequences.
And that lack of scale means it has little mass monetization potential. It's actively hostile to weirdos just trying to make a buck. I don't get randos from who-knows-where trying to hawk their wares in my private group chats, so those recommendations actually carry weight.
We're between a rock and a hard place with the internet's collective signal to noise ratio plummeting and no easy way to fix it.
I really don't see how these people making the ebooks are making any money on these books either since they cost 99 cents and they pay for advertising. Maybe they're using llama2 idk lmao, either way I've seen a ton of embarrassing titles pop up on there and don't want people thinking I'm interested in that slop.
SEO was annoying but there was still value in the results. Now, as soon as I detect or start to believe some result is AI bullshit I immediately close the tab out.
It seems like the leaders of all of these online platforms are going to just bury their heads in the sand and pretend that this massive problem isn't one.
This AI trend is absolutely going to destroy many (most?) of these companies. For better or worse, the new age of the internet is already here.
I have noticed a lot of these articles that just make no sense at all starting to appear in my weekly email to the point that I probably won't renew next year if it stays this way, so it isn't just "nobody reading them".
I'm not sure of the exact machinations, but the times that my posts have been boosted, it's always been by a publication on Medium where a human editor has read my content and 1) asked that I submit to their publication and 2) submitted to Medium for boosting.
Seems like this type of filter -- as long as they actually keep humans in the loop -- can potentially work to keep high quality content more visible.
Of course, even mentioning the f-word is forbidden, so...
* humans creating things
* algorithms emulating (humans creating things)
* human-in-the-loop guiding (algorithms emulating (humans creating things))
* ai that mimics the (human-in-the-loop guiding (algorithms emulating (humans creating things)))
* what next???
Why is that not sufficient for site operators to separate "AI Slop" from other articles?
If existing sites choose to tolerate "AI Slop" (why doesn't HN have this problem?), does that create a market opportunity for competitors with better filters?
> why doesn't HN have this problem?
HN comments aren’t easily monetizable.
And it's only going to get worse. I don't know what the limits of LLMs are, even before we start discussing future architectures and capabilities, but it certainly seems to me that if they just scale a bit farther on their current trajectory it is going to be very hard to detect AIs simply by the content.
One of the adaptations I've adopted is that if your content looks like it was generated from a simple prompt, no matter how long the resulting content, I just toss it. "Write me a beginner's tutorial for Python" isn't going to produce a document that the internet needs another copy of or that social media sites need to spend time linking to. Unfortunately that means even if the "tutorial" was human-generated, it's probably getting the axe too, unless it has a very interesting or unusual approach, but it isn't clear to me that that is actually a large problem.
Why isn't it? As far as I can remember, all the forums had lot of outright spam, even forums which are orders of magnitude less popular than HN. You can even see spam in usenet. Reddit/Discord is full of it. I wonder are dang/moderators the only reason for this?
I've seen a few comments on HN that looked likely to have been generated by LLMs, but they tend not to get any replies or up-votes so there's not much incentive for people to keep on trying to post them.
LLM-generated content is not distinguishable from human content by any simple rule, so you lose the ability to police the platform that way.
Tell that to Elon Musk. He seems to unable or unwilling to stop a certain type of fake female bot spam on X that has extremely simple to spot and unique patterns.
It probably drives engagement, so there's no need to interfere. Perhaps it should be encouraged even..
---
Please refine: To be honest, I do use GenAI to help me format my comments including this one. I do, however put in substantial input outlining my thought in detail, so I use it more like a glorified grammar checker. It helps me a lot as I'm not a good writer (as being non-native, just sufficient to be able to thrive in business environment, just not a good writer. It helps me tremendously to get my thought across efficiently.
Writing is not just about syntax, but also about communicating who you are. This is your true voice.
Some roughness in my recent writing comes from being on mobile, which makes reviewing even harder! My writing has improved, though, even when I’m working without AI, likely from exposure to better writing flows. I’d say I used to rely more on edits (20% input to 80% edits), but now it’s closer to 75% input and 25% edits. Hopefully, I can refine it even further.
"Consider the following message. Distill the main thrust of the message, and list any implied statements made within. Be particularly sensitive to any negative connotations, and summarize the writer's feelings towards the reader:"
In the case of blatantly false images like on Facebook, they get a surprising amount of organic engagement: https://x.com/FacebookAIslop/status/1806416249259258189
Because it's already filled with russian bots.
just absolutely embarrassing for us all.
We missed the boat long ago on dismantling or neutering the ad industry before it did massive irreparable damage. Whatever we do to fix this mess is still going to leave prominent scars for generations.
It was always possible to cocoon yourself into news and opinion you agreed with.
AI slop does seem a little different, if only because it's machine generated and the sheer amount of it would not have been possible when rubbish was human generated and the gatekeepers of publishing were newspaper editors.
Like for example a TikTok, but all the videos are generated by the platform itself using likes/skips to drive its algorithm to generate more addicting content. That seems like the logical conclusion of where all of this is going.
They can then drop the creator side of the market and a lot of moderation issues that come with it, while gaining full ownership, control, and monetization of the underlying content.
https://theglobalistperspective.substack.com/p/the-rise-of-i...
The CEO accused me of trying to extort him because I sent a short email with our findings prior to the publication of this article.
He also casts doubt on the efficacy of AI detection (fair) but I think our AI detection model is orders of magnitude more accurate than the others and I stand by its predictions.
I noticed something, you achieved near 100% accuracy on most domains in every domain but scientific which made me wonder, how much is that could be due to how "strict" and "profrssional" these papers could be, or maybe how a slightly disproportionate number of the training data for these LLM could be from science based articles and papers, as they are generally viewed to be "high quality"
Interesting read either way, best of luck on your project (:
More specifically, humans who feel the need to make a quick buck.
Actually, the problem isn’t those humans, it’s the societal structure creating incentives and pressures that lead to a large percentage of humans feeling like they need to make a quick buck. Whether that quick buck be for cost of living or for social status gain.
> The 12-year-old publishing platform
basically
Amongst the others things that walled gardens like Facebook, Instagram, Twitter, and TikTok have taken is the open availability of social signals to rank content. Previously Google used links to build some trust metric, but almost no real person creates public links anymore.
Sexual content
Violent or repulsive content
Hateful or abusive content
Harassment or bullying
Harmful or dangerous acts
Misinformation
Child abuse
Promotes terrorism
Spam or misleading
Legal issue
Captions issue