Publishing AI Slop Is a Choice
daringfireball.net
daringfireball.net
Already now, a lot of users will associate Google AI Overviews with nonsense; eat rocks, cook spaghetti with gasoline. Google also showed how easily these AI Overviews can be manipulated, since they use RAG.
I think the Slopception is only going to get worse. Slop in. Slop out.
And also, Google Research just happened to be sitting on AGREE[0]:
> a learning-based framework that enables LLMs to provide accurate citations in their responses, making them more reliable and increasing user trust
Which was published yesterday.
I tried to submit it[1] but it didn’t get much love.
[0]: https://research.google/blog/effective-large-language-model-...
We are in the rough and tumble time of new tech. As an analogy, trains didn't always have the distance between their tracks as uniform (called gauge). It strikes me that we are in a similar time of invention. If you have ever read old patents (my experience being older pop book like "Strange Stories, Amazing Facts" and random blogposts) you'll see the truly bizzare.
It would be regrettable if the result of rolling out this type of product too soon has the effect of eroding trust in AI-based search, and delaying people from engaging with it when it actually is ready.
nowhere in that sentence did you mention AI, though. If you take AI out of the picture, isn’t it equally problematic if they had some other mechanism, say some kind of analytical/stochastic mechanism, that selected data retrieved from actual results, and cited those?
Because, that’s what they’re doing. Google isn’t using an AI. They’re using a stochastic textual generator driven by results of the search. Sometimes it spits out weird and incorrect answers, especially when high-ranking sites suddenly spoof malicious results - that is a general problem with any attempt to pull data out of the results automatically, including the lesser variants of this approach they’ve deployed for decades.
The problem is (a) google putting out a poor product, similar to the pre-google era of search engines being really bad, and (b) high-ranking sites engaging in malicious and disruptive behavior, for which they probably should be de-ranked or blacklisted.
But you cannot gave this one both ways, it can’t be an “AI” when you want to be an maximalist and “just a textual stochastic algorithm” when you want to minimize and downplay. By the latter terms Google isn’t even using AI here, just a stochastic algorithm, driven by the search results. It’s just a bad one. So you’re arguing against an approach that doesn’t exist anyway.
Again: why specifically is google answer problematic, given that it’s been rolled out for like 15+ years now? The new algorithm behind it just kinda sucks, but the feature isn’t new. Citing and summarizing results has been a thing for a long time, and naturally it’s not always correct/accurate.
This isn’t the first time people have been upset over it either - newspapers got mad about this summarization so they got a law passed in Canada which outlawed it.
https://www.reuters.com/technology/canada-news-industry-body...
But, like Google's garbage, the user experience for shipping or riding was abysmal.
No regulation, rapid development, no protective measures even for the very basic attack vectors.
A lot of tech-minded folks who actively follow this stuff on Twitter et al make this association, but the average person was already defaulting to reading the first (likely incorrect) hit on Google and moving on with their day. I'm not sure they'll notice a dip in quality until it directly contradicts something they know about confidently.
Most of the world is primed to Google and Bing isn’t any better. Bing is doing the same shenanigans.
A new player could beat them at their Game, but Google has a huge ML team and millions of TPUs. They could either buy them or quickly copy them.
Google did this to defend against OpenAI and Bing.
The big question is whether their bottom line will suffer. Google like every other big corp is beholden to their ever growing market cap. A few negative quarters and they face the fire.
I've never looked something up in an encyclopedia. I've never visited a library in pursuit of a research goal. There's no "going back to how things were before" for me - there was no "before"!
Life will go on, but the demise of search engines is quite a terrifying prospect for me. Generative AI is a search engine killer but not in any positive sense.
It's a tough nut to crack.
Google, WhatsApp, Facebook, Tinder.
As if those services are the only options. No they are not.
The world isn't one dimensional nor binary. Everything is a multidimensional spectrum.
I don't use Google search and I don't have to. There are reasonable alternatives like DuckDuckGo and Brave Search. They aren't as good as Google was in the past, but there is also Wikipedia, Reddit, GitHub, Phind, StackExchange, Discord, IRC, forums, newsgroups, Claude, ChatGPT, LLama, and what not.
Sure, it would be awesome to get all the answers from a single input field, but honestly, even though Google was better in the past, this has never been the case for any sufficiently complex question.
Speaking as someone who was born long before the public Internet: you can never go back to how things were before. Similarly, you can never really stay with the way things are now. The world changes. Even if you stick to the ways of your lifetime, or go back to the ways prior to your lifetime, the ways have change simply because the context has changed.
I'm not saying that we should jump the AI bandwagon. I'm just saying that we need to recognize the world is in constant flux.
I simply don’t believe you. The human mind can only take in information so fast.
Does that include Wikipedia? I mean, the problem with encyclopedias is that they were small (despite physically taking up entire bookcases). I didn't use them much either, even back in prehistoric times when the internet wasn't invented. They tended to disappoint by giving a crappy, superficial overview of something that was only vaguely similar to the thing you wanted to know about.
But I use wikipedia instead of doing an internet search a lot of the time. Or sometimes archive.org, for things probably in books. Cut out the middleman, go directly to where you know the information is.
I now like the fact that they, as books, are frozen in time. And that their knowledge is well packaged and consistent. Unless you're a researcher, you don't need the latest. And as a child, the clear writings and the great illustrations was key to internalize the knowledge.
Maybe we need a need a new paradigm for search ranking. Voting, experts, introduce organic feedback back into the loop ?
With the March updates I thought there was a chance that Google could pull it off, but now I am leaning towards it's all doomed. Personally I wish I never got into building websites when I have to compete with ugly wordpresses from SEO companies in India taking my #1 spot for English content.
I wish I could get angry about the forbes articles in my niches but at least that's written ostensibly by a professional-- when you start getting beaten by center-aligned text from South Asia you know there was -never- a chance for quality to win.
Wikipedia won over competitors because it has an anti-business-model. Only the passionate write for it, which paradoxically ensures quality. My favorite content on the internet is always non-profit, sincere stuff. Like this comment thread :)
That's a scary prospect for creators that live off the internet, but if the ecosystem is already doomed it may be the last option.
Wikipedia won over competitors because Jimmy Wales honed his marketing tactics from years of being a pornographer. Not because of any inherent strength of the wikipedia platform. In fact those first articles on wikipedia were all horrible and far worse than the other online encyclopedias if you remember. Wikipedia's success was mostly due to good timing, good governance and great SEO.
And btw it is hysterical some think HN is non-profit. Sure, perhaps technically. But this is exactly why I think we are doomed. If even the self-proclaimed guardians of the internet (ala the IT Crowd) cannot grasp the simple concept that HN is as astroturfed (if not more) than reddit there is simply no hope for the internet as a whole to understand these concepts. So long and thanks for all the phishing attempts.
AI could be something that helps prevent google bombing. AI could be something that lets them serve up one or two ads per page and get a similar number of click-throughs.
Google's lunch is spread out on a picnic blanket, just waiting for someone to come along and eat it.
The big question is: can it get better fast enough to save search? I think the era of search is ending, and it doesn’t look like LLMs are a good replacement.
I've stopped using google search entirely in favour of Phind and OpenAI.
>There’s no reason Google had to enable this feature now.
Given that Bing beat them to the punch I'd say they're already late
If google had any sense they'd just buy them and use them as their new front end.
The users and even the employees took a back seat to making Wall St. happy.
It was ever so, and so only the naive put their faith in mega corporations. Google was extremely useful for about 10 years from around 2002 through about 2012, but even then they were conspiring with Apple to suppress wages for top web browser and other programmers, and they were abusing their overwhelming search monopoly to gain illegal advantage in email, web browsers, and office productivity apps.
Fuck them. They were always evil, but the trade off was better back then when their evil gave us useful stuff and not AI slop. For over a decade that trade off hasn't been worth it.
Once enough of us realized that Google search results are usually just paid ads, what's the point anymore?
It's insane, and "AI" isn't going to fix any of it.
The rot in our tech economy is astonishing.
I recently need to find some details of hardware implementations for a feature. I used Google to find specs for this hardware. The specs were hard to read or didn't explain things as I needed. So I asked ChatGPT and it gave me what I needed. It first regurgitated info from the specs but I was able to ask it to explain pieces in more detail and it was great! At what point do I just stop using Google search to find an answer?
Or course AI answers have their issues. ATM I wouldn't search for reviews of some new video game or movie on ChatGPT. I also wouldn't search for product reviews. Even though it will happily spew out recommendations they'll be old at best. Maybe that will get fixed but it will have all the same issues as search and more. People will write their SEO type techniques to try to get the AI to surface their product just like they try to get them to the top of search results. I guess we'll need a new name. ARO (AI Results Optimization). Searching for products at least I get the illusion of lots of various opinions. With ChatGPT I just get the bot's one opinion.
Anyway, the point is, Google has to do something. Too late and everyone will switch. Which pushes them into too soon and hence the issues we're seeing.
IMHO, fully cosigning that google search has gotten worse and the AI shit is only going to make that even worse, using ChatGPT as a replacement feels crazy to me. Like yeah, Google gets it wrong sometimes, sure, but ChatGPT just makes shit up. Is a dumbass better or worse than a confident liar? Guess that's up to the user to decide.
Close enough is good enough. It doesn't need to be perfect, it needs to just not be catastrophically wrong.
I really want to give this a pass but the worst thing is that you can't really see where it turns from correct to incorrect, like you would usually pick up with a human teacher or peer when you start to get a feeling they really don't know what they're talking about. ChatGPT will gladly keep hallucinating references and reasonings that sound superficially plausible until you look them up, because that's what it's trained to output..
If they could just find a way to cut it off when it starts to become too unsure, it would be a big improvement.
I recall a time when a popular joke amongst tech people was that a tech enthusiast was someone who had every new smart home accessory and every new gadget and used them all, and a tech engineer was someone who had nothing more advanced in his house than a laser printer and he kept a gun in the same room in case it ever made a noise he didn't recognize.
I guess we're just more enthusiastic than we used to be. I like my smart switches, but I don't like the notion of all human knowledge being only accessible to me through the filter of a word generator. If that makes me a Luddite, then Luddite I am.
But was it accurate ? Chat GPT just provided you with the statistically most likely specs. If the actual doc for this hardware is lacking, this is 100% hallucinations.
My experience is that ChatGPT is simply bullshitting you (sometime accurately), Google is drowning the info in 20 links (where they try to make you buy "Hardware you're searching for"), and you have to go to DuckDuckGo to find the thing you're actually searching for.
That gave me a laugh. I don't trust Google. Their entire purpose for many years now has been simply to enable SEO trash in such a way that users see as many ads as possible while providing a minimum amount of useful information so that users are not totally frustrated.
If Google could legally sell your organs, they would. They are that nefarious, and deserve all the hate they get. AI is just the next step. Frankly, if Google went bankrupt today, I think civilization would benefit immensely.
I’ve developed an instinct for when to ignore them. Typically it is when “edge information” is involved: facts / data that are only very sparsely published in time / web space, perhaps by only one or a few web authors, perhaps only very recently.
As an AI developer for over a decade, I feel like if we’re calling this “slop”, I fear what these authors think of human error. I’m also doubtful that these people who are damning this project have ever in fact innovated or developed anything significant from scratch themselves.
Makes me want to start a company called Slop.
There is precedent. You can start a modern lifestyle brand like Goop.
As a defender of the new LLM-thingies, do you think they're doing a reasonable job of promoting AI-output literacy? I think it's their job to do so when they are the ones generating the content, whereas general media-literacy was not really their problem when Google was just a directory for the web.
Google has been pushing to "answer questions" over return web sources since at least 2010 (when their public messaging changed), so that is nothing new. But it is a seismic, human difference between aggregated results from trusted web sources and reformatted comments from Reddit posts.
AI is great for targetted tasks and even for general web usage, with caveats. I would love to see something like an accuracy/trust score associated with the results or at a minimum a beta flag with a "hide results like this" option.
These are all basic UI procedures any reasonable person would make, but Google has been riding high on hubris for a while now (zero human support, just see the Google Cloud issue on the frontpage...) and this won't change until the stock starts getting slammed, which seems imminent.
Calling it slop at least acknowledges that the AI isn't trying to lie to you. It's not merely making a error, either. It just doesn't care either way.