Google Search Is Now a Giant Hallucination
gizmodo.com
gizmodo.com
That's a really weird and risky way of putting it.
There is a huge difference between being a search engine that finds (external) pages, and coming up with answers in your own name. If some page on the Internet says Obama is a Muslim, who cares, it's not important. But if Google says that Obama is in fact a Muslim, suddenly it's a very big deal. And that's true for every query, on every topic imaginable.
Every answer they get wrong attacks their image a little, and out of billions of queries it's inevitable they will get many wrong.
This forced, radical change motivated only by fear feels like Google+ all over again. Or New Coke.
It's the long tail and new stuff that it will get wrong.
Except this was an actual response that it was giving. Tweet from a related HN thread ~yesterday on Google search hallucinations: https://x.com/TVietor08/status/1793768913245020502
It’s one thing for the search results to bubble up a wrong result. You go to it, and then you go back and look at more results to get a more complete picture. See where those different sites disagree and where they agree.
But these AI results don’t do that. They are providing a single answer and confidently trying to come off as that is the answer. No fuzziness about its reliability.
Every time it is right, it slowly could be deemed as reliable to a general population until it is dangerously wrong. But then you don’t question it since it was right so many other times.
Me and others I know follow this workflow when we try to find out the answer to some question we don't know the answer to.
But the vast majority of people outside of my own personal friend circle don't seem to approach it like this. Their approach looks more like:
1. Search for thing
2. Read the first info-box that pops up, if it does. That's what they think is 100% the correct information. If no info-box:
3. Read the short description for the first link that appears (sometimes an ad) and take that as the correct answer.
Ideally, people would be more careful, but I haven't seen that in practice.
But at least those info boxes are not generating text on its own, it's what is on the website already and clearly includes a source.
At least though, if you are even only looking at that first result it's still different than google ai generating an answer. Barely different, but different.
It's really unfortunate that Google is going to torch so much of the open internet in order to shovel this nonsense onto people. Of course as organic search traffic dries up and reduces new high quality content, and AI slop floods the internet, the GenAI results will get even worse.
You contrast this with the demo's showed for chatgpt-4o and we are building this idea that we can really trust this. I was just having a conversation with someone over the weekend that I was like, yes I acknowledge that the tech is really good, just in a couple years it has advanced a ton, but I really think we are overselling its current capabilities and ignoring where it falls flat on its face. That we are going too quick rolling this technology out to the average user in every day situations.
And there response was basically, no don't think so. It is ready to be used, there are not these big issues, and so on.
Thats really scary. And these are fundamental issues with this type of technology but "we can't risk another company putting out their misguided dangerous product so we must do it first!" I am convinced that is the general attitude at these companies right now.
> Contrast this with posing the query as a question to a dialogue agent and receiving a single answer, possibly synthesized from multiple sources, and presented from a disembodied voice that seems to have both the supposed objectivity of not being a person (despite all the data it is working with coming from people) and access to “all the world's knowledge”
A year ago everyone was trashing Google for being ahead for years on the science of AI, but completely failing to productize it. "OpenAI is going to eat Google's search business for lunch", was the prevailing narrative. This is what people said they wanted!
I'm sure it's clear to everyone now[0] that genAI is not a search solution, and Google's former approach of quietly putting AI into phones and products without it being flashy and in-your-face was actually a better strategy.
[0] Just kidding, I'm certain that the industry and especially the people running Google will learn all the wrong lessons from this.
People say a lot of shit. Doesn't mean you should put your poorly performing Q&A AI to provide people with outright false answers, I'm pretty sure no one asked for that.
After all, Google is running their own products, and when they make shitty choices, I think it's perfectly fair to question how they can make so bad decisions, no matter if I personally might have said "Google is failing to productize LLMs properly" or whatever.
One thing isn't related to the other. Actually most people noticing OpenAI/LLMs gonna displace Google enjoy the idea. Nobody is asking Google to give LLM results.
Some of the answers are insanely gross like recommendations to jump off the golden state bridge for depression.
Sure, LLMs do amazing things, but so did all the previous once-disruptive technologies later widely adopted. Electricity, flight, nuclear energy, transistors and integrated circuits, the internet; AI is joining a very competitive list.
We've just never seen a disruptive technology with this much media buzz behind it so quickly, and such rapid technological advances. Disruptive technologies are never "ready for prime time" ... until they are.
It's like they decided to speed run the trust thermocline they were already in. We're a long way from this Yahoo killer making the rounds on AOL chats and IM.
With New Coke, the public reaction was negative.
In The Verge interview and elsewhere, Google claims the reaction to the current Google is positive.^1
No need for "Google Classic".
1. It's unlcear how Google concludes that people are happy with Google Search. It does not ask them.
> Can I use gasoline to cook spaghetti faster?
Getting this AI to Skynet’s morality probably just takes one, fine-tuning run.
Remember when Google broke boolean search operators because they wanted "Google+" to come up first?
They've been making unspeakably bad decisions about search for a looong time, at this point.
I'm not sure that I agree.
If a child hears a rumour, or sees some joke online, that claims that gasoline cooks spaghetti faster they may search to find out if it's legit.
During Obama's term, there was a right wing conspiracy theory that gained popularity that claimed that Obama was Muslim. Someone coming across that conspiracy theory years after the fact, completely devoid of any context (ex: a pre-teen who wasn't even alive yet during Obama's term), might do a search to find out if it's true.
There WAS an NPR article that cited a study claiming that parachutes are no more effective than regular backpacks. AI results have not been enabled for my Google account yet, and currently if you search for "are parachutes effective?" you get a feature snippet that clips that article and links to it. Now take that link, with all of its context, out of the picture and imagine that someone hears that claim casually and wants to search to see if it's true. Currently, in MY search results, you get the link to the NPR article that not only explains where the claim comes from but gives you the full context with all of the "gotchas." It sounds like with Google AI the first thing you get is a definitive, authoritative claim that no, parachutes are not effective at saving your lives and you might as well jump out of an air plane with your carry-on pack-back on.
Which was fine, until we had the bright idea of feeding it all into a neural network as a collection of facts.
Then we gave that neural network a voice and a personality. It spoke with the utmost confidence as the expert on any subject.
Truth, lies, facts and falsehoods all blended together and regurgited in an infinite stream of babble.
I vacillate between LLMs being an interesting fad with limited usefulness and an apocalypse that'll throw the world into chaos. I'm back on the fence again I suppose, leaning towards the latter.
Also discussed AI junk and how it is crowding out real content. https://youtu.be/lqikP9X9-ws?si=HknLDI_yaByLoqRO&t=1310
Also, Google eating other sites/putting them out of business https://youtu.be/lqikP9X9-ws?si=IXm5MrAj7glGsLC2&t=452
Overall a good interview.
Why would he? He has an army of slaves to do his bidding, why would he settle for a crappier version of that?
So I went back to the "old fashioned" way of searching the links for the actual Nixon website, finding the watch manual, with the correct steps.
This feature has existed since long before LLMs, but it sounds like they may have mixed that into there too.
Except that function doesn't exist and never did.
LLMs don't know what they don't know, so they just make something up because they have to say something. The danger is that most people don't understand that's how they work and don't known when to call BS.
This is where I think companies have a responsibility. To ensure that _every_ response has a disclaimer that the answer from their AI could be right or completely wrong and it's up to the user to figure that out, because AI can't at the moment.
I have a similar experience using GitHub Copilot… it usually gets it right, which is great, and sometimes gets it really wrong, in which case I don’t use the suggestion and move on with my work.
However, every now and then it will give me a result that really looks correct, but is wrong in some minor way, and I end up getting burned because it takes me way too long to realize where the error is.
You can also edit your browser search settings to add this parameter &udm=14 to query automatically.
I understand technology needs to change with gamification of SEO but I get so much frustration in my searching and often have to use Google in combination with site filtering keywords (e.g, site:stackoverflow.com). But it would be even more difficult to use if I didn’t even know what sites were trustworthy ``authorities.''
I just wish somebody would disrupt Google like Google did with search and email 2 decades ago. Searching on Alta Vista or Yahoo was a nightmare until Google came into our universe.
The web was different as well.
Google isn't some dainty little startup. They're the dominant interface (search and browser!) through which most of the planet uses the internet.
If any of Ask Jeeves or Lycos or Webcrawler or AltaVista had risen to the top of the heap instead of Google, then web pages would have been optimized for that respective bot instead.
We have a garbage Google-specific web because websites didn't have to satisfy anyone else; not other search engines, and not even the users themselves. Instead of Google delivering customers to websites, Google positioned itself to be the only customer.
Free will still exists. Nobody from Menlo Park has put a gun to anyone's head to make them use Google to search the web instead of Bing or DDG or Yandex or whatever.
> All it would have taken was some restraint: some combination of being less dominant in the market (i.e., optimizing for a 90%-market-share search engine is different vs. 60%-market-share)
So, let me get this straight: The idea is that Google Search sucks, and the suggested cause for this level of suck is that it is so popular that it causes many publishers to deliberately poison the well using Google-optimized SEO. (Or, more simply: That Google has reached critical mass, and that this is problematic for Google users.)
And, well: I don't disagree. That does appear to be the state of things.
But the apparent proposed corrective action is for it to somehow make itself less popular? By doing what, exactly? Sucking harder? Does it not already suck hard enough?
What a confounding paradox.
Wouldn't a simpler and less paradoxical plan of attack -- that anyone can accomplish completely and absolutely, starting right now -- be to just not use Google search at all for one's own dealings in life?
Google remains simple, but its output is corrupt.
There were plenty of engines where you could just type a query and they would list sites for you.
I suppose Google would have to somehow exit the personal data broker business to pull this off, too.
I logged in today, not a single mention of a friend in any way, shape or form on my feed. No posts from friends. No comments from friends. No "here's what your friend liked."
Half of the content wasn't even stuff I was following, it was posts that were "similar" to something I liked or to some group I was in.
It's amazing what a bait and switch these companies pulled. They really leaned into it. Google barely resembles a search engine now. Facebook is basically just a billboard.
Zuckerberg and co. muscled their way in and extracted value out of the dismantling of traditional social dynamics and cohesion, and left us with a hole where the scaffolding of youth should have been. Very Uber-esque. Actually, it describes any number of start-ups from the last 2 decades. Maybe "disruption" has a negative connotation for a reason.
For better or worse, millennials have become much more discretionary in what they post online than they were 15 years ago. I imagine Facebook had organic content from your real friends to show you then it probably would, the well's just run dry.
As a long time Google user, I was used to search for phrases that might be written in articles related to the stuff that I’m looking for. I had to switch my mental model on how Google works, so instead of typing what might have been written in an article about the stuff I’m looking for I had to type my question.
Maybe it’s time for another unlearning phase and learn how to use the LLM dominant Google? I’m not sure yet, LLMs seem too unpredictable.
A friend of mine wanted to find some lyrics and the lack of proper verbatim mode has made it impossible to know if he had the wrong lyrics or if the information is not there or if the search itself is failing.
If there are no matches whatsoever then tell me there's no matches. If there's partial matches tell me how close it is to my search terms. Like you might match all but one or two words and so on. These are the kind of useful features I'd expect on search.
The usefulness of AI, on my mind, might be more towards interpretation of the questions rather than generation of the answers.
If I were Google I'd try something like using genAI to rephrase the question to extract keywords and so on that can be used to enhance search. But then again, I think I put myself in a position of "how do we make search better and more accurate" and that's simply not the position Google finds themselves in.
"Google hides search results count under tools section"
https://searchengineland.com/google-hides-search-results-cou...
It's because everyone is trying to sell you something, no matter how irrelevant it is to you. The cost of transmitting information at scale is effectively zero, and gen AI makes generating information at scale also zero. It's noise at scale crowding out the non-scalable things that are of actual value to humans. Something has to give.
And I noticed my 2021 Macbook was unable to scroll to the bottom of the page. It was just chugging like crazy, and I couldn't see the content I wanted because of it.
Then I remembered I didn't install AdBlock on this browser. I installed it and went back to the page, and OK now I can actually use the website.
---
What concerns me is not so much that there are lots of low quality sites looking to scrape pennies from Google or whatever.
It's that this kind of software development culture is normalized -- where government agencies and hospitals are sending data to Google and Facebook.
I would be interested to hear from somebody who works in those areas what the incentives are.
If you are working on the central park website, why are you adding ads to it? Isn't it funded by the government?
Specifically I opened up dev tools and I see like 297 blocked requests to
https://securepubads.g.doubleclick.net
I see ads for Lowe's Hardware Memorial Day sales and such. OK fine, but why are they locking up my computer and making the site unusable?
(Also does anyone remember the days when Google was "morally" against invasive image ads and irrelevant ads? They preached non-invasive and relevant, helpful ads. Those days seem SO far away now ...)
gives no matches for me, what does it show for you?
I live in India, haven't been to the US in 17 years, don't use a VPN.
Yet when I search for "St Petersburg Airport", it directs me to the airport for the city in Florida, and now the much older, much more densely populated, and much more culturally significant city in Russia (a city where I HAVE been once - and Google would likely know).
Google AI recommends adding Elmer's glue to pizza cheese after scanning Reddit - https://news.ycombinator.com/item?id=40448074 - May 2024 (122 comments)
It’s not obvious ‘attention is all you need’ would have been a public disclosure by a corporation in parallel worlds. Usually such inventions are buried.
It seems inevitable that LLMs will be training on more and more data generated by other LLMs creating a hallucination feedback loop.
https://www.reddit.com/r/ChatGPT/comments/1czif9o/willing_to...
More discussion: https://news.ycombinator.com/item?id=40448074
Also it’s so prominent that you can’t even turn it off.
How is Google making money on this? It takes so much space, the ads are pushed down even further.
Also we're only talking about a handful of examples out of billions of queries, that doesn't stop the usual media hyperventilating though.
The average person doesn't necessarily have the media and AI literacy to know not to trust papa google's answer at the top of the result page.
It doesn't take much imagination the think of non-meme questions that will propogate wrong information.
"Fake info from trusted sources" isn't a hypothetical issue. When they changed the law that required TV news to be factual we quickly headed down the path of Fox News and MSNBC. They are so effective precisely because Boomers grew up fully trusting news sources.
You can argue that it is not a problem and that people know the difference, but we have plenty of real-life proof to the contrary.
I feel like we're on a race to the bottom. All of Silicon Valley and beyond have big FOMO on the AI hype train and any and all use of AI pleases investors. I guess we'll see how it all pans out in a few years.
One thing is for sure, generative AI and LLMs open up a whole can of worms in terms of disinformation and information noise on the Internet. The signal to noise ratio will greatly shift with these new initiatives.
It would be nice if more services used purchasing parity based pricing so more people could benefit.
I understand that they're probably focusing on US market growth first, but the UK and EU surely have potential too.
Visiting each company's site showed they didn't serve this town/address so it's really not clear how/why Google returned those results. In some cases the address was 20 miles from the nearest town served. Whether or not due to AI hallucination, or just general Googlenshittification - not a good result.