Wired has removed "How Google alters search queries" story
wired.com
wired.com
The author appears to have gotten the slide exactly backwards. She said the slide showed a query of “children’s clothing” that Google rewrites to be a “Nikolia kidswear” query so that it can sell more ads. But in reality, the slide is describing a fuzzy keyword matching system that takes a query of “Nikolia kidswear” and allows it to match ads with “children's clothing” keywords.
I’m surprised WIRED allowed such an obviously incorrect article to be published in the first place, particularly when it was by a known partisan (the article discloses that the author is a former Duck Duck Go executive with an obvious bias).
Wouldn't it be better to add a correction at the top of the article instead of deleting it?
Deleting it I'd think would just fuel (current or budding) conspiracy theorists, as they can point to something that existed but got removed, rather than something that got corrected.
But in the end, it's probably a loosing battle anyway...
I generally find that anyone expecting error-free perfection in any subjective field from an entire industry can be safely dismissed as a complete idiot. Intelligent people look at the processes that produce outcomes, not the outcomes.
(In this case, the process that produced the outcome was a former DDG employee mailing in an oped of incredibly questionable quality that criticized a DDG competitor, that was eventually pulled. Wired should look into reading op-eds before publishing them.)
Admitting it was trash is proof of inaccurate reporting, whether they remove TFA or not
As for the conspiracy folks, they'll think what they want, regardless of evidence presented to them. That's the whole point.
The way to solve it is for "investigators" to get access to said source code and give us an honest assessment. Get someone neutral, get NDAs in place, have them look at the source in an air-gapped clean room, anonymize DB access, etc etc.
Instead we go around in circles. Journalism at the end of the day then becomes "we got some info that points to X, but we have no proof, and we won't look for it. We'll just wait till someone, somewhere, at some point, who knows when, finds some proof and puts it out in the public, only then will we say it's conclusive. Until then, you all go nuts, especially the conspiracy theorists. Don't worry, we'll ride this gravy train and report on those nut jobs too!".
It's a mess, and I'm not at all surprised that conspiracy-theorists have a field day with it.
Hey, let's get a bounty up so that a whistle-blower from Google comes forward. Maybe Mozilla foundation can take some of that sweet sweet endowment they use for "Social Justice" and pony up to fund this investigative effort. After all they claim to be fighting for a free internet. I'd say this is a much better use of money than paying random DEI consultancies huge speaking fees.
No evidence: obviously the evidence is being suppressed, proves the claim
Evidence against the claim: shills obviously trying to control the narrative, proves the claim
I showed him the IMDb and Wikipedia pages for it and a clip. Without missing a beat - remember they literally had never heard this before and have already decided it’s misinformation planted by the CIA - he said “Well it’s an independent show he put on that happened to be shown on RT.”
I mean what can you say to that?
They should have mentioned 1) what the article got wrong and 2) that the article was written by a former Duck Duck Go executive who should know better.
Making your mistakes suddenly disappear as if they never happened is tempting, but it's going to leave users feeling gaslit, distrustful, or even just misinformed.
“Ain't nobody got time for that”.
The role of journalists is to provide facts, but the business model of media in the internet age is about “creating content” and “engagement” and they don't work well together …
[1] 'The gun went off and the shot hit Mr. Smith's head' versus 'Officer Sloan shot Mr. Smith in the head.' [2]
[2] Both of these are 'just the facts', but one of them blames the gun for going off, the other blames the person who made the gun go off. Passive voice versus active voice, very different presentation of the same facts, both are correct, and both are biased.
Yes, that's why you need press freedom with multiple point of views. But it is orthogonal to the point I was making.
I think you might just be disappointed in humanity, rather than any particular agency or person. Humans come together to get the right answer eventually; not multiple times a day, every day, without fail. Sometimes the software doesn't do what you want. Sometimes that article has a factual error. Eventually we figure it out. It's not the end of the world.
Think about the opposite. Imagine we pickup a bug in a live application and update it citing "our QA standards". This magical word does not absolve us of the fact that we let a bug through to production. And it's implied that we'll do better next time to pick up this kind of bug.
Is that really something you are challenging? Which pieces of popular software you can think of are bug-free? And are you thinking of more than a handful?
And they themselves were responding to someone criticizing media for using "editorial standards" as an excuse when they "retract a release".
And I explained how we are the same and supposedly strive to not release bad content/software, we just don't get to magically absolve fault with an excuse that "it doesn't meet our QA standards" as an analogy to "editorial standards".
Anywho, the whole thing breaks down when we have to beat it with a stick. It's a discussion, assume a charitable interpretation.
They got an opinion piece with an extraordinary claim about something the writer saw in a slide presented at a trial. She wrote the entire article about it creating theories and then how it needs to be stopped. Not to mention, she had previously worked with a competitive search engine. The central piece of fact checking was seeing that slide, that would have been the second question after "Are you sure you saw something like that?". The entire article hinges on that one slide she saw and what she understood from it.
This is not a case of missing a boundary condition. This is missing the central premise. At any level, it would be inexcusable.
PS: In a sense, as the artifacts are becoming public, i think more such confusions would surface in the coming weeks where merely misunderstanding what is on the slide would lead to a lot of rumors. Probably best we have an initial case and some intuition that this can happen.
There's a wide spectrum of how fucked your software can be when you release it without it being a huge deal, like if I push a broken release for my hypothetical build system that jumbles the error messages a little that's probably fine, if I push a broken build of the google dot com landing page that's raising a few more sirens and I really shouldn't make a habit of it, if I deploy critically broken firmware to your pacemaker or your crewed rocket ship I should probably be exiled from the field. I imagine there's a similar spectrum for journalists and we can't really figure out where on the spectrum to the article from the OP should be with just analogies to a different field. Complaining about a trend of publications slipping towards the yolo'ier end of the spectrum doesn't seem on its face hypocritical.
Editorial standards should stop bad stories from being published. Saying "we retracted this story because it doesn't meet our editorial standards" begs the question "why are you publishing things that don't meet your editorial standards in the first place?"
It doesn't take responsibility for or explain the mistakes in the article. It doesn't state that the article had factual errors. It's a frustrating cop-out. I sincerely hope this is a temporary measure while Wired gets a more comprehensive retraction put together.
Publications can simultaneously be praised for “rolling back a bad build” while being criticised for letting that build roll out to prod in the first place.
This would be like the SWE equivalent of pushing to production without code review. It's not an excusable mistake like when a bug gets through.
Maybe there were lots of articles that were never published due to editorial standards, but we only see the ones that slipped through.
They retracted it so shrug
So it's showing "Nikolai kids clothing" being rewritten to "Nikolai kidswear".
1. (because of [kids → children]) ads with keywords “+kids +clothing” would also match searches like “clothing for young child” and “newborn children's clothing”
2. (because of [kids clothing → kidswear]) ads with keywords “+kids +clothing” would also match searches like ”nikolai kidswear” and “kidswear outlet”
3. (because of [clothing → apparel / outlet]) ads with keywords “+kids +clothing” would also match searches like “creative apparel for kids” and “kids outfits”
That is what the slide's title (“Advertisers benefit via closing recall gaps”) refers to: the gaps in recall (matching) are being closed, by being broader.
The WIRED article misunderstood the slide, and was entirely based on the premise that if you searched for “children’s clothing” you'd get results for “NIKOLAI-brand kidswear” which is not true (and would indeed have been “startling”, not to mention obvious, if it were true). In fact, the organic (non-ads) part of the search results in Google are always completely independent of anything in ads, something that the Search team in Google have maintained for several decades as a fundamental principle.
Are you sure? There's an email from the ads team proposing multiple measures to increase the number of search queries so they can reach their target revenue. One of the mentioned measures include "ranking tweaks."
Yours should be the top comment
Why are people not talking about this document more?
“I also don’t want the message to be we’re doing this because the Ads team needs more revenue…but what is the best for Google overall?”
Clearly, the ads team has influence over search to the point of saying more-or-less screw culture and team morale, let’s do what’s best for Google overall which is hitting our quarterly targets.
Ads didn’t get its way here. It doesn’t drive every decision. Especially not with search, but also not with chrome.
reading through the email chain, it seems ads did indeed get its way, and the product was indeed made worse to drive revenue numbers – chrome team was unable to say "no" when pressured by ad team
That email thread you linked is between the Ads and Chrome (not Search) teams, is about the number of search queries (not the results of search queries), and “ranking tweaks” there refers to the ranking that Chrome uses to show the suggestions in the omnibox (address bar). (To get a sense of these “ranking tweaks”, try this experiment in a (new?) Chrome profile with default settings: type "flowers" in the Chrome address bar and don't hit Enter, and look at the suggestions: what mix of search suggestions, entities, and bookmarks/history do you see? Try again with other commercial queries like “insurance” and “mortgage”, and also some less commercial queries like, I don't know, “Minnesota” or “economics”.)
(And FWIW, I think that whole email thread actually shows Google in a “good” light relative to the popular impression here on HN as a company whose every action is some Machiavellian scheme to increase ads revenue: it shows that Chrome actually launched something to production before its negative impact on revenue became a concern, that Ads leads had to work hard to persuade them to either roll back or find some other way to undo the decrease in search query volume, that starting to include search query volume as a launch criterion would be a “cultural shift” for Chrome, etc: that Ads having an influence on Chrome is a rare occurrence.)
>"Thanks Anil (Google Chrome Lead) for pushing your team and being open to this whole line of thinking... We are short REDACTED% queries and are ahread on ads launches so are short REDACTED% vs. plan...The Search team is working together with us (ADS) to accelerate a lunch out of a new mobile layout by the end of May that will be very revenue positive (exact numbers still moving) but that still won't be enough. Our best shot at making the quarter is if we get an injection of at least REDACTED%, ideally REDACTED%, queries ASAP from Chrome... I also don't want the message to be "we're doing this thing because the Ads team needs revenue." That's a very negative message. But my question to all of you is - based on above - what do we think is the best decision for Google overall?
>In that spirit, do we think it's worth reconsidering a rollback? Or are there very scrappy tactical tweaks we can launch with holdback that we know will increase queries? (For example, can we increase vertical space between the search box/icons/feed on new tab to make search more prominent? are there other ranking tweaks we can push out very quickly? Are there other entry points we haven't focused on that we could push on soon?) Just to be clear, the reason I haven't pushed harder on a rollback so far is because I don't want the message to be..."
(source: https://web.archive.org/web/20230919185431/https://www.justi...)
Is wired supposed to be a very accurate news source? I'm not surprised to hear bullshit from the media. This statement implies wired is supposed to be better than normal?
https://www.wired.com/story/google-antitrust-lawsuit-search-...
HN discussion: https://news.ycombinator.com/item?id=37740425 (4 days ago, 119 comments)
Sure would be nice if the material was provided to us too. Or write an article explaining clearly how the original info was misconstrued. Simply deleting it (edit: instead of replacing it with a clarification so anyone who goes to read the wrong info gets the corrected one) really sends a wrong message, both about Wired and Google.
https://twitter.com/searchliaison/status/1709726778170786297 (via https://news.ycombinator.com/item?id=37802302)
Uh, other than being pushed off the visible page, I guess?
Maybe the score is simply highly-correlated with a site that shows many google ads? Think about it, someone that shows Google ads happens to also be very very keen on optimizing everything that affects their rankings in Google searches.
Hey Google, perhaps "displaying many Google ads" should be a negative on page-rank?
Again, your point is one I strongly agree with, but it's also taken out-of-context with respect to the article Google was responding to.
Or, I guess "below the fold"
I wish WIRED had included this context in their retraction statement…
ETA: Okay, weird. The op-ed itself made it sound like she was a current DDG executive, but her LinkedIn profile states that she no longer works there (and is more or less self-employed now). No idea how to interpret that.
In any case, she was DDG's legal counsel and VP of policy for three years.
We don’t know, but it does seem like a reasonable suspicion?
Little DDG would necessarily have to do with it either.
Okay, but where is its correction? If you, as a professional publication that considers itself to publish news, unintentionally spread misinformation, you shouldn't just take down the old links, you should put out a new piece with the correct information (and give it just as much attention, don't hide it away).
The news came out, everybody got mad, and not half the people who read the original article (or headline) will revisit the articles and see the retraction.
Maybe the correction will come later, but I'll stay sceptical of anything Wired has to say until the correction comes out.
Everyone knows Google has been swirling the tank for over a decade for actual web search and AI will likely finish the job. The main vector I look at is how much control does the user have over the results they see? That has been in decline at least since Google+ took away the + operator.
Remember "did you mean" results? Now we get "let's assume you meant" with fewer and fewer ways to force "actually I did mean..." This is the obvious trajectory for any corp that needs to show quarter over quarter growth for 20+ years, at some point you hollow out from the middle until the whole thing collapses.
Everyone knows this? GOOG stock is near an all time high with search being more than 50% of its revenue.
1. GOOGL is the voting share stock ticker, not GOOG. GOOG is nothing more than class c shares.
2. Search as a line item is more than 50%, but search is not searching, searching itself provides not much value, ads are where Google makes its money.
3. A stock being near all time high has no indication that it's a company that is "swirling the tank". Often a stock does not decrease until after a series of bad earnings/negative outlook. But with that said, there is no evidence at the moment that Google is not doing great.
As for why their market share has not degraded as a result, it is likely that the folks currently prosecuting Google for antitrust violations have the best argument for why this might be so.
"It looks like there aren't many great matches for your search"
Good job, Goog!
I remember when it was new, and one of the killer features was the dead-simple "innovation" of having AND searches instead of OR. To this day, I think that any search engine that queried based upon the exact provided search terms would eat their lunch. Probably not, but I would like it.
The modern internet is full of hot garbage.
This seems orthogonal to the suggestion. Presumably the search engine would still filter out noise.
However, even if the other poster means they want the raw internet as well, there's no technical limitations preventing search engines from offering a 'filtered' or 'raw' option. They already offer filters based on what you're looking for (books, news) or to exclude explicit results.
Kagi is probably what you're looking for then. :)
If you assume something ("oh they meant X not Y", then when your assumption is wrong you (Google) just look stupid.
When it acts as if I'm wrong, it's wrong.
Though Google PR is brings up "relevant ads" every time they're criticized for cyberstalking, they seem to care less about relevant search results. They insist on showing me pages upon pages of links that completely disregards my search queries.
Fortunately, for you, there's a link you can click on that will force the query using your precise search terms, on every auto-corrected search page.
You still get what you want, it's just not the default. And it shouldn't be the default.
Sounds like your problem is crappy UI enforced through attempts to minimize production cost at the expense of enabling the use of the human processing medium to effectively integrate with an electronic device.
My error rate typing on mobile skyrocketed with the loss of haptic feedback.
Don't knock ideals. Most of the world has a vested interest in convincing you they aren't possible. They most certainly are.
Google has been steadily increasing its "fuziness" to the point that it considers words like "motorcycle" and "bicycle" synonyms. It's made it more and more difficult to get the results you're looking for.
They even do it to search terms you put in quotes.
Google’s continued erosion of their core user facing product “search” (the real core product is advertising but that’s not the majority of people interacting with google are interacting with google for… and an argument could be made that googles only real product customers care about anymore is YouTube but the quality of that experience fluctuates wildly depending on how stupid they are being any particular day due to asinine policies and abusive relationships where they seem to desperately want to destroy the goodwill of the creators that upload content like it’s some kind of fetish and they just have to know the creators hate them otherwise the job of working at YouTube isn’t satisfying…) is indicative of a complete failure to care about the core competency of the company… and companies that fail to care about their core competency are rotting hulks doomed to die…
It's a computer, that's what they're supposed to do. Precision and determinism is what makes them great.
This is like saying if you had say, a pocket calculator but hit a key at an incidental angle and the calculator then presumed you meant the next number over and gave you that answer instead. It's incorrect - that's not what computers are supposed to do.
> This is like saying if you had say, a pocket calculator but hit a key at an incidental angle and the calculator then presumed you meant the next number over and gave you that answer instead. It's incorrect - that's not what computers are supposed to do.
This deserves wide consideration. (Replying because hidden upvotes can't convey that.)
Optimization exists though, and an interface and search algorithm isn't a simple calculator. Suggesting the correct term when you misspell or mistype something is precision -- it's both identifying the lack of results for your erroneous input, and suggesting the correct input to get the result you're most likely searching for.
That's literally the point of optimization. If Search was still the same as it was in the late 90's, Google wouldn't be able to do half the things it does.
Are you going to make similar gripes about autocomplete, or GPS that reroutes when you fail to make the planned/"correct" turn?
Comparing an intelligent and contextual search interface and result, with simple arithmetic, is a patently false analogy.
That'd be great. The newer half of it is terrible.
> Are you going to make similar gripes about autocomplete, or GPS that reroutes when you fail to make the planned/"correct" turn?
Absolutely valid. I never use autocomplete as it is vapid and incorrect. Also I don't use gps routing because it does this.
The "smart" rotates and "smart" zooms around the screen ignoring my input isn't desired.
These systems presume the user is profoundly, unbelievably stupid and can't, for instance, understand cardinal directions.
It's why when you enter a url like "http://somesite.com.:80/" it will be like "well I think you meant https://somesite.com" and then just ignore all your very explicit protocol and port instructions and whisk you off to an https, even if it's broken and doesn't work or similarly if you explicitly select a subsection of a url that starts at the first character, it will invisibly tact on the protocol to the beginning of your selection to be helpful ... as if the user is helplessly befuddled and perplexed by the protocol syntax.
These aren't optimizations or improvements. They're diffusive and reductive interfaces that disempower the user, they're everywhere now and it's why everything sucks.
Here's what you're advocating for in the physical world - a smart flathead screwdriver that can't be used to pry or wedge anything. In fact, if you try to do that it will have special built in motors and then work against your intentions, wobbling around looking for a flathead screw and then refusing to work if it can't find any.
Presuming the user is a completely incompetent clumsy dumbfuck and ONLY working under that modality is not an improvement. This has somehow become a core design assumption in the SV and it needs to die.
Not that I disagree with your calculator example, but "what computers are supposed to do" changed radically when AlphaGo hit the scene, changed some more with ChatGPT, and will rapidly become a meaningless notion going forward.
We don't have to like it, but we do have to accept it.
You can ask absurd things to gpt and it will try to respond to your absurd request.
For instance, I asked it "What year in the 1900s had the most tuesdays" and it spit back a python program that tried to figure it out. On Google, to compare things, I get "1900s" crossed out and the wikipedia entry for the Ruby Tuesday restaurant chain as the first result, maybe because it's around dinner time.
The difference here is between that and the "guess what I mean based on crude demographic information and popularity" interfaces that ignore the user's clear intentions in favor of gross statistical markers like how google changes the position of images, shopping, maps, etc in the results based on the query presuming your intention based on crude vague guesses or the search systems that seem to only return 50% of what you asked for and the other 50% is simply what it thinks you want to see instead.
Those are not tools, they are broken trash. It'd be like if you had a knife that randomly turned into a spoon or a fork based on what time of day it is and what room you're using it in.
ChatGPT tried to do exactly, precisely what I said regardless of the fact that there is no answer since it's a list of years and not a single one. It's the "you told me to do X and I did exactly X" interface, the kind you get with a good tool.
Too much software is the opposite - basically as if someone walked up to a dinner machine and said they're vegan or kosher and the machine was like "American male. Eats cheeseburgers. Here's cheeseburger"
These modern systems just straight up ignore you in favor of some "big data" approach. It's trash.
Recently I slept through a silenced alarm on a Saturday because my phone decided that people don't want wakeup alarms on Saturday and Sunday without extra configuration. I had set it Friday night for the next day and it was off by default because of some grand assumption. Asinine... This stuff is everywhere.
It still does that now if you put your search terms in quotes.
This has been the trend in so much technology and I think that's why there's such a slow, grinding, complete distrust of Silicon Valley going on. When all this stuff was new and interesting, you did a Google search to find stuff. Sometimes, by law of probability, that stuff was indeed something you wanted to buy. Now, all queries are subject to change so they can put a product that paid them for exposure in front of you, completely irrespective to if that product is relevant to your search or will solve a problem for you. Social media feeds used to be a chronological timeline of things your friends posted; now they are a selection, made by an algorithm you cannot interrogate, of the "most interesting" (judged on metrics you do not set and have no control over) things your friends have posted. Some of them today, some six weeks ago. And, between nearly all of them... is a product that paid for exposure, that is probably at least tangentially related to things you're interested in, but could not be, and more importantly, was not requested.
And, if you make the terrible, awful mistake of browsing any of these sites without an account, prepare for an absolute FIREHOSE of the worst, crummiest, most exploitative, total and complete bottom-of-the-barrel content, the widest possible net designed to ensnare anyone who passes by to watch a marketing grad in his 30's give 500 people in the developing world dental care, or whatever the fuck, set to copyright free music with as many ad placements as they can stuff into the thing.
Increasingly the Internet is not for us, it is certainly not by us, it is simply where you go when you are bored, the only remaining third place that people reliably have access to, and in true free market fashion, it is wall-to-wall exploitation. People selling their bodies because they can't get access to enough money to live, people desperately trying to sell things they've made because time cannot be utilized anymore without a financial benefit if you want to remain solvent, and of course, massive corporations posting billboards large enough to cover the sky, in every direction, every place. A new spot springs up, a new gathering spot that promises to be better, and it gains traction because everyone is so sick of the rest of it, and then in short order once enough people frequent it, the ads go up, the beggars appear, your friends are hidden behind a cylinder of recommended bullshit that surrounds you, and it joins the cavalcade of endless irrelevant nonsense that makes up the spaces you fled, just in time for the next one to pop up nextdoor promising it won't do the same.
God I am sick of the Internet.
Also apologies for the tangent.
The SEO attack surface is functionally infinite for AI, and I'm not sure how any search engine is supposed to be good and also have revenue. The only way to make money is to get people to give it to you, and it's hard to imagine a world in which either a) people will pay for the right to use a search engine or b) companies will pay you to rank them fairly.
I think that's their hypothesis either. I don't think it is a coincidence that we are seeing more lines of ads in search results than ever. They are squeezing the cash cow. I don't think this observation is subjective but could depend on your location, do you observe the same?
I've never seen a set of search results that felt like a different search.
My issue with Google results stems from them being low quality. Like all the stack overflow and GitHub issue clone sites before the real results.
I don't think this is wired getting bought or bullied- I think the author was mistaken and they retracted the story.
Google serving specific ads and optimizing that algorithm to maximize profit is expected imo
They'll certainly ignore search terms, modifiers like quotation marks, and so on. If the result looks like what you meant to search for, that means they're doing their job right I suppose, but they are definitely only using your literal query as a suggestion.
(I did not get to read the original article, I have no idea if this comment of mine is off the mark with respect to the claims of that article)
Disclosure: I work at Google but not on Search.
Since I don't use Google search much anymore, and have my search history turned off anyway, I can't recall any specific examples from my own life. So, I searched and found internet threads like this one:
https://webapps.stackexchange.com/questions/144207/google-fo...
Where the complaint is that searching for "map accuracy" (in quotes) results in pages without that literal string in them, implying that Google is ignoring the quotes to give the user what it thinks is a better list of results.
But, when I try to duplicate this three year old problem, I can't replicate it, at least not on the first page of results.
And when I try to contrive my own example, searching for a literal phrase that would not exist in any web document (an example: "He beheld nonlinear radish-scented vestments") there are 0 results, which is exactly what I'd expect if Google was obeying the quotation marks.
So, this is probably not true anymore, and you've got me doubting my memory now, but I am still fairly certain that it has happened to many on multiple occasions. /shrug
Google search will return results that match the quoted query only through invisible text. So you you might get a result that seems to not contain your text, but if you check the source code or the DOM, the text will be there, hidden.
When checking for equality, search will ignore certain punctuation and HTML. So your text might not be there exactly, but once punctuation is stripped, it's there.
In the time between Google crawling the page and you viewing it, the page may have been edited.
That SO example may be a case of the 2nd reason. For example a page saying
<title>Map Accuracy</title>
<p>Map accuracy is a measure of...
technically contains the text "accuracy map" once you strip out HTML and normalize whitespace and case.It's long been rumored that Google coerces news outlets to publish (or not publish) certain topics lest they find themselves downranked or missing from search queries for a while or get penalized with adsense.
With the removal of an article besmearching Google, I begin to wonder if there's any truth to those rumors.
(my overriding thought as I read it being "I'm sure google does assorted morally dubious things in this area but this really doesn't seem like their style of evil, and I suspect the article is either a misrepresentation or mistaken")
I find it highly likely that Google -do- try to influence coverage, mind, but I don't think that was the primary factor in play here.
https://archive.ph/3a0wY#selection-715.112-715.649
When I read that I was floored - not because I expected even remotely that it might be true, but because I couldn't believe just how far Wired's quality bar had fallen. It doesn't read like a piece written by someone who's familiar with how the internet works.
Something important about it is this article, by Megan Gray was in Wired’s Opinion section which usually means the writer has more freedom to express their own opinions, and Wired does not claim this it is accurate reporting. But still it was removed.
How Google Alters Search Queries to Get at Your Wallet
Testimony during Google’s antitrust case revealed that the company may be altering billions of queries a day to generate results that will get you to buy more stuff.
RECENTLY, A STARTLING piece of information came to light in the ongoing antitrust case against Google. During one employee’s testimony, a key exhibit momentarily flashed on a projector … [1]
This onscreen Google slide had to do with a “semantic matching” overhaul to its SERP algorithm. When you enter a query, you might expect a search engine to incorporate synonyms into the algorithm as well as text phrase pairings in natural language processing. But this overhaul went further, actually altering queries to generate more commercial results …
The “10 blue links,” or organic results, which Google has always claimed to be sacrosanct, are just another vector for Google greediness …
Google likely alters queries billions of times a day in trillions of different variations. Here’s how it works. Say you search for “children’s clothing.” Google converts it, without your knowledge, to a search for “NIKOLAI-brand kidswear,” making a behind-the-scenes substitution of your actual query with a different query that just happens to generate more money for the company, and will generate results you weren’t searching for at all …
[0] Complete article here - https://news.ycombinator.com/item?id=37802265
[1] A link to the actual slide, that she saw. This image was supplied later by Google apparently - https://news.ycombinator.com/item?id=37802302
Edit: my opinion is Google should respond to the accusations. Removal of the article without a detailed explanation looks real bad for Google and also for Wired
So the path to a potential sale is shorter and you’re more likely to buy (less time to get decision fatigue), and certain vendors might be prioritized.
I think that’s what she is implying
(2) How would you feasibly create a link between every brand who advertises with you and every brand whose site you're trying to uprank? What happens when two different brands who advertise with you appear in the same results?
(3) Most importantly, is there any proof at all that Google is upranking organic links on behalf of brands who advertise with them? (I don't think there is.)
And not only is it incorrect, it is obviously incorrect. The website owners do not pay Google for clicks on the "10 blue [organic] links"; so it gives Google no business advantage to make them more commercial.
Google doesn't get to decide what Wired does with content on its site.
Having said that, the fact that Google goes back and forth on whether they respect the use of quotes to search for literal results doesn't make them any favors.
If I wanted to do what the article claims Google did, I would do precisely what Google is doing. So no surprises there.
I think the only way to know this is to understand whether google’s rewrites also impact what ads are eligible to show.
E.g. if 50 advertisers are bidding on “children apparel” and only 10 advertisers are bidding on “kids clothes”, then rewriting to the query with more eligible ad impressions is misleading to both the user and the advertiser.
Certain keywords are more profitable than others. Google knows this. One sure-fire guaranteed way to increase revenue is to reduce the % of long tail searches (since long tail searches typically have no ads at all) by modifying queries and swapping words around.
If I were a desperate exec at Google, that’s the first place I would look if I was in a crunch to boost revenue short-term.
>I think the only way to know this is to understand whether google’s rewrites also impact what ads are eligible to show.
The other way to know is to ask Google directly (i.e. ask the company, not the search engine) and for them to explain what they're doing and why, in the name of transparency. Google could do that. They won't, but they could.
Isn't their entire slide deck doing just that? Explaining what happens and why?
Sounds like they didn't give the answer you are looking for, so you choose to ignore it.
I just wish they would leave alone the modifiers we can apply to actual search results. (And, for Google Translate, allow specifying that adult results are acceptable. Normally, with a term that can be adult or not you only get non-adult results but it's easy enough to throw in an explicitly adult term to fix that. Throwing in extra terms doesn't work so well with translate--that means it's basically impossible to get the adult result for an ambiguous word. What's the dirty word for the male reproductive organ? English has no unambiguous word for this, thus the translation is impossible.)
At some point they went even further and started doing it even when the words were quoted.
https://www.searchenginejournal.com/google-execs-scheme-to-i...
No one knows exactly how Google alters search queries except Google. That's arguably the issue. It's not transparent.
But the point of the Op-Ed, if I understand it correctly, was not to describe the precise mechanism of alteration. It was to raise awareness that Google is altering search to boost ad revenue and the changes do not necessarily lead to better search, but they lead to better ad revenue.
ISO 8601 tackles this uncertainty by setting out an internationally agreed way to represent dates: YYYY-MM-DD https://www.iso.org/iso-8601-date-and-time-format.html
Thanks for USA for holding back a sensible time-stamp format and the metric system.
I just wanna say that removing an article is not something they do lightly. They never do if because of PR.
They do it if facts do not match the premise. That is all.
>replaces them with ones that monetize better.” We don’t.
How is injecting names of brands (that likely advertise) like "tj maxx" not monetizing better?
[0] https://twitter.com/adamkovac/status/1710041764910846061
"tj maxx" isn't the part being matched, it's the "kidswear" in "tj max kidswear" that is causing it to match to {kids, clothing} because of the aforementioned transform
If you publish an article you have an obligation to archive it. Put a huge flashing disclaimer on the top of you must, but people should be able to read it on your site.
Removal is going to needlessly fuel conspiracy theories. It increases the power of the original story. You’d think wired would understand that.
https://www.seroundtable.com/google-deletes-queries-replace-...
Personally, I dislike this approach by Google (from their response):
> "It’s no secret that Google Search looks beyond the specific words in a query to better understand their meaning, in order to show relevant organic results. This is a helpful process that we’ve written about many times."
This is patronizing, manipulative, and condescending - and 'organic'? They've also been boosting low-quality corporate media outlets to the top of their search rankings for some years now (Youtube search is far, far worse incidentally), as well - part of this 'daddy knows best' mentality, I'm sure.
Google does provide a 'verbatim' option under its tools heading but this is incompatible with time-restricted searching, although perhaps not on the advanced search page. Practically I find that to get what I want from Google (search results from a broad range of sources) you have to jump through many hoops and write complicated queries for no reason other than to avoid their enshittification defaults, and that's a time-consuming process.
It makes Kagi look more and more attractive, certainly - the time-saving alone might be worth the monthly fee.
With web publishing there is no cap on space. You can throw up garbage that wouldn't have made it as a letter to the editor.
And if it gets clicks the temptation to give it a higher and higher profile on the landing page is likely too great to resist.
Now it makes sense!
0, https://twitter.com/doctorow/status/1709221318284173443?s=46