The world needs more search engines
0x65.dev
0x65.dev
For instance:
"the actor that plays the news guy in spiderman"
ddg:
Tom Holland (side bar)
Spider-Man Homecoming (imdb)
Tom Holland (wiki)
Spider-Man: Into the Spider-Verse (imdb)
google:
Jonathan Kimble Simmons (splash, link to wiki)
J. Jonah Jameson (wiki)
J. K. Simmons (wiki)
The google result is exactly what I want. And the results were made in incognito mode so Google wasn't able to cheat with privileged information about me as a user.
At the end of the day, most people care about the product. I'm only willing to sacrifice so much to satisfy the ideal that there should be less concentration. Make a better search engine but trying to pull at the heart-strings of users about how Google is an empire and too powerful just won't work and it undermines your product and mission.
Google also censors some results in controversial categories, where Bing/Yahoo return what I am looking for. I don't trust a search engine that censors entire categories of results.
I am not an adwords user but anecdotes in another HN thread seems to affirm that Google is due to be disrupted: https://news.ycombinator.com/item?id=21667484
"free energy" use to produce countless results that now require extra keywords.
Under "free energy suppression" you find professionally crafted hatemongering. The actual list of claimants is huge, non of it is here. https://peswiki.com/directory:suppression#Wiki_2796702
"free energy device" also produces really crappy results compared to what it was.
It only seems like things argued not to exist may or must be scrubbed from seach results. Astrology wasn't scrubed nor was any religion. Everything has its history too!
DDG's first page was only fake sites and wrong answers.
Google gave me the right answer right away (in the positions 2 and 3)
(for those who don't know legacy electronics: I was looking for the box with labeled pins, like on first page of https://www.turus.com.tr/class/INNOVAEditor/assets/PDF/MM584... )
that's not what Incognito mode does. It prevents your search from being included in the browsing history and doesn't send cookies from active sessions, but that's about it. Google still knows this is you being unauthenticated. You don't need to be logged into google to be reliably targeted with ads that fit your profile.
There are advertising companies that use fingerprinting for ad targeting, but Google doesn't.
(Disclosure: I work at Google on ads, speaking only for myself)
Seems like a pretty good reason to think it still knows who you are.
As to how that works, you tell us.
One way you could see something similar to this would be if you opened a clean session, logged into Google, logged out of Google, thought you closed the last incognito window but didn't, and then opened a new incognito window? Then the user cookie would still be in client-side storage
It's also a good example of their monopoly position; Android, Chrome, Chrome OS, advertising and analytics code on almost every website, ownership of multiple of the most popular websites and services on the internet puts them in a unique position that no one could ever hope to compete with realistically. Competitors have to rely on imperfect fingerprinting whereas Google can probably detect you with more accuracy than a DNA test.
That's literally the opposite of what s/he just said. The person you're responding asked "How would it know?", implying that they (while being on the Google ads team) think there is no way to know without fingerprinting (or cookies from non-incognito mode).
Personally, Google ads give me mixed feelings. I see how personalization is useful for everyone involved and, so long as only machines look at my data, I don't have any personal issues with it. But at the same time, Google collects everything on everyone worldwide to the point where I feel like the USA would have an easy time conquering any country they please (if a nation already has live data on pretty much all its enemy's subjects, war would be exceedingly efficient for them to start and quickly win), so that kind of threatens our freedom if you see what I mean; and secondly the data is not necessarily 100% secure, so in the event of a breach it might be seen by humans, specifically people that I would not want to know what I searched for (or pages I visited that have Analytics or an embedded YouTube video or ads or a map on their contact page or ...). So it's a mixed bag of feelings and your position (job) seems like the kind that would make one think about before accepting. I'm curious to hear your thoughts on it.
I've written some about this: https://www.jefftk.com/p/value-of-working-in-ads
"Many people would put ad tracking on this list of downsides: sites pass information to data brokers that build custom profiles for each user and allow personalizing ads. From my perspective, however, while having this information collected seems a bit creepy, it allows showing ads I'm more likely to be interested in. This makes publishers more money than showing untargeted ads, and I'd much rather fund them through better ad targeting (invisibly intrusive) than through more obnoxious ads (visibly intrusive)."
I chose this team because I thought the work would be interesting and I liked the people on it, and they were interested in me because of my prior work on mod_pagespeed rewriting websites so they would load faster.
> if a nation already has live data on pretty much all its enemy's subjects, war would be exceedingly efficient for them to start and quickly win
Lots of thoughts:
* I think you're dramatically overestimating how much data Google has and how well that is mapped to the kind of identity the military would care about.
* I don't think Google would share this information unless legally required to, and I don't think such a request would be constitutional.
* Many other countries are in similar positions; for example Criteo is based in France and has a similar ad tracking reach to Google.
* I'm still not sure how this is especially useful militarily. Military targets are mostly not in the data one of these companies would have, and none of these countries would go to war targeting civilians.
> I chose this team because I thought the work would be interesting and I liked the people on it
That is fair! I guess most people would make that decision if you already know people there and you think you'll enjoy the work as well.
> none of these countries would go to war targeting civilians
Not as if people in the army are somehow exempt from tracking though?
As for whether Google would share it in the first place: I don't think the government cares much what Google thinks if they're willing to kill (us) over something. Laws can be made by the same people that decide on this. I don't mean to pose it as a simple matter, but I'm pretty sure that's how it works in principle.
Now that I think of it: aren't "national security letters" exactly this? "It has something to do with the safety of the country, just give us that data [e.g. Lavabit private key]"?
Of course, the chance is remote in the first place. Much more likely, if it is ever used for this kind of purpose in the first place, it'll just be posturing and threats, and people will protect themselves better before it ever gets to armed conflict. Just imagine, though, if you're not in the USA, China, or Russia, and one of the three (the most democratic one of the tree, it is fair to add) has the rest of the world's data. That's kind of uncomfortable when I pause to consider it.
> how well that is mapped to the kind of identity the military would care about.
While not readily available, I expect that it's not hard to find a few datapoints to filter them out. Following someone for 10 minutes as they go through traffic and matching the coordinates against location history data is probably enough to find a subset of 1-5 possible accounts. But I doubt physical following is even necessary to find enough datapoints to find them in the data.
And if he really is saying that Google doesn't track you in incognito mode, then I'm going to go ahead and assume he's either lying, or he's not in a position to know about that system. This is Google we're talking about here.
https://www.theguardian.com/technology/2017/nov/22/google-tr...
For the record, I'm not saying that I expect Google not to track me when they detect some privacy mode. It'll sure try to set cookies, and it may use my IP address and connect whatever that IP accesses as a weak indicator of interest for anyone else with that IP address (for a limited amount of time, since IPs change in many countries). What I don't think is that, when they say they don't do fingerprinting, they're lying. This person may not be privileged to know and say "I don't know", but that's different from saying "Google doesn't".
Also for the record, I didn't downvote you (and when you reply to me, I can't; I don't have an alt account with 1k rep or whatever it is one needs to downvote).
Try visiting from tor
It’s in their best interests to also use it for ad targeting (in a plausibly deniable way so they don’t get in trouble).
We’ve seen them using dark patterns to coerce users into opting into more data collection, and another advertising company got caught using phone numbers for ad purposes even if they originally promised to only use them for 2FA, so why should we trust them this time?
If it was being used for targeting it would be practical to run an external study demonstrating that.
I would be very curious as to how you’d prove this is or isn’t happening with a reasonable degree of accuracy considering all the factors involved in ad targeting. Unless you’re willing to give us access to all your source code and SSH access to the systems running it, it’s reasonable people have their doubts.
And surprisingly for most of HN readers, Google has been pretty transparent on the policy of its ads business. In fact, Google has pretty strong incentives for transparency in this area due to advertisers, who give all the money anyway.
An external study to evaluate whether Google is using fingerprinting would be some work, but pretty doable. Targeted advertising is generally very blunt: if someone thinks you're especially interested in a valuable category they'll often pay a lot to advertise to you. So you could set something up where test browsers visit pages related to high-value categories (mattresses, asbestos cancer, credit cards, ...), clear client-side data, and then visit a site that loads ad scripts only from Google (to make sure you're not getting someone else's fingerprinting) and see whether the ads differ from a control group that never visited those pages.
IMO, ads are probably the least worrisome way the data could be used. A boring but scary example is that aol search history leak (which is still searchable today):
This person is identified by name for example: https://searchids.com/user/19431784-joann_whitman
https://meta.stackexchange.com/questions/331960/why-is-stack...
Anyhow, based on what I've seen, when you go into incognito mode it definitely doesn't use your profile when you go to Google services. Search results, suggestions, etc... are different. In fact, one of the use cases for incognito is using a different Google profile on a someone else's or a public computer. Keeping contextual search results from the 'main' profile logged into the browser would be counterproductive...
Another example is a query for "elm dict". DDG has little idea what you're looking for while Google links you directly to the docs of Elm's Dict data-structure.
Most of the time I see comments like this, people don't provide hard examples of queries they found unsatisfying with one search engine compared to another.
You know you're looking for a data structure, right? Add structure to your query, and DDG will do fine. Easy fix.
The opposite case, when google thinks he knows what you're asking but it's wrong, is impossible to fix.
I don't want them to be smart, I want them to be predictable and search what I say, not trying to reinterpret the query.
If you want DDG to understand natural language queries, I think their privacy policy may need to be adjusted so that our queries can actually be used to develop that, and then they need a boatload of funding to catch up with Google's semi-ethical money-generating practices.
!bangs are the main reason why it's always going to be difficult for me to switch from DDG to anything else.
By contrast, the new privacy friendly search engine from the article, with a lot less money and users, can answer a simple question like "news guy actor in spiderman" with
Cliqz:
- And the new actor playing Spider-Man is... this guy (www.foxnews.com)
- J. Jonah Jameson (J.K. Simmons) | Spider-Man Films Wiki (spiderman-films.fandom.com)
Not bad for a small German company uh...
And even if it didn't, when I know it's Jonah Jameson, I can search for the actor that played Jonah Jameson. Search engines are not about having "the answer to life the universe and everything" but helping users refine their searches until they find what they were looking for.
But Google fails as well.
Most of the world population is not native english speaker.
Try "attore che recita la parte del giornalista in spiderman" (the actor that plays the journalist in spiderman) in Italian
- Spider-Man: Homecoming - Wikipedia (https://it.wikipedia.org)
- James Franco - Wikipedia (https://it.wikipedia.org)
- Spider-Man film: tutti gli attori | Popcorn Tv (https://popcorntv.it)
- Martin Sheen, dieci ruoli per scoprire un grande attore ... (https://www.consigli.it)
Or the simplified version "attore che fa il giornalista in spiderman" (same meaning as before, just more down to earth)
- J. Jonah Jameson - Wikipedia (https://it.wikipedia.org)
- J. Jonah Jameson - Wikipedia (https://en.wikipedia.org)
- Spider-Man: Far From Home, nel cast anche l'attore ... (https://tg24.sky.it)
Refining is still a very present need, it's simply that for common searches on common topics in english it's less so...
It's more or less the same in German, French, Spanish, Portuguese... I can't even imagine the results in languages like Arabic, Indonesian or Balinese.
Local, culture aware, search engines are the future, despite Google efforts, the generalist "one size fits all" search engine is not.
Yeah, and due to economies of scale and network effects it will stay that way, unless some people are willing to suffer the minor inconvenience of using a slightly inferior competitor.
But is it? I find myself often having to fight it to search for what I entered into the box, rather than what it thinks I really meant. If you know what you want just not where it is, DDG and Bing are both superior.
I know, Bing.. But I have used Windows for about 30 minutes in the last year and that was just to help my mom fix her printer. I don't think I have ever had a Microsoft related account since Hotmail. I know they track me but I don't use MicrosoftDrive or WintowsTube so I can live with it.
The way they "cheated" was to look at the tens, hundreds, thousands, or tens of thousands of previous queries that matched yours in substance.
If the competing websites for "web search" had the same volume of submitted queries that are substantially similar to yours, then they too would be able to give you "exactly" what you are looking for "9 times out of 10".
If all people submitting queries on the web are somehow convinced to visit only one website and submit a majority, (hypothetically let's say 93%) of their queries there, then it should be no surprise when that website "magically" starts becoming more "predictive" than other websites in returning "exactly" what searchers are looking for and becomes "by far the best at search".[1]
The question raised by the blog post is whether enabling such "magic" is worth the trade-off of also creating a "panopticon" in that single website. Giving this level of visitation and query traffic to a single website makes any competing website (working off less than 7% of people's queries) seem irrelevant, maybe even pathetic.
Without the enormous traffic, I would surmise the glory of the "empire" (to use the author's chosen term), and its appeal to "followers" (e.g., those who marvel at its "magic"), might dissipate rather quickly.
1. Especially the part of "search" that involves dealing with repeated, similar queries for popular information.
Google's advantage is simply billions of dollars and 20 years of R&D into NLP tech.
I doubt that users have a strict expectation of finding the correct (interpretation of "corresponding") answer. Even if the most clicked answer[s] are still not correct, they may still be the ones who suck less, which can still be acceptable/desirable.
Keep in mind that in this perspective, the concept is similar to Google Translate - Google built a translator that, at least originally, doesn't understand language, instead, it applies (applied) a static model on a large amount of documents. While they certainly poured a large amount of money on it, it's success can't be centered purely on the economical factor.
There are at least three kinds of queries that a search engine has to handle. Requests for websites e.g. "Facebook". Traditional keyword searches across the web e.g. "Twitter ban political ads", and questions "Who was the guy who voiced bender?".
Not directly, but it does mean that, say, the second 50% of people that ask the same question will get a better answer than the first 50%.
Google has been able to build fairly accurate instant results based on which sites users were clicking on before. I'd say that a majority of simple general knowledge queries are solved by quoting the first 3 sentences of the wikipedia page that match the search query.
But let's say that there is no easy match to show a quick result for. But after 1000 queries, 98% of users clicked on one particular site on the first page and never went back to the search results. Google then A/B trials how many clicks result from showing that website at the top of the results page in an instant result window. If clicks drastically drop, that's a sign most users are satisfied with that result. They were only able to do that, because of the 1000s of times people typed that query and interacted with the site.
So I'd say that yes, having common questions asked over and over again does help you find the answer that users are looking for.
The funny / scary part here, is that this may not be the correct answer. But it's the answer that satisfies the most users, and is therefore the one that will keep the most people coming back to the Google search engine.
Plus, you're underestimating the percentage of queries that google has never seen before. Google processes trillions of searches every year, and still, 15% of those queries have never been seen by Google before. [0]
[0] https://searchengineland.com/google-reaffirms-15-searches-ne...
No. I am highlighting that 75% are queries that they have seen before. Note also that the 15% is a figure that is declining, based on what it was in 2007.
When you see google giving you a one/two word answer as a card, that’s very likely coming from it’s knowledge graph.
What we need is more of this open knowledge graphs. Google and Wolfram Alpha are both closed sources but have deep understanding in niche domains.
For example, say the film credits for Spiderman is a "primary source", and a cast list for Spiderman derived from the credits at IMDb is a "secondary source". Google extracts the information from IMDb and substitutes itself as the secondary source.
Does this raise an issue in that ideally users should sometimes be retrieving information from (i.e., accessing) primary and secondary sources directly, whereas a third party always acting as a universal secondary source, e.g., a third party funded by advertising, might introduce (more) bias into the information retrieval process, e.g., in competition for "eyeballs".
From the OP references: https://sparktoro.com/blog/less-than-half-of-google-searches...
It's also another matter that so much of the internet is now filled with copy-pasted crap and artificially inflated content that it's actually hard to find what you are looking for on the so-called primary source webpages. If I want to find George Clooney's age, I don't want to sift through a 2000 page ad-ridden Page3-esque gossip blog about him.
At the same time, when I want detailed info about something, I will go in and try to read through the primary material.
In any event, neither a blog nor Wikipedia would likely be a "primary source". Maybe something like a driver license would be a primary source.
It could be that you have a "philosphical objection" to "copy-pasted crap and artifically inflated content". I wonder if Google could have a role in encouraging the continued existence of this stuff. The effective opacity it creates seemingly justifies having an entity like Google.
Even when I entered the actor's name plus the term "age" and was redirected to the Wikipedia search results, I could still see the actor's page as the third result and his date of birth in the summary text.
As for why anyone would want to search some things using Wikipedia versus Google, I can think of a few reasons. I cannot speak for other users however.
I have a script I wrote to search Wikipedia from the command line. For those who care, one can strip out "X-Client-IP" from the returned page html before opening the page in a browser.
"Google's dominance is almost entirely due to the fact that its by far the best at search."
is, and I'm sorry to say, a make-believe
Google's quality is better than anyone else, that is a fact. Let's go to the point. [[I work on Cliqz, and in search]]
Do you know how much Google pays apple to be the default search engine?
According to you, nothing, because people will go to Google because it's the best.
Well, it turns out that last year was more than 9 Billion (with a B). Quite a lot of money poorly spend, someone should really get fired :-)
The other option is that Google does not pay to get the Apple users, which might come otherwise, but to prevent Apple to try something funny, either directly or indirectly.
In any case, the "build a better product and people will come" mantra is flawed when companies in the space pay each other billions for distribution.
Instead, Apple has a thing to sell -- "position of default search engine" -- and it is selling it to highest bidder.
I bet if Bing would offer 10 Billion, they'd make Bing default instead. Hey, if DDG could offer them 10 Billion, I am sure they would set it as default, easily ignoring the fact that many people say they do not like its results.
But please focus on Google, which is the one putting the money!
Why Google pays so much money if they already have the best product and what the users want? Are they just giving B10$ as charity?
There are many reasons for doing that, but none of them is aligned with the make-believe statement that "build a better product and people will come"
I would agree that is a necessary condition, but by no means sufficient, which is kind of sad.
Google pays billions to get those 90% of don’t care users. They already have the users who care, but more users = more money, so presumably it is worth it.
Cliqz, DDG, and others have no chance for those 90%, but they fight for the rest. It’s a hard fight because google is so good. If one spends a hour searching for stuff on alternative engine, then goes to google and finds the results in a few minutes, they will be unliketo visit other engines again.
Yes, when what you want happens to be the lowest common denominator - in this case, the most recent Spiderman movie. When the next movie in the franchise comes out in a few years and you don't like it so much because the news guy is now played by someone else, you'll be complaining that your preferences have stayed the same but Google no longer handles them as well.
In the end, I think most people could use either Google, Bing or DuckDuckGo and be happy with the results. It's just that we're stuck on "Google is the best search engine", and while that may be true in some technical sense, many of the other search engines are just as good for most of us.
I don't really notice any infrastructure problems for ddg in Europe though, so Finland may be atypical?
That said, it does't prevent the situation where searching for certain gifs or images gives you a page full of softcore porn, even with the moderate safe search enabled.
You can almost always, by phrasing more carefully, get exactly what you need with DDG.
I also think that it's better to learn to search properly than finding the best search engine.
Maybe it’s denial or weirdness on my side... after all I am also a vegan because I don’t want to hurt animals or other people just for a little bit of taste and convenience and most people seem to think that’s stupid on my side.
For me their mission is pretty clear: Google ate Burda's ad cookies and now they are trying to get their hands into the cookie jar again. Given that today we have widespread TLS adoption the war about the endpoint has begun. Cliqz, alias Burda Media, is just another combatant - the one who controls the browser controls the ads.
Previous thread with more info:
This "feature" should be off by default on any software that claims to be a privacy respecting alternative to Chrome.
Obviously, if you can get the content of what they're typing, it gets much easier still. I think I've seen papers where they ID programmers based on the code they've produced. This applies to other types of writing too.
You can build a model from keystroke timings and figure out people's SSH passwords too. https://www.usenix.org/legacy/events/sec01/full_papers/song/...
The main problem here is that even if you could do it, why would you? There are over 3.4 billion Internet users in the world. Given that people share IP addresses and even browsers, what's the actually gain identifying someone through typeahead search keystroke jitter? This would spend a lot of effort, and and then tell you what that a cookie doesn't?
I can't imagine that it's actually worth the effort.
And just one more thing: In Google Chrome at least I can turn off auto-suggest. This was not possible in the version of the Cliqz browser I tested.
Disclaimer: I work at Cliqz.
GET /api/v2/results?nrh=1&q=pythuxs=UTlnJzAh4ULkORaiZgLPO6LW9LOWZZoO&n=5&qc=0&lang=en&locale=en-US&platform=0&o=%5B%5B%22custom-search%22%5D%5D&country=de&adult=0&loc_pref=ask&count=10&suggest=0
HTTP/1.1
Host: api.cliqz.com
1. I saw UTlnJzAh4ULkORaiZgLPO6LW9LOWZZoO in other requests too, but not all of them. Can you explain what this parameter is good for and which information it encodes?
2. Can you tell if it is possible to disable auto-suggest (or Quick Search how it seems to be called in your terms)? If it is possible then how do I do it? I couldn't find it in the UI.
3. Can you specify what the exact legal relationship between Cliqz GmbH and FoxyProxy LLC is and if there is any shared ownership or if there are any common parent companies?
4. As you can see above the requests go directly to api.cliqz.com. While the terms go into great length to explain that information is routed via a third party owned proxy the information we are talking about here is exempt from that. I quote the relevant passages here:
> This channel collects signals about WHAT you search and where you land. That is why we do not collect any personal identifier here, which makes it impossible to associate searches with users. Moreover, all query entries and clicks on website suggestions are evaluated only as a single event, disentangling these signals from everything else. Thus, we are neither able to combine data from multiple entries or multiple clicks on website suggestions, nor to link this information with personal information like your email address or an IP address, either.
> Query logging data is used to further improve the Cliqz backend. More specifically:
> To be able to suggest websites in real-time while you are typing into Cliqz’ combined browser-and-search-bar, Cliqz sends your keystrokes to our servers. With every new keystroke, our backend scans our index and predicts the most relevant results for your search query.
> “Relevant” to that regard is (very simplified) defined by the frequency a given website is clicked on for a given query. In other words, Cliqz predicts the most probable site you will navigate to, based on the (partial) query that you type. In order to further improve this mechanism of relevancy, Cliqz logs the clicks in its drop down menu and the respective queries.
I wish the terms were more clear about the fact that crucial information is indeed sent directly to Cliqz and is not sent via the FoxyProxy route.
Thanks for taking the time to dig into things. Comments like these are why I continue to regularly read Hacker News.
Thank you the questions, we are always looking for constructive feedback on and off HN.
1. These random values are used for grouping partial queries together, and they reset when you press enter or start a new query. Source code on how it's generated: https://github.com/cliqz-oss/browser-core/blob/master/module... We actually take one additional precaution of using crypto random and not plain Math.random(), which could potentially be used to link multiple sessions together.
2. There is no feature to disable auto-suggest. But I will pass your feedback to the team.
3. No there is no shared ownership, we don't have access to their servers. We also do additional encryption with bucketing on the payload sizes that we route through foxyProxy, so that the proxy provider cannot learn anything about the content of the message. We will have a blogpost explaining this on Wednesday - 4th December. Also, we are looking to add an option, where user can choose their own proxy provider too.
4. There are two parts: a. You can select the option from Control Center (Q menu) icon in the toolbar -> search -> search via proxy. (Now the calls should go through FoxyProxy) b. All calls to api.cliqz.com go through proxy when in private mode. The only reason it's not default is latency. As to what goes through FoxyProxy by default is: all Human Web data.
Once again, we appreciate you looking into details, and please keep digging, we would be happy to answer, improve our documentation and if there are bugs specially related to privacy and security they are on our uttermost priority.
Well, that's how you implement an autocomplete feature in the 1st place. If they'd sent every keystroke you typed outside their URL bar, now that's something to be aware off. So, did they?
We will also be sharing more details on the process on data collection in the blog posts scheduled in the next 3 days. Of course, you don't need to trust what we say, all the code used for collecting data is open-source for transparency and auditing: https://github.com/cliqz-oss/browser-core
Lastly, privacy policies are legally binding documents, and we take the law very seriously. We are located in Germany, where privacy laws are as tough as they get (e.g. GDPR was loosely just a re-wording of the existing data protection directive in Germany).
That's called typeahead search, and that's how it has always worked since the invention asynchronous HTTP requests.
If you know of someone that sends a trie down to the client containing every possible completion a priori, I'd like know about it.
I see two problems with their approach:
1. The product is not built with the 'grandma test' mindset.
More sliders and widgets is not what your grandma wants in a search engine. This is why building a search engine is hard. You have to guess with very little information what the user wants and get it in front of them at first try, without the user having to tweak anything.
2. Google must not fall because it is a monopoly. If it was to fall it should be because someone built a better product.
Similar to how ICE cars had "monopoly" over transportation and the time for change has thankfully come. Not because monopolies are bad, but because electric cars are so freaking awesome.
Google perfected what the 'best search engine' is to 99% of population. This comes at an expense of really annoying 1% of users but it is the price they are willing to pay. To de-throne Google you really need to cater to broad population with a product that will be better both at capturing intent and delivering and presenting relevant results. This may or not come with a different business model.
> One could rightfully counter that Google has a good product. They even offer it for free.
But Google Ads are not free. If Search is the product, then advertisers are the customers, not users.
I second that. It's somewhat reassuring to know that others realize that as well. Most people just have no idea how marketing and now 'deep learned' algorithms (applied psychology) make puppets out of human beings.
I consider "user feed control" (that users choose what, how, where, when and why things are presented to them) of equal importance to e.g. democracy as far as human freedom is concerned. But the current manipulation is so insiduous that from now to mainstream awareness to generally implementing solutions is a long, long ways away.
> at least $50, in addition to whatever the services I use already cost me
I think this tends close to an upper bound — not many people willing to pay, and even fewer at that level — but certainly one more anecdotal proof that premium services are totally viable for a certain group. The question is 'how big' that group, what's the market for that, but I'd wager it's enough to sustain a few 'premium' businesses (or alternative plans) for most common services (some are harder; search notably).
When a Wikipedia/encyclopedia article is what I'm looking for, why is Google showing articles?
Wikipedia used to be the top result.
I've noticed things on the spectrum of articles <-> blogspam crowding out the top results in many queries, and that it's substantially more difficult to find forum content discussing related topics to the query. It's a shame, as I think there is often more interesting content on such forums.
Google's algo changes since about panda have been burying good web sites and content while bringing quicker answers and 'sanitized' aka semi-censored results to the top.
Some of this is over reaction to SEO and trying to out do the spammers - but the collateral damage to the results and thus the end users who believe that google brings the truth is hard to calculate.
Combine that with regulation like dmca, right to be forgotten and others.. results are even more censored, and the general population does not know what they are not seeing, as they still trust google to be bringing the truth.
worry about bad PR from various factions - tweak the algorithms.. and you can say for sure that the results the more adolescent google brought a decade ago were often more of what people were looking for.. and the results today are often like cheap irradiated / sanitized snacks, not the full enchilada that was once a G search away.
Much of this started happening when whats his name became that adult in the room and started putting the finger on the scale to change what millions could find, it's gotten more and more censored every update since then, and less transparent about that.
in my biased opinion.. your searches and the results will vary. I still use other engines for different things, and I feel strongly that we need more search engines. Anyone who wants to create a better adult engine, let me know.
One of the reasons for breaking up monopolies is that they make it harder for newcomers to build something better.
So You get Google Search Inc. Adsense Inc. AdWords Inc, Gmail Inc, I imagine.
This would be just restructuring, Googlers are too smart to get hold back by this and they probably already have an emergency plan if it remotely comes as a possibility.
My point is that if your company sells information rather physical products then breaking it up is pointless.
When Bell Telephone was split it was cut along geographic lines. That would also not work on a tech company for obvious reasons --- software knows no borders. I don't know of a sensible way to break up tech giants because the efficiencies of scale create a natural winner take all market.
Search, inc is still going to want ads, presumably, so they'll contract with someone. Setup rules for the contract. Maybe require at least N ad providers with each getting a minimum of Y% of pageviews, and contract terms have to be FRAND.
Strongly restrict personal information passing between the companies.
Or, i guess you could go all Bell on them and divide the US into different territories and have Pacific Google and Southewestern Google and what not. Would be kind of weird to geofence search and ads though.
Bell was broken up because they owned Western Electric and used it to vertically integrate the telco stack for the entire country.
A hypothetical breakup of Alphabet could mirror the breakup of the Bell system, as Alphabet controls the data collection -> advertisement stack. Spin off Search + AdWords as an independent business, while services targeting data collection services that feed it (Chrome/Chromium, Android, GSuite, Google Home, etc).
Personally I could see a lot of consumer benefits.
Is Google somehow suppressing the creation of a good search engine? No, on the contrary it has created the best web search engine we've ever seen. I would personally be sad if you took it away from me, and I suspect that 99% or more of its users would be in the same boat (don't judge the zeitgeist by web forum echo chambers). You break up monopolies when they are harming users, not in order to cause harm to users.
With the world being what it is today, Google Search is in an extremely powerful position. We depend a lot on information that we find on the web. And these free form search queries are the best UI we have at the moment to retrieve it.
It is a good thing that Google Search is good and is getting better, but that doesn't take away from the fact that they wield a lot of power that needs to be checked somehow.
In the case of petroleum fueled cars, they are significantly contributing to a massive climate emergency and should be banned regardless of the existence of a superior product. In the case of the Google search engine... Well in the case of the Google company, it should be forced to split up because it is a massive monopoly that hinders competitors from entering any of their dominated markets, and they use their domination of one market to increase their dominance of another.
And this is should be done regardless of the existence of superior products. Goggle as a single company is making the world a worse place, and the right thing to do in that scenario is to split it up.
Would smaller farms with smaller fields and a larger variety of crops be better for the environment? Would those farmers grow more organic food?
I don't think that this is black and white. But smaller companies would likely be a whiter shade of gray in some cases.
But highly engineered products are often ideal for:
1. Having a longer shelf life. (Ideal for countries without the same standards for food preservation) 2. Having yields large enough to feed populations. 3. Increased ability for transport.
I agree that there are gray areas, but it's these qualities that help feed a world.
You should at least acknowledge your bias by being transparent enough to show you have a vested interest. It's really disingenuous otherwise, and makes it hard to take what you have to say at face value.
Though it's important to remember that not everyone on here has English as their first language, and we should generally be considerate of that when reading people's comments.
They've clearly optimized for the mass user base, since the average person is probably fine with a result like that. But I'm willing to bet no one on HN would ever click a link like that, just because of how "spammy" it feels.
When I search for something like "best headphones", my ideal results would be forum/discussion posts (eg. like "Ask HN") where I can read what other programmers/hackers are using, their experience with it, etc. And that's the problem with trying to make a one-size-fits-all search engine - it isn't possible to make everyone happy.
I've actually been working on a product to solve this exact problem. The best comparison would be the old "discussions" filter that Google used to have until they removed it. I'd love to show more people and get feedback. If this is something you'd be interested in trying, drop your email here ( https://degoogle.typeform.com/to/QzVy7c ) and I'll send you the beta
Reddit is a great option, but there's a TON of hidden forums out there that are gold mines for information, which is something I've been also keeping in mind while building this product.
What bothers me is searching for a specific product and having Google promote a competing product to the top search result (or as a "featured" listing that looks like a normal search result).
I feel the force of small sites like HN is precisely to be in the shadow and too small to interest spammers.
I made some typical searches I do in my job. They returned reasonable results. Impossible to know if they are better or worse than Google. I need some days of usage. I'll do my best to keep using it this week.
Maybe autodecting the language would help. It gave me a German page without any obvious way to change to English.
I went to the settings, changed the language and discovered that it wants to know my country. I left Germany because it looks like profiling and I don't want to help them at it.
Then I disabled every feature (news, weather, etc). I'm interested into a search engine, not into a portal from the 90s.
However I'm afraid that Cookies Autodelete and other privacy extensions will delete those settings and I'll have to do it again. I'll probably hide them with uBlock Origin. For the language a URL ending with something like ?lang=en would be great for bookmarks.
And finally, do we need more independent search engines? Yes, definitely.
This is still beta, so please keep on using it and we'd love to hear more feedback [beta@cliqz.com].
If not that, then at least IP, considering I’m sitting in the US.
You should create a contact form for the search engines. Everything reachable from the Contacts link at the bottom of the page is about other products.
* search engine is a cloud service, which is not controlled by user of that engine
* search gives final results in milliseconds (why? because google cannot spend many seconds/minutes for you, that's why)
* search creates information bubble (that user cannot control) because it tires to satisfy user's expectations
Most likely this new 'google-killer' will be:
* open source, because user should trust the code
* self-hosted (easily deployed in a click), not a cloud SaaS. Our search preferences have ultimate value, no one should have an access to it
* more useful than google because of accumulated data about you that are processed by computational knowledge engine (something like WolframAlpha)
* background reasoning - this engine can work continuously and utilize your own computational resources on notebook/PC and bring you brand new search insights that google never will be able to deliver (because they cannot dedicate a lot of computational resources for each google user).
Sounds good, isn't it?.. Maybe this kind of software already exists, could someone point me out?
>search engine is a cloud service, which is not controlled by user of that engine
So is hacker news.
>search gives final results in milliseconds (why? because google cannot spend many seconds/minutes for you, that's why)
Do you want latency to be larger??
>open source, because user should trust the code
That would make it easier for SEOs to game the system. Or perhaps it would put everyone on a level playing field. I'm not sure.
>self-hosted (easily deployed in a click), not a cloud SaaS.
Do you mean users should keep their own index of the whole web by themselves?
>more useful than google because of accumulated data about you that are processed by computational knowledge engine
First I think you greatly overestimate how useful information about you can be. Second if that worked then it would make they problem you mentioned before (information bubble) worse...
Instant results are good for sure, however very often few more seconds - in addition to instant results! - is not a problem if late, more carefully processed results can save minutes of my time - for now I spend it for opening the links and scanning the content with my eyes.
The same is about not very often but important searches that may be described as 'research about something', in this case I'm ready to make complex, well detailed query and wait even hours - then back and get well organized and intelligent results.
> Do you mean users should keep their own index of the whole web by themselves?
oh no! Users should keep only their personal data - in wide meaning, this includes all history of searches, search results, refinements, anything that ML currently uses to bring personal search experience. In addition to that, relatively small index of important content may be saved. For internet search this 'personal search' will use API of anything that can be used manually for now - google, bing, consume direct API of Twitter/FB/Medium/WolframAlpha and hundreds of connectors to other cloud services. It is important to say, that this 'delegated search calls' may be anonymized.
It will be important that search results are not limited only by what google decided to be 'top results for this user'. At this moment I can do all this manually - open N tabs, query many services, compare results, open most 'relevant' (from my human point of view, not google) links and scan them for most interesting information. I believe that all this can be automated.
> First I think you greatly overestimate how useful information about you can be. Second if that worked then it would make they problem you mentioned before (information bubble) worse...
As for now, all this just thoughts. I'm a programmer with almost 20YOE; I have understanding about how google works in general, how lucene works, how WolframAlpha works, modern approaches to NLP and search-driven queries processing, and I think - without a MVP that works, this is more belief, of course - that value of this 'personal computation engine' combined with modern ML approaches might be ultimate. Challenge, but nothing impossible!
I will grant that Google has a more intelligent indexing and ranking of Stack Overflow. However, DDG is making major progress there, and I rarely need to add !g
Pro tip: with the bang shortcuts, you can add them anywhere in your query, it does not need to be in the beginning of your query string.
“This result isn’t that great. I’m going to !g just this once”
Then, one week later I’m adding !g to literally every singe search so I switch back to google because what’s the point?
I really do hope to one day get off of google though.
An example would be something like "Rick and Morty episodes." I know google will give me a list with recent/upcoming episode names and air dates for pretty much any show. DDG will link me wikipedia and fan wikis. I make a "<show> episodes" query anytime I want to know when the next episode of something is released.
DDG's knowledge graph (and/or query parsing) is just so limited I skip it anytime I think google will be able to produce the answer directly. Similarly, there are things I'm confident asking a voice assistant, and there are things I won't even bother trying. If it's something I'd ask an assistant, I'm skipping DDG.
We're better served by breaking them up and enforcing and strengthening the pro-competition laws on the books. The government has a pretty mixed record with regulating industry, and a far more impressive record of investigation, enforcement, and prosecution - corporate break-ups are far more in the wheelhouse. Regulation requires constant vigilance, whereas a breakup is self-executing once the case is won.
It's not anti-success rhetoric, so much as an acknowledgement that tech has become increasingly concentrated and anti-democratic. The Sherman Act was passed in 1890 - these aren't "knee jerk" solutions - they've worked effectively in the past to curb corporate abuses.
But the tech giants run a single system: Amazon in the delivery, FB in the social graph (DB), Google in the knowledge graph. Those would be incredibly hard technically to break up. The less technically complicated split would mean that one new company has all the revenue and the other has all the costs.
I haven’t heard any credible proposal on how this “breakup” would actually physically work. If anyone has links to good sources I’d be interested.
All of the virtuals you mentioned have different stories. Cloud is profitable, and is not the top company anyhow. News is not a product it's a grouping of news stories. Youtube makes money through ads and better access and could be considered a loss-leader but shutting it down will not make the field more competitive. Android is open source.. and perhaps could be seen as dumping to prevent others. Gmail is a mail service, others exist.. and starting a new company will not cost you billions unless you plan on serving billions of people. Drive is one of many companies that didn't cost a billion to start but might be worth it now.. try dropbox or box.com or rapidgator.
The facebook conglomerate could similarly be broken into Instagram, Facebook, etc. Each would get a portion of their current ad business unit.
In terms of software, each baby Facebook could get a flexible royalty-free license to all the software currently owned by Facebook, and they would be free to derive from it to differentiate over time.
If there is a will, there is a way.
Google has:
- Android OS
- Nest
- Search
- Advertising
- Chromebooks
- Pixel
- YouTube
- Google Cloud
There are clear lines where they could be broken up.
I’m not saying they should be broken up.
Fair point - there are certain cases where software platforms are too valuable and complex to break up, and should be regulated like utilities with rules for fair play. Amazon shouldn't be able to use it's marketplace data to monopolize entire categories, Google Search needs to be a public platform, with regulated rates for API access (which Google itself will have to abide by).
You don't have to break up many software systems to prevent the worst abuses. Facebook can keep the knowledge graph, but they can't keep WhatsApp or Instagram. Amazon can keep their store, but they can't own a FedEx competitor or AWS - those need to be broken off. Google can keep Search, but Ads, Doubleclick, Analytics, Waze, GCP, Google Home - all of that needs to be broken up.
Some of the resulting businesses may have to update their model or may not remain profitable, but historically break-ups have resulted in an increase in value for the new businesses.
>Some of the resulting businesses may have to update their model or may not remain profitable, but historically break-ups have resulted in an increase in value for the new businesses.
All of this teaches one lesson: be more like Apple. Treat your customers poorly and massively overcharge them. Make sure you are never the biggest in your field. Essentially, don't offer products that are too good for the price.
I'd also like to point out that if those companies have to be spun off on their own, then the only way for some of them to make money is to sell your data to third parties. Right now Google doesn't seem to do that, but how else would Analytics monetize itself, if it cannot offer ads?
Or just don't abuse your customers by hiding the ball on what you do with their data. Don't buy up all your competitors, so you can squeeze more and more money out of small businesses who have to buy impressions because the organic ways (which the same companies own) don't work any more. The products aren't too good for the price - we don't even fully know what the price is.
> I'd also like to point out that if those companies have to be spun off on their own, then the only way for some of them to make money is to sell your data to third parties. Right now Google doesn't seem to do that, but how else would Analytics monetize itself, if it cannot offer ads?
The same way Mixpanel, KissMetrics, Adobe and dozens of other web analytics vendors monetize (including Google's own 360 Suite). Or make it clear, up front, how the data will be used on free accounts, and give the user a way to monitor and revoke that agreement at any time. As an independent business, they will be free to find the best path.
Charge money for the product. Plenty of other companies in the space do so successfully. There's ample business value provided by good analytics.
Google Analytics was originally the product of Google's acquisition of Urchin Software. At the time, Urchin's self-hosted version cost $895 with additional optional "modules" that cost up to $3,995 extra. Their "on demand" version cost $199 per month.
You’re wrong on both counts: Apple has its own share of potentially monopolistic behavior, and there are billions of people worldwide who disagree with you on how the company treats them.
https://en.wikipedia.org/wiki/Small_but_significant_and_non-...
Does it also have a monopoly on apps that can be run on the “HomePods”? Does every smart TV manufacturer with their own OS also have a “monopoly” on their ecosystem?
Let’s go further down the rabbit hole. Does Tesla have a “monopoly” on software that can run on its cars?
If non insignificant switching costs defines a “monopoly” every software as a service app would have a “monopoly”.
Unfortunately, neither the EU or the US has ever defined “monopoly” like HN posters...
I just explained one possible scenario how this could be defined as a relevant market and you're now throwing a bunch of random assumptions on my comment. I would read the link thoroughly (it explicitly mentions DoJ?) and study more on the history of the antitrust law and its applications before doing such.
If that were the case, every single console maker since the mid 80s would be declared a monopoly. Why hasn’t that happened?
I already told you that the market defining process is a very case-specific one and big techs now are unprecedented. Why are you trying to find applicable prior arts on such cases? You asked how Apple can be monopoly on a market and I gave you one possibility which may or may not materialize due to its uncertain and complex nature. It's pretty hard to understand why you're being so defensive on this issue?
> If that were the case, every single console maker since the mid 80s would be declared a monopoly. Why hasn’t that happened?
There has been multiple antitrust lawsuits and investigations on Sony, Nintendo and Microsoft for their gaming consoles while their exercise of monopolistic power was nowhere comparable to Apple's nowadays. Why don't you google just 2~3 words before making such a false claim?
Maybe because I don’t believe that people can randomly make up definitions instead of citing precedent?
There has been multiple antitrust lawsuits and investigations on Sony, Nintendo and Microsoft for their gaming consoles while their exercise of monopolistic power was nowhere comparable to Apple's nowadays. Why don't you google just 2~3 words before making such a false claim?
Well seeing that Nintendo specifically has been very strict about what was allowed on its platform, forced third parties to use its manufacturing facilities since the 80s, forced all software whether distributed physically or virtually to be licensed and to pay a fee and had a much larger marketshare, where was the government intervention? Where were the consent decrees? Lawsuits?
Please provide one citation where any of the console manufacturers were ever forced to change their business practices?
If my claim is “false”, you should easily be able to find a citation.
While you're trying to frame my argument as "make up definitions", the reality is not; this is a standard practice since 1982. Spend your time on searching and studying the topic, not mine. This is an area of vast complexities and I don't think it's effective to spend my time to enlighten you.
https://www.justice.gov/atr/operationalizing-hypothetical-mo...
> where was the government intervention? Where were the consent decrees? Lawsuits?
Your ignorant in the topic doesn't necessarily mean an actual lack of a prior. In Nintendo v. Atari case, there was not much arguments on the market definition, but its practice was illegal or not. Yes, I've been talking only about the market definition and you're intentionally conflating the concept of the relevant market definition and antitrust violation. Don't do that. And please don't even try to say "come up with evidence". You can spend your time on studying this.
And your citation were about hypothetical arguments, and went to show more of my point. All throughout the article it speaks about the government’s arguments being “defective” and still doesn’t show a single example where a vertically integrated minority player was called a “monopolist” nor where government imposed remedies or sanctions were implied.
In fact, Atari vs. Nintendo affirmed that Atari did in fact violate Nintendo’s copyright when it tried to circumvent Nintendo’s control over its platform.
So in fact, you still haven’t come up with a single precedent where Apple could be considered a “monopolist” on its own platform or where they are violating “antitrust”.
But I rarely see Apple being included in these calls for breaking up the tech companies. Also, I'd be surprised if Apple even has a billion customers, let alone a billion who are happy with Apple.
Turn the relevant parts of it (the ones which are affected by network effects, or other barriers to entry) into an open platform that can be accessed on an equal basis. That's quite a bit easier than breaking up a mining company. For example, Android is already "broken up" from this POV, since the likes of Amazon can use AOSP to create their own Google-free platform, and applications written to run on AOSP will work on either version.
Is it the index itself? There's already a fairly decent open-source equivalent to that in the form of Common Crawl, which has been out since 2011. People (including me) have tried to build search engines off of it, but it never seems to work quite as well as Google.
Is it the computing infrastructure? That's already been commoditized and offered as a service by multiple providers - AWS, Azure, GCP, SoftLayer, etc.
Is it the serving infrastructure? Commoditized by ElasticSearch, which uses many of the same techniques as Mustang and in many ways does it better.
Is it the ranking algorithm? It used to be that Google would agree with you. However, the ranking algorithm changes basically continuously - the one in use now is very different from the one that was used when I left in 2014, which was different from the one where I learned how it worked in 2011. I've heard the new one is heavily machine-learning based: given a suitable training set and some learning-to-rank papers you could construct something similar.
Is it the log & clickthrough data? Imagine the privacy advocate conniptions if that were open-sourced. AOL got in huge trouble when they open-sourced their click-logs circa 2002.
Is it the evaluation system? Mechanical Turk exists, and Google's rater guidelines are public.
I'd argue that the real competitive advantage of Google now is the brand and associated consumer habits, and it's really hard to break up a brand. Same reason Coca-Cola remains dominant 150 years after they started selling cocaine-laced sugar water. This is a recurrent problem in the economy today - brands fuel not just Google, but also other giant monopolies like Coke, Nike, J&J, P&G, DeBeers, McDonalds, Wells Fargo, and so on, and in many cases the companies that own them get to practice some exceptionally bad behavior. But short of reaching into each consumer's head and getting them to consider each purchase on purely rational factors, I don't see how to fix this.
It's a huge advantage and critical ingredient for ranking in all major websearch engines as far as I know. Google gets billions of queries and clicks every day. Your startup? Pretty much none of that. How are you going to train your ranking algorithm with no data?
> Imagine the privacy advocate conniptions if that were open-sourced. AOL got in huge trouble when they open-sourced their click-logs circa 2002.
Right, logs with such level of detail as individual user sessions would never get public for this reason, too much legal risk. But even simple aggregated datasets of the form "how many clicks did this query-url pair get" would be very useful to bootstrap a competitor and can be anonymized much more effectively.
Doesn't this run counter to private property though? Or what would you have happen when Google says "no"?
If you really want to change the landscape and make it actually feasible, force companies to expose APIs or impose standards on them. Both of these are very hard things to do in practice but at least they give a hope of a better future - increasing competition.
Breaking up big tech, depending on the specifics, will cause either worse productivity or (in the best case) change nothing.
>My biggest problem is: How do you break up a software system?
There are people who have created entire careers about exactly that: antitrust remedies. I defer to experts where I can.
We entrepreneurial wannabes on HN are only going to provide unrealistic answers that become battles of will for the rest of this Sunday before all is forgotten overnight.
Counterargument: you can't copy the people who know how it works.
A democratic country has an interest in ensuring that our labor and goods markets also remain open, competitive, and democratic in nature. We've made this choice as a country repeatedly throughout our history, and it's time to do it again.
“Shady data practices” is hand-wavey and non-specific. What do you mean? Many companies today (again, across All industries) have had issues with keeping user data secure, selling it to untrustworthy third parties, not giving users transparency or controls in what information is shared, etc. Google’s track record in these areas is far better than most. How does this necessitate a breakup?
Concentrating industry through M&A is a legitimate issue, in my opinion, but it’s also probably the easiest to regulate.
Our tech companies are the most competitive in the world, bar none. I’m not sure that cutting them off at the knees will help with that. ”We did it before,” isn’t a convincing argument that we should do it now, under much different circumstances.
I used a cover-all term, because there are too many to really name. Facebook is a co-conspirator in defrauding American elections, Amazon uses their data to kill small businesses, Google has been repeatedly fined in the EU for using search anti-competitively. Those are just 3 tech companies, and doesn't get into how data is used abusively by credit agencies, banks, and other financial institutions.
Regulating industry is a fools errand, and part of how we wound up here. Breakups are the only self-executing, corruption-resistant solution.
Are our tech companies the most competitive at in the world? The largest companies have been cutting off our startups at the knees for a decade, so we have no idea how competitive we could actually be. It doesn't really seem like our big tech companies have to compete much at all these days, actually.
So, suppose Google Search is broken away from alphabet. What changed?
I can understand that argument about Youtube (which is losing money and would probably fail), but google search will keep being dominant and their ad revenue won't change significantly.
Every week we have posts about new products, often self-posted. If they're upvoted, it's that people find it interesting. Quite ironic writing the blog is an emotional rant (that I would disagree with, I find it quite argumented), while your whole comment is whining that people upvote things that you don't agree with.
The problem with monopolies is when the company decides to e.g. price gouge or gets lazy by not innovating. Companies can only really get away with this and survive when they have a monopoly.
Monopolies also kill competing products and deter other companies from even attempting to enter the market - so you might love the product of a monopoly right now but if the monopoly had never existed an even better product could be available today.
I think people generally love Google's products but it would be good to see more competition. The upfront investment you would need to create a competing search engine is prohibitively high so most companies aren't going to attempt to make their own.
(Disclosure: I work for Google, not on search)
Google Search is pretty well embedded into Android. You can use Baidu or Yandex or Bing / DuckDuckGo on an Android phone, but there's substantially more friction. Considering mobile is now more than 50% of search traffic and Android is 80% of devices, that's close to 40% of all searches basically going to Google for free.
Combine that with the fact that Google is embedded into Chrome -- which is used by ~67% of web traffic -- and there's SOME friction to using a different search engine:
You have to set your default to a different search engine, or go directly to that page, rather than just type into your URL bar -- like most people do.
It's not hard to see that Google has a huge advantage.
Desktop is dominated by windows, where bing search is embedded in the start menu, from where it is non trivial to remove it, and changing it to something else is not even possible.
How is there more friction on using another search engine on Android. What other platform has an even lower friction in switching search engines?
Which they mention a line below your quote.
https://blog.mozilla.org/press-uk/2017/10/06/testing-cliqz-i...
- a search engine for blogs
- a search engine for dev questions
- a search engine for shopping
- a search engine for diy
etc.
Im really tired of sites trying to centralize and own this content. The key to decentralizing and empowering individuals to own their content is enabling distribution and discovery. New search engines are essential to the web we want.
That would be nice especially since hopefully it would mean there would be a search engine I could use and not get shopping results. It seems like anytime I search for something I get shopping results so I have to be more specific and type more which is inconvenient. Wikipedia used to be the first result for many things I search for now it pretty much never is. It's been getting worse every year.
On a pretty unrelated note, it would also be nice if I could search the play app store without getting games in my results. There's pretty much a game related to everything I want to search for. Very annoying.
I'm pretty sure that we won't see a significant blog search engine again until blogs regain some of the power they lost to social media.
I thought Mozilla would loss all their credibility after this, but somehow they still market as privacy-focused..
[EDIT]: Just to clarify and not have anyone create the wrong idea - I defend my earlier point. But your question is loaded. Here's why: We do not collect browser history, which by definition implies being able to piece visited urls back to a profile in our servers. That is impossible - to us, each single URL comes as a detached datapoint - devoid of any information that can be used to aggregate them back to a user profile.
>That is impossible - to us, each single URL comes as a detached datapoint
IPs could be used to aggregate those datapoints, and you obviously cannot avoid receiving these. It is only promises that you or your proxy provider doesn't store them. (though maybe it is possible to implement P2P mangling network? encrypt UDP data packet, send to randomly selected peer discovered from DHT, peer delivers it to your server. Or directly send UDP packet with spoofed source address, but this is not possible for browser sitting behind NAT)
I also stick to my original point: those users who had cliqz had significantly more privacy than those without.
Having said that: I don’t think, you and me are that far away from each other. But: If we, who care about privacy constantly criticize or even shout at those who also care about privacy, those who build better products, but maybe don’t follow an idealistic “no data at all paradigm”, then we will always end with the worst data collectors, because non of the alternatives will ever have a chance (or people get frustrated and decide they can make more money at Google or ad tech).
By the way, we have a post about data and how we collect it in our blog today: https://www.0x65.dev/blog/2019-12-02/is-data-collection-evil... - you might find it interesting).
In any case thanks for challenging us. I don’t believe we’re perfect. But we’re trying!
We love to get scrutinized and get feedback on this - we’re very serious about privacy.
Does this mean the article is insinuating that DuckDuckGo is dependent on another search engine? Is this true? Or are DDG results organic?
The truth is they rely largely on Bing and Google results and mostly just rerank them. That's why it's hard to take their PR seriously.
It's also clear that they try to obfuscate this fact by claiming "more than 500 signals / sources", which means they feel that being more upfront about the truth would work against them.
I wouldn't be surprised if it is the primary source of their search results either.
Developing a new search engine from scratch that can compete with Google is almost impossible unless you are at the scale of Microsoft.
It is much easier to white label an existing search engine so you can have good search results on day one.
> Paraphrasing Monty Python: “What has Google ever done for us?”. They have done an awful lot for society, for the web and probably saved it a few times a long time ago. But it is time to also see them as the empire they have become. And empires must fall.
I think people generally do not understand well enough the far-fetched consequences of a monopoly in the search business. We really need more diversity and competition in this space, not just for quality of results and convenience of services; but for society as a whole.
The article rightly points out that Google is a new "Cambridge Analytica" (even on steroids), we should not passively wait for our democracies to break down because information is controlled by a single private entity. Let's step out of our comfort zones, try new products, support the ones we like!
There are plenty of alternatives. But people don't use them as much because they are worse.
Consider Bing ( http://www.bing.com/ ). It is backed by a company even bigger than Google. MSFT's market cap is 1.155T; Google's is 899B. MSFT is 30% bigger than Google. Why haven't they put more resources on their Bing search engine? Spend a few billion dollars, hire away some of the top search engine experts from Google, and build away!
No ads, no algorithms, no bullshit ranking and metrics, no tracking, just pure boolean search with a heavy handed spam filter and jam packed with features for the sake of utility and not engagement.
>just pure boolean search
So if I just repeat the word Facebook a hundred times in my personal website, I'd rank higher than facebook.com when people search for "facebook"?
I think simple boolean search works for things like searching books in a library, where you can assume good faith from all the authors. But in the web you can't do that. It's just too easy to game.
Maybe you should've disclosed that you work at this company?
'far-fetched' (meaning implausible) should be 'far-reaching', right?
One of the most amusing patterns, both in newspapers and on HN, is how often whenever Facebook does something wrong, the response is "this is why Facebook and Google should be regulated/destroyed/broken up".
EU should have its own search engine. They should subsidize this effort and have Google pay for it.
But the European taxpayer should absolutely pay for it.
I'm unsure if that continues to be true if the EU siphons money off of them to build a direct competitor.
Or in some other other gradual way. Anyway: it's not a binary over-night thing based on a whim of some dictator.
Either way, I fully trust the competency of the people running these things; they have proven themselves in the past. I'm writing that as a Swedish citizen who is a little bit skeptic about whole the EU thing. This part though - I trust these EU bureaucrats. My impression is that there's a team of very sharp people working on this.
A basic web crawler is not. A semantic web crawler and associated search engine capable of mapping real human input to useful information and dealing with the heterogeneous structure of data on the web, with low millisecond latency?
Bing has been working on that problem for years now, how come they haven't caught up more with Google?
But then again, there are also a lot of other problems Microsoft haven't been able to solve...
Need not to be more complicated than that. It is not like Google is the same darling it once was at home either. The public now has an appetite to go after US big techs not vice versa.
And Google won't have a valid defense against it, unless the US government comes to play
It will also help insofar as site owners will be less inclined to block "EUROPEAN SEARCH INDEX CRAWLER", just as they're less inclined to block google, despite being inclined to block "small time search index".
The creative, non-capital intensive part where having a diverse ecosystem will help the most is building a layer on top of the index to actually find stuff - e.g. a bit like how duckduckgo is built atop bing's index.
Google already payed a bit, now EU can begin :D
I look forward to reading the rest of the series.
I see no mention of their technology being open source, which makes it a lot less interesting as an alternative, to me.
Their search UI seems OK, once I turned off all of the noisy widgets, including the terrible pretty background image. I wish this trend would go away, but I'm glad they provided an option to turn it off.
I don’t trust a browser nor a search engine if an ad imperium paid for it and gives it away for free.
Does the Cliqz browser still disallow plugins, such as ad blockers?
Hey, I work on Cliqz' adblocker (open-source here: https://github.com/cliqz-oss/adblocker). Cliqz browser comes built-in with adblocking, anti-tracking, anti-phishing and private search built-in so that users are protected by default.
You can also install another one if you prefer since all Firefox extension are also available from Cliqz! But the one built-in is pretty good already[^1] and we're happy to get more feedback about how we can improve it!
Finally, Cliqz search is also available standalone now so feel free to give it a try! https://beta.cliqz.com/
[^1]: https://whotracks.me/blog/adblockers_performance_study.html
Yeah, that’s why I don’t trust chrome or google search.
Still, I was referring to the Burda conglomerate. Why would we want to replace one ad giant’s tech with that of another (more regional) ad tycoon?
Lol, yeah. The HN headline is like "hell yeah, break up and/or nationalise Google/Alphabet," but the link is "get our product please think of us poor little oppressed advertising agencies" and, uh, no.
Yeah, I'm going to stick with DDG. Questionable ownership aside they at least have a functional website without running javascript.
“Skate to where the puck is going, not where it has been.”
Google will not be replaced by another search engine. It will be something else. Something we can’t imagine yet.
I’ll give Cliqz the old college try, but I hope you make it a priority to get into the next release of MobileSafari as an option if you’re intent with taking on Google and taking them on seriously.
Here’s a data point of one: my default method of “searching the web” is to start with Search on iOS, what you get when you pull down on the home screen. When the results I want aren’t there, and I can usually anticipate when that will happen, I shift it over to Safari which uses whatever my default is, DuckDuckGo but mainly because I can change my search engine on the fly with them.
I also hope you incorporate the same syntax as DuckDuckGo (if you haven’t already) because to be blunt, every search engine I’ve used in the last few years is terrible. I can’t use just one, and I end up doing a lot of fancy site searches. Not even Google today is as good as Google from 2007-2008.
I don't disagree with the premise, after all I was part of a team that tried to do the same thing (well, without the ad network) and discovered that the adtech business has become so corrupt (not that it is all corrupt just that enough of it is) that the only way to cover the costs of running the thousand or so machines you need to hold a decent index and cache is hard to achieve.
If the answer is privacy, that’s a niche already occupied by DDG.
What other ways could a web search engine differentiate itself to be worth customers using more than one?
These vertical search engines, at at least in the comparison shopping space, were never that good to begin with. Shopping (and possibly travel) is just the easiest vertical to monetize. All the sites were spammy, pages were optimized for clickouts, and results were optimized for yield, not relevancy, because everything was an ad. There's a reason none of them were recognizable brands and they all depended on Google for traffic.
So yes, not just what niche is there, but how can you make it profitable--or even break-even. Would you donate $10 per year to a non-profit search engine?
[1]: https://techcrunch.com/2012/06/08/nextag-ceo-google-is-a-mon...
i really miss using an image itself for query terms.
Hopefully there will be even more stringent antitrust enforcement to prevent Google from buying favorable search placement everywhere, but Google has such a big edge both in consumer awareness and people adapting their searches to Google that I'm not sure how it can be dethroned.
This is the reason they're pushing their browser and their extension. Privacy wise, this method sounds OK if they really don't send the data (or what ads you see personally) out. I don't mind seeing ads if they are somewhat relevant, non-obtrusive, and not stalking me across the internet.
My first impression about Cliqz search is that it is somewhat viable, but the index is pretty shallow. There are actually a few other search engines with their own index not mentioned by the article (Google, Bing, Yandex, Baidu):
* Mojeek. Sometimes really good for obscure sites when you want to "grep the internet" but seems very vulnerable to blogspam and other SEO. Maybe I just haven't figured out how to query this one yet.
* Yippy. Pretty decent index. Cool feature: categorizes search results with a tree on the side, so if you search for "cobbler" you can just remove the pie or shoe results depending on what you're looking for.
* Gigablast / private.sh: private.sh is the new Gigablast proxy run by PIA. Shallow index as well, haven't used it as much.
The search engine market will be up for grabs if Google keeps getting worse at the same pace.
most it's saved, some is lost forever
The first attempt was in 1999, i built a meta search engine using coldfusion.. I didn't have any server to run it on so it was on my dialup running locally. Worked pretty well, i was a big fan of metacrawler at the time.
The next time was back in around 2003.. I had a server running on my adsl connection at the time. I wrote a crawler of sorts that indexed word positions to search by proximity. At the time, Google didn't do this, it was a feature i was looking for. Eventually, the project died. If i had a 200k startup fund like google had, it could have probably turned into something.
I did build it on mysql and perl but it needed more servers.. Something i didn't really understand back then with only about 4 years web experience.
Then, i built another a meta search engine, ran it on a hosted server, gained some limited traction around 2011. Died as the sources for the api dried up or became too costly. It did make some money from advertising at the time, which is a pretty big deal. The coolest thing was the fact that i kept the searches open, each search had an rss feed that you could use to keep track of certain keyword sets.
Something i need to say about Google is this: it seems simple, but there is a lot going on in the background. Its doing a lot of things like checking for malware, spam, link farms, etc. but also its running the page in a real browser. Its pulling in data from google analytics and clickthrough/bounce rates, its always split-testing result sets, showing results based upon geographic location, language recognition, date recognition, etc. They use machine learning with real humans training those models, that's not to mention their internal systems for managing servers.
I think you shouldn't try and copy what they are doing, in the end it doesn't matter if you do.. Look at bing, its kind of like a joke compared to Google.
Facebook did the right thing, that was to bring the community and data onto the site itself. Identity is the future. Reddit is pretty good as well, you can talk about any link on the internet.
User-moderated content is the future, as AI/ML can always fool another AI/ML but not a human.
You only have to look at this site hacker news to tell that.
Also Cliqz belongs to Hubert Burda Media, a company which I trust even less.
Others that might seem to be in market are niche players.
Scale is good for consumers, benefit of scale is a thing, so 'trust busting' should happen if there are fewer than three.
In the us search market, we have google, ddg, and Bing.
Wanting more choice is very American, but it's hard to take seriously.
Others disagree, including Google Chief Economist Hal Varian, who said (in 2008)[1]:
“Scale is pretty critical [because] search technology exhibits increasing returns to scale.”
Or Jonathan Rosenberg, Senior Vice President of Product Management and Marketing: “So more users more information, more information more users, more advertisers more users, more users more advertisers, it’s a beautiful thing."
Or Eric Schmidt: “Scale is key. We just have so much scale in terms of the data we can bring to bear.” “We are a company that operates at scale . . . trust me it is a scale company.” “We think search is about comprehensiveness, freshness, scale and size for what we do. It’s difficult for [Microsoft] to copy that.”
It seems obvious to me that internet search has both economies of scale and network effects, both classic barriers to entry.
[1] http://fairsearch.org/fact-checking-google-scale-is-a-barrie...
Also,
> With 93% of the search market, Google’s algorithms decide what becomes truth. Can you think of a TV channel with a 93% audience? Would you find it acceptable if there were only one TV channel?
Seems to me like Google is more analogous to the TV remote.
The mystery-box-that-reads-my-mind-then-poops-out-links model feels increasingly clumsy as the years go by.
For those who don't know, .dev is operated by Google.
*- and by "handful" I mean tens of thousands of sites, which is relatively small.
They do not and it would be a sad day if they did. Please don’t help push us towards that any further.
but:
> since the internet is so vast now that only Google/Microsoft/Yandex can afford the server/bandwidth fees to crawl it.
Is it a vast space that search engines need to crawl, or are there only a handful of sites that matter?
Disclaimer: I work for Cliqz.
We are aware of this and it's unfortunate, although the lists are not the main ones from community like Easylist, uBlock Origin etc. Cliqz does need to collect data to build it's search engine and other features. Even though we do with by privacy-by-design in mind and without collecting any PII (we will be sharing more details on the process on data collection in the blogposts scheduled for Monday, Tuesday and Wednesday),still we do end up getting on such lists.
Of course, you don't need to trust what we say, all the code used for collecting data is open-source for transparency and auditing. Additionally, also happy to help setup network debugging incase you want check what exactly goes out. - https://github.com/cliqz-oss/browser-core
We are also willing to discuss with the maintainers of the list and explaining what and how data collection is done.
> In the TV world, this would be the equivalent of 100 different channels, but they all show Fox News 24 hours a day and just replace the logo. This clearly cannot be good.
Unfortunately that's already happening, just look at Sinclair [0]. They operate 193 different TV news stations covering 40% of US households and syndicate very similar content between stations (with the same political biases, just different logos).
But a search engine is only a smart part of the whole, where and how we get the information.
If you slice your market in convenient ways, you can make any company a monopoly.
SEO doesnt even count that much, instead it's curated "reputation" that matters.We might be at a point where DMOZ is probably viable, because there aren't now 1000 entries in each category, but 10. In fact a Dmoz would probably be a mightier competitor to their search.
But the key is , we need competition in advertising. After http referrer was removed from queries, we left all the information about the user's intent to google, and, unsurprisingly , they ate it all, and then some more, leaving a progressively diminishing piece of the advertising/intent processing pie to the rest of the internet. Is that ethical? Is it OK that a website doesn't know why the user visited it? After all, the user typed it in the browser's bar , not on google. Perhaps search queries should be initiated in the browser, and available to subsequent clicks for processing.
And as for tracking, is it OK that google is the only one who is tracking users? Shouldnt tracking be a user's decision, provided and mediated by the browser to any website who asks for it? Like it or not, advertising is not going away and it's healthier for the net if a large number of publishers share that (growing) pie rather than if google eats all of it. Tracking doesnt kill people, but anti-tracking hysteria causes many people to lose their jobs. Perhaps we should be pragmatic, because being irrationally anti-pragmatic is only serving Google's long-term interests
And personally I'd rather have no internet at all than an internet that's a weird murky tracking soup where a hundred different parties are creepily looking over my shoulder, knowing more about me than what I would tell even my closest friends and family.
Well, DuckDuckGo is pretty known and independent of Google. Not long time ago was an article here on HN about how DDG operates and they have their own web crawler, not relying on Google.
Why?
Brittle empires fall.
The Chinese empire has a history (including mutations to its government) spanning thousands of years.
But no, it wasn't just the end of the government, it also meant losing almost 30% of its territory.
https://imgur.com/r/mapporn/Cdi6KOL
> In the same sense that the US and UK are empires.
UK was an empire, and it's already fallen, it is now on the brink of losing even the Kingdom and becoming an Island hosting borders and customs.
U.S. is not really an empire, never was. It is more the mandator of many bad things that happened in the last 70 years, that turned many countries into U.S. colonies through "exporting democracy" which actually translated too many times as "planting dictatorships befriended with U.S.A.". But it was good for U.S. and only U.S.
That's why Cliqz has reasons to exist, U.S. is an unreliable ally, even more now and Europe needs to
> building the foundation for a sovereign digital future of Europe
> develop digital key technologies as a European alternative to the market dominating US platforms.
> If We Don’t Act Now, We Will Become a Digital Colony
The point was that every empire falls sooner or later and that it happened to China as well.
Disclosure: I work at Cliqz.
So is it necessary that every grouping of man, "must fall", because there exists "coercion" in those groupings?
Just as a matter of full disclosure:
My own view is that any grouping must fall only when the grouping provides no, or very little, actual benefit to it's constituent members. Western Rome fell when people in Lutecia, Londonus, or even Ravenna were no longer deriving real benefit from being inside the Roman polity. Which arguably was long before Odoacer finally had enough and put an end to the charade.
While it's ridiculous to view wealth as a finite thing, it's just as ridiculous to imply hysteria over wealth inequality is based in ignorance. There is a long-standing trend of capitalists treating lower classes as exploitable resources undeserving of the means to have a healthy and happy existence beyond what keeps them productive inside the constraints of the current system.
Median salaries haven't been stagnant while the rich get richer because that's what's best for everyone, it's because that's what's best for the people with control.
I recognize that there is still room to improve DDG's search quality. However, I wonder if DDG or any competitor can ever get as good as Google without collecting user data and using that to contextualize searches.
But we shouldn’t trust google, what trust existed has been repeatedly damaged by their desire to grow and profit.
In my opinion we need transparency. Why is a result / snippet number 1, 2, 3, etc for a query?
“The algorithm did it” is not an answer.
Any new search engine should be accountable for the results they provide, and there should be a mechanism to dispute the results.
Kinda like road signs vs gps
This article was negative to the point of being unreadable. This is not how you disrupt.
While I'm wishing, it would be nice to hear about how they anonymize their logs. Their privacy notice only really mentions scrubbing associated IP addresses - which may not be sufficient.
I’m extremely disappointed in the HN discussion here. The blog post is by someone wanting to compete with Google. Great. The entirety of the post makes ZERO compelling arguments about what features, guarantees, or outcomes their product will provide. In the last couple paragraphs, they mention the word privacy - okay but how so? What’s the business model? Why should I trust you?
The blog post is nothing more than an attack on Google (and other tech companies, but that comes later in the article). That’s the only “substance”.
There’s no question that key regulation is missing in the space. But what exactly is Google even doing that is anti consumer and killing competition? What anti trust laws have they broken? Should we identify them or maybe determine what is missing in our legal framework? If the author wants to compete with Google, perhaps they can share that insight. That be productive.
>Google is Cambridge Analytica on steroids.
That’s a pretty damning statement and unproductive. You want to make an emotional argument and have keyboard activists work for YOU instead of sticking to objective facts.
Disappointing this is #1 on the front page. We are turning HN into Reddit.
True content is buried or listed in a row or two out of 10 plus results.
Mobile is even more flagrant. Inswitch engines, just to get a less mainstream and digested view of my queries.
[1] https://www.comscore.com/Insights/Rankings#tab_search_share/
These guys call for banishment of "search monopoly" yet are in it for the money and business share. Privacy is just a currently sound competitive angle.
I would prefer the UI in English and I want hits from wherever they may be found not restricted to some small area.
Good try, but some way to go.
There is a very strong national security argument here: you don't want your citizens' web searches to be controlled by a foreign company, even if you're nominally allied with the country, because of the immense political power that control of web search entails.
This kind of policy would unite the populists on both the left and the right. The left would like the idea of redistributing money away from big American technobillionaires to local businesses, while the right would appreciate putting the interests of the nation first.
"MyOffrz is opening a brand new marketing discipline: Browser Based Performance Marketing. This innovative technology makes it possible to show users discounts, special offers and valuable information with true added value within the browser. As a 100% subsidiary of Cliqz GmbH..."
"These include independent search engine, browser and privacy technologies as well as new techniques for responsible advertising and statistical data collection for the benefit of all users."
Seems a bit odd for a company that brands itself as a privacy search engine.
Cliqz is nothing but a parasite disguised as a pro-privacy product, funded by Hubert Burda Media (https://en.wikipedia.org/wiki/Hubert_Burda_Media).
Their strategy to move into the couponing/voucher market comes as no surprise, since injecting ads (which this basically is) is a very lucrative business - and they have the customer base to do so. The line between malware and "legitimate" browser (extension) becomes increasingly blurred.
So even if they don't do "tracking", injecting ads surely isn't something the average user "signed up for" when downloading the browser.
What are our options? Government? How will it be even remotely close? Data will be needed, How will the government get data? They will probably make it mandatory in schools of course. Why not? its for the good of the children.
Interesting times indeed.
“The individual is in a dilemma: either he decides to safeguard his freedom of choice, chooses to use traditional , personal, moral, or empirical means, thereby entering into competition with a power against which there is no efficacious defense and before which he must suffer defeat; or he decides to accept technical necessity, in which case he will himself by the victor, but only by submitting irreparably to technical slavery. In effect he has no freedom of choice.”
― Jacques Ellul, The Technological Society
Don't you have fiduciary duty for the owners of the company to maximize profits? Why do you want competition?
Whether Cliqz would become a bully like Google if successful or not, I believe it's impossible to answer.
But at least, we know, that the data we collect from our users to build the search engine is totally anonymous and is used only to build search and the browser.
So, in the case that Cliqz would turn evil, I would stop using it, knowing for a fact that there are no sessions about me on their data, none.
Someone on the thread mentioned fiduciary duties, true that companies must maximize profit, but that does not mean becoming a gear on global surveillance conglomerate.
Rather unfortunate typo, that.
I was around during the “search wars”, and what’s usually forgotten is that well, they were by far the best behaved as a corporate entity.
This “romantization” of this ideal state, devoid of historical context or acknowledgement of the world we live in is extremwly counter productive.
Should google, Facebook, Amazon, Apple do better? Absofuckinglutely!!!
Regulate them, hold them to a higher standard!
But Will breaking them up do anything except destroy them and just empower the next competitor ( who will have to learn the same lessons?
Even worse, likely in a country that won’t match your same values of “freedom”
Are imbalances in power great? Nope
Are monopolies awesome? Most never
Are billionaires awesome? Nope
But this is the stage we are at due to an interconnect global economy. We built this.
Knee jerk solutions don’t and have never fixed anything
It is not. Apparently some people are still fooled by the doodles and dinosaur jump game.
It’s easy to say MORE REGULATIONS... just as after any tragedy people claim “we just need one more law”! The truth is government has proven toothless in tax collection and regulation of these giants.
They were well behaved when they had competition to worry about. Once they effectively achieved a monopoly, ethics slowly got relegated to the back seat. They started leveraging the search monopoly to take over other industries, to undermine the open web, to engage in political censorship (hello Project Dragonfly) and much more.
Everything you say here is an excellent argument why search should not be a monopoly. We need competition to keep the search engines honest. Regulation won’t help, since orgs like the FCC are vulnerable to regulatory capture, and there’s no other way to "hold them to a higher standard" than taking your business elsewhere. As long as you continue using their products and thus earning them money, they don’t have any reason to care what you think about their business practises.
That's why they want Google and Facebook broken. And of course, they're unwilling to spend any money, change any law anywhere, accommodate anyone, or put in any kind of effort whatsoever.
Fundamentally they want the internet to disappear. They want the control of information out of the hands of these idiotic American nerds that refuse to control information !
Right now the only country that could actually do anything about Google is the US, but they won't because we benefit from having the number one internet search company be American.
Also, the rank and file employees in Google are largely liberal and tend to make a big fuss over things like censorship. If Google is broken up a company that is more friendly to state interference could emerge and take it's place.
Google was trying to gain the Chinese market by implementing political censorship. And once they knelt and swore allegiance to the Chinese regime, every other totalitarian regime around the world would be lining up with demands that they do the same in their countries.
Expecting Google to stand up to EU on anything is naïve. They already got massive fines, they don’t have the spine to risk more. They will do as the commisars dictates, whether they’re split up or not.
I'm super glad that google didn't go through with dragonfly and it's really sad that they even went as far as they did. However, going through with it would have just made them the same as basically every other company.
Apple, microsoft, and basically every other company all operate in china, following chinese "law". Ironically, it's two of the worst/most hated companies (google, facebook) in the US, and most frequently demanded to be broken up here on HN, are two of the very, very few companies that haven't bent over for china.
In other words, their refraining from Dragonfly wasn’t because of moral courage or a principled stance. It was just to avoid a PR disaster.
And it was only because Dragonfly was leaked that that we ever heard of it. Imagine what other interesting programs they’ve got going on that we haven’t heard of yet.
Even as a Google Employee I agree that the concentration of marketshare and infrastructure that Google has distorts the market. But without radical intervention either by normal humans collectively or by governments using legislative and military power to shape corporations into their vision (again, probably collectively to be effective) the problem can't go away.
At the very least, we need to stop pretending that pointing our the problems of rent economies and monopoly tendencies in capitalism is not "an anti-success rhetoric" but rather a candid and informed discussion. Markets do not function well (defined as delivering utility to the largest number of people, we could of course define this in other ways) when you need to render the player's assets on a log scale.
For many people in the world, the US already is a country that matches this exact description - what with toothless antitrust laws, anti-union actions being tolerated, and what the ICE camps look like is something we don't even have to talk about.
Sure, China would be worse, but you should strive for more than just "the second worst developed country(tm)"
In fact, yours is the perfect argument why we in Europe should try to end the dominance of Google and break it up.
In the case of search engine balkanisation (I'd be in favour of it), I'd think that eurosearch would, on average, be more censorious than amerifind.
Every European engineer I talk to seems either amazed or resentful about American software engineers’ salaries and benefits. How have the unions helped you in this regard? How have they helped your tech companies be more competitive in the global economy?
E.g. we get health insurance even if we quit or switch jobs.
The thing about European software companies tho is that they are hardly unionized at all.
>Every European engineer I talk to seems either amazed or resentful about American software engineers’ salaries and benefits.
I am not, to be honest. Sure, at a first glance it might seem the US devs make tons of money. But once you account for cost of living (like the astronomical rents in every major tech hub city) things do not look that rosy anymore. I know US people who make twice or trice what I make, and yet they live in dumps that you can hardly call apartment and they say they cannot afford any better. That, combined with the at-will nature of US employment, really doesn't strike me as desirable.
I like living in a country where I have a pretty awesome standard of living, a nice apartment that I don't have to share (unless I want to) and a safety net that is (at least for now) able to catch me should I ever struggle or become sick, so I won't go homeless.
Those things are largely thanks to work the unions did in the past and keep doing.
If breaking up the American tech companies helped disperse jobs outside of Silicon Valley, that would be a good thing.
The title may be, but the article was not. Was a reasoned statement for why those folks bothered to build another search engine.
It was an inspirational company and story that many rooted for due to its culture and mission.
Were. When though? The founders said in their original university paper that selling advertising was wrong and would inherently corrupt a search engine. Of course once they realised how much money was being made in search, that fundamental tenet of don't be evil lapsed before the phrase was even coined.
History and Google's progress of updates seems to have borne their claim out. Google is worse than it was, in good measure as they have twisted results to push more sales. It's a shopping search first and foremost, and frankly not a very good search any more. With an obscene overreach of data gathering that is almost impossible to evade. Every opportunity to gather data they gather more. Remember the Google war driving guy who "accidentally" gathered all that extra wifi info from streetview cars? Entirely by accident. How many believe in that "accident" who isn't on Google payroll?
So no, I don't see it as anti-success, more that they became far too large, far too greedy, and far too abusive - e.g. demanding everyone train the self-driving fleet to pass a captcha. Scanning the world's books was a far more appealing proposition.
Given they have the data, and the holy trinity of search, advertising and monitoring, regulation is not enough. Split off advertising or analytics and monitoring, or break it up some other way. That's to serve the public interest, not as some anti-success crusade.
That proposition failed precisely because it was regarded - rightly or wrongly, it's not clear that it matters - as a "far too large, far too greedy, and far too abusive" endeavor.
I disagree here. Billionaires are often pretty awesome. Once someone gets way more money than the could ever need, they tend to pursue other activities that they enjoy more than making money, which is frequently giving it away (Bill Gates, Warren Buffet) or starting moonshot space companies like SpaceX, Blue Origin, or Virgin Galactic. Even Zuckerberg donates to projects like these.
Rich people take on projects that may not be viable yet, but are important nonetheless, essentially filling the role that government used to play in R&D.
generalizations are false a lot of the times
Now that they have that success, lacking threats, the value is evaporating.
Lots of people do not care how much others make. If the value is there, world better, game on!
But, getting very large amounts of money and those things aren't really happening and lots of people will definitely express negatives about it.
As they should.
I see links are back in my search results now. But my trust is not back with those links. I have found all the other search tools I need and am using them, encouraging greater success.
These companies with almost impossible to think about type large amounts of money can be doing better by people. Expecting that makes great sense.
And where that is not effective, seeking various remedies does too.
Competition, use of the State, protest, all on the table.
Anyone concerned about the potential impact of those things is completely free to avoid them. Just make sure the value to people is there and they won't really care about the money.
Knee jerk, break them up helped create this net we are seeing consolidated and increasingly devoid of value.
People calling for that again is understandable.
Finally, breakups are super expensive. A real, material threat of that happening makes for a real, material cost and risk assessment, which can justify hedging all of that with better value to the people too.
Oh, come on!
Google is not doing its job well, other could do parts of it much better, but Google is too big to allow others to enter its same space, so we need alternatives that can put pressure on Google more than we need Google.
The "people hate success" rhetoric is so sad it can't even be measured.
search isn't a naturally monopolistic market, because although there are some initial high fixed costs, much of those costs can be roughly scaled with size. there aren't really any other significant barriers to entry.
marginally but noticeably better search results (e.g., pagerank) did provide an early advantage for google, which is how it established its monopoly over time in the first place (combined with the ad auction model from overture, which gave it economic & political might), but that advantage has eroded with its lack of focus on search. google also has a giant suite of data gathering products but it's not clear that those provide a true competitive advantage in search (again, search results aren't much better than the competition).
the other interesting characteristic of search is that its competition isn't necessarily zero-sum. a user might do the same search on multiple search engines, providing each competitor with a full marginal revenue opportunity, rather than only one competitor winning the business (as in classic competition).
because of all this, i'm neutral on the breakup argument but am bullish on the "increasing competition" argument for search.
So yeah, I'm down for more search engines. If there wasn't only one dominant one, SEO wouldn't be such a problem, either, I imagine.
Realistically if I were trying to beat Google I wouldn't be building a search engine. What can you offer that they can't? And if you even begin to compete with them there, what steps can they take that you can't recover from?
Instead, I would focus on beating them first, building search later (if search is what you actually care about). What's Google's weakness? Figure out "the cord that makes them run" (to paraphrase Meredith Vickers in Prometheus) and relentlessly attack that (legally, of course, this is business not cartel turf war). I have no idea what that might be, but Rome fell, IBM fell (well, stumbled, I guess?), so Google must have its Visigoths or Bill Gateses or whatever.
You built something cool, that replicates the functionality (80%). Cool. Good for you. You proved you can do it. Now decide if you really are serious about "beating Google", and figure out an effective way to do that. Search engine is not the way.
Anyway, Google built the (past of) search. If you really wanna do search, build the (future of) search. I dunno, like AR/VR search? Search for FBs AR platform? I have no idea but text search on WebPages is the past. Google already won. Get over it and move on. Build the future, or go build a better steam engine, you know, for like fun and stuff. Because it's cool, but it won't ever beat Google. Sorryz Cliqz
We are all ears on ideas on how to beat Google :-)
But on a serious note; i do not think that the aim is to beat Google because of Google but because on the monopoly they have on the access to the information. And that's why we are building a potential alternative there.
Attacking from a different domain, might beat Google on that area, or let's go wild here, perhaps even on revenue, but would not fix the problem we wanted to fix to the begin with.
"Replicating" 99% of the functionality with a quality to be good enough is a very valid option IMHO. Otherwise if one decides to leave Google, where do they go? If they do not want to advertise on Google, where do they go? There are too few alternatives, there are many names, but most of them are aggregators (full or partial). That said, if besides Bing, there would be 4 or 5 like them (or better), then yes, I concede that Cliqz as is, would make no sense.
Apple is not competition. They serve a lax customer base.
We need a service that would organize the information for us. Instead of us telling the service what we want to see, the service should tell us what's worth attention. Some people care about NFL, so they would be given the most valuable news about NFL. Others care about superconductors, so the service presents them a daily summary of most important advances there.