On the growing, intentional uselessness of Google search results
neosmart.net
neosmart.net
There is not a single day that goes by that I am not searching for something specific and particular in google and am treated to pages and pages of search results that are missing at least one of the terms, thus rendering the results useless.
The worst part is, the strikethrough "missing: search term" identifier does not always appear, and you click through to a page that is useless without knowing it.
My habit has become to immediately ctrl-f on the resulting page and look for my terms so I don't waste my time.
Further problems:
- "allinsite:" is just a toss-up whether it is respected or not. Who knows why, but it does not fix this problem.
- "quoted strings", such as for programming or naming conventions, are completely ignored and are useless.
- there is no "not" operator, which is desperately needed.[1]
The only function that actually works as advertised is the site: prefix which limits searches to that particular website. I won't be surprised when they break this too, because it's not producing enough search-result-revenue.
I am not a teenaged kid searching for Justin Bieber and perfectly happy with whatever "relevant" or "related" results pop up. I am a professional. I am an engineer. I need tools that work, and google is shit as a search engine.
[1] https://support.google.com/websearch/answer/2466433?hl=en
When you use a dash before a word or site, it excludes sites with that info from your results. This is useful for words with multiple meanings, like Jaguar the car brand and jaguar the animal. Examples: jaguar speed -car or pandas -site:wikipedia.org
That is the definition of not working. The point of the quotes is to allow searching for literal strings.
You have elucidated the entire problem. Unless, of course, you're just searching for movie star names, in which case the rubbery nature of the search results really doesn't matter.
It appears that -negative searching does work, which is amazing - that's nice to know. But it immediately reminds me of a terrible, glaring problem: searching for command line switches. You cannot effectively search for things like:
rsync --ignore-existing
... since they are interpreted as -negative keywords, even if you quote them.
You can also remove the double dash and search for rsync ignore-existing
I appreciate the point of the posted article, but for the most part I don't actually have any trouble using Google as a programmer. You just need to format your searches correctly.
I believe s/he makes a good case.
Google provides no concept of time. Best case conception of time seems to be either provided by the site directly, e.g. "article date" or something to this effect, or t0 = when google first learned about the page.
Top anser for: "nodejs" + "2016" mongo api
returns top hit: 2015, 2nd hit 2014.
and that I can't give it context myself:
Qouted strings can't possibly work. There is almost certainly an algorithm in the background to provide confidence that they could work before being run. Even with the best case optimization, how could they check every page in index for:
"does google respect qoutes?", I mean even if it finds pages with "google", "respect", "quotes" then runs query on all of them, how could it be fast enough for a human.
In above case, it is possible, however smaller phrases with common words can't ever be respected.
> Google is a natural language search engine
Eeryone has a google search profile. Allowing someone to calibrate theirs would provide them best results. Outline down thread, I explain why it would be nearly impossible for google to do this better than you.
Your appraisal of the not or "-" operator falls into the same problems in terms of time. I can't define my percentage of crawl index that is authoritative, have to have the applied "-" operator applied to a larger base than google and this may not be possible in limited time.
One idea I had is that google could actually take longer, and update itself on the rolling basis. So immediate best guess would come up, and then it would continue running in bg and rerank if you waited another 3 seconds for better results.
Not true. It can be achieved by simply adding the
-
Operator to a search. For example, to get Google Search results about Tesla, minus all the stuff on the internet about the cars / motor company try... tesla -"tesla motors" -carrsync version freebsd 9.2
First, note that the first four results (for me, anyway) do not contain the string "9.2" at all.
Second, note that google does not give you the "missing: 9.2" notifier on those in the results page.
Later results (for me, 6 and onward) do contain all of the terms. Google decided that search results that did NOT contain all of my terms were more "relevant" than ones that did.
Bonus laughs: changing the search to:
allinsite:rsync version freebsd 9.2
does change things up a bit, but results 1,3,4 and 5 still don't contain the string "9.2".
I am not waiting in line with my girlfriends and searching for Britney Spears Songs. Related pages do nothing for me.
They are bad results.
Google search will generally learn your preferences and get better at vending you the results you can actually use if you're logged in so it can build that search history for you.
I always assumed that such cases mean that the webpage changed since Google crawler's last visit. Did you try ctrl-f'ing cached version of the webpage in such situations?
https://en.wikipedia.org/wiki/Amazonen-Werke
> the company was founded in 1883
I'm reminded of this:
I don't see how you can complain when you have the solution to the problem, don't use it, and are an incredibly niche (<1%) case.
And there will not be any time soon. Writing an efficient crawler for what we call the "modern" web is not something a small or even median-size company can pull off. Google enjoys a tremendous competitive advantage: people specifically optimize webpages for what it can and cannot do. So any newcomer to the field will have to replicate tons of technologies Google had years to perfect (in addition to solving problems like storage, search logic and bandwidth management).
Besides, regardless of what you do, you would need to have tons of storage, bandwidth, CPU power and a high-availability infrastructure.
A novel solution could be designed and implemented by a small company but no one dares.
You're not going to get the first unless you're Facebook or Amazon or God, but maybe you can build a smarter algorithm. You are up against an army of some of the smartest computer scientists and mathematicians ever assembled -- but what you have going for you is a complete lack of inertia or legacy. You could try crazy things that Google might not, because they won't think it'll work. If you get lucky, one of those blows up. But you have to get very lucky (this is the Innovator's Dilemma in a nutshell).
Anyone want to give it a shot?
Or a limited search space. For myself, it could be HN, SO and Wikipedia.
Content can even be static and downloaded once or regularly.
Entrenched businesses are RARELY displaced from their top position. Instead, what usually happens is the world changes around them, and the entrenched business is ill suited to compete in the new world. Nobody ever managed to seriously challenge Microsoft for desktop/laptop OS dominance: https://en.wikipedia.org/wiki/Usage_share_of_operating_syste.... The issue is that desktop/laptop OS dominance is no longer as important as it once was.
[EDIT] And as an addendum to this fun tangent, even if the AI was capable, in principle, of reading billions of webpages fast and well, we don't know the power requirements would be for this. This hypothetical small company may have stumbled upon the right algorithms and the right training data to produce a true AGI, but they may simply not have the hardware or the engineering know-how to scale it up across multiple processors. Or scaling it up may require too much (i.e., more than the company can afford) data bandwidth if the robot is controlled. Again, it depends on the details of how they arrived at the AGI.
memex-explorer + blockchain is the model. A company can not do it. A company can design a platform, and many companies can sell on the optimization & information marketplace to fix problem.
local cache ===> personal cloud cache ===> centralized crawl repo ===> small specialized data cache sold into marketplace ===> faiil
heirarchical crawl index and DNS rebuild(or similar model) will lead to search platform and fix discovery, monopoly and monetization.
note that Google has solved a HUGE problem, but their task is now impossible. You can't simply use a textbook and some booleans and quotations(which I am not sure they even respect) to deliver results for a billion people.
Individuals and the market will calibrate their own results.
>elastic
hahahahahahaa
you have not used Elasticsearch I see :)
Eh. I switched my default to Duck Duck Go and I'm pretty happy with it. Not quite as magical as Google was in its heyday but then neither is Google. Set up keywords for your searches ("g" for Google, "b" for Bing etc.) and they're all just a keystroke away anyhow.
And if you're too lazy to set up your own custom searches or you're using a borrowed machine, Duck Duck Go has some slick built-in "bang" searches: "!imdb aronofsky", "!msdn system.diagnostics" and so on.
Though DuckDuckGo does do its own indexing, AFAIK it is limited and they mostly rely on other search engines (Bing and Yandex most probably). So it's more a meta-search engine not quite in the category of Google.
Today, I just posted a Show HN for a search engine and feed reader I've been working on. It also has "bangs" except they are called activation codes and start with a ?, like "?jq appendto". It's easy to add your own "?" handler, as it doesn't require any approval to do so. I'm just getting started so I would love to get any feedback.
Link to the Show HN: https://news.ycombinator.com/item?id=11174127
>OK, confession time: the article linked to in the fourth result – the one that says “no retina support […] Deluge” actually talks about another app’s lack of retina support on OS X, but just go with it!
That's the result he's using to say the results are worse, that that result should be higher, but says in the footnote that's not even a result relevant to what he's looking for. What am I missing? He's saying the first results are worse, but they are better results. They return the actual app he's looking for, where he could presumably find info on retina support, not some completely different app's lack of retina. All of the links in the second screenshot are wildly irrelevant to what he's looking for.
In fact, if you look at the results with "deluge.app" and "retina" in quotes, there are no results that he would be looking for.
I think this article misses its own point: it's not that the results are worse, it's that what he actually wanted was for google to tell him there were no results, not to realize there were no results and so expand his search for him.
There are two choices that google can make in this type of situation, show you that there are no results for your search terms (helpful in this situation but not when there are obvious synonyms that do have results ("help" vs "aid" example in the article)) or google could expand your search to include synonyms or not contain all terms (helpful when there are lots of useful results close to your terms in search space, not so helpful when you're looking for something very very specific).
I'd imagine it's incredibly hard to tell the difference between these situations with just a few words as a query, but if I had to bet, I'd guess they probably hit some useful level a large portion of the time. It's not clear the problem is universally solvable, however, and going back to add quotes to terms is annoying to type. We're probably never going to get the beloved + operator back, so how about just making verbatim search much quicker to toggle, google?
(I'd settle for a more consistently triggered "Search instead for [original search term]" at the top of the results)
"The" post that I wanted to be first place had the perfect summary in Google, discussed deluge on OS X, talked about the lack of retina for a few different apps, and explicitly mentioned a few without retina support but did not outright include deluge in that list of apps without retina support. It was the most-relevant result in that it actually discussed the topics being searched for. It was, for all intents and purposes, the correct result that should have been returned - only pedantically it did not provide a point-blank answer to whether or not deluge itself was retina-ready.
I agree with you 100%, the results in the first image which do include all the search terms are more relevant than the results in the second search. But Google, for some reason, chose to prioritize the results that did not have all the search terms over those that did. Now from the results in the first image, the first of the displayed results that did use all the search terms (i.e. did not say "Missing: deluge") was the most-relevant of all the results that were obtained from either listing (important pedantic note: whether it actually answered my original question or not does not detract from the fact that it was the most relevant. Because the other links neither answered my original question nor were relevant to it.)
I think a comment by "Robert" from the blog post (if I may re-post it here), best summarizes my disappointment:
Imagine if I told you I have someone who might be the perfect soulmate for you, but unfortunately because the pool of candidates for “perfect soulmates” is so small, I’m also including people that are maybe compatible with you or maybe not – a kind and thoughtful act, on my behalf…. And then I proceed to introduce you to these latters while holding back the perfect match until a random time that I saw fit?
Regardless of whether or not the suggestion for potential soulmate ends up working out, the fact remains, you don't say "I have a result for your search query, but let's look at these definitely irrelevant results first"
If you want to over-analyze this, let's look at the "blurbs" returned by Google for the search results:
1) Deluge's main download page; blurb: open-source cross-platform torrent client. Site includes screenshots, FAQ, and community forums. MISSING: RETINA
2) Download - Deluge. Latest release <url here>. Release... <link to ubuntu.png here> Deluge.app. MISSING: RETINA
3) Installing/Mac OS X: A deluge package is available which works on Mac. MISSING: RETINA
4) From Linux to OS X: Meet your new apps: OS X Mount Lion ships with an app similar to AppX and AppY ..... [sic] It has one notable shortcoming: no retina support .... [sic] There are plenty of great Bittorrent clients on Linux - Deluge, KTorrent, Transmission, etc.
Of these four results, only one specifically talks about Deluge.app and Retina. It's the fourth result. Based off these four blurbs, which do you think is the right page to click on with the highest probability of answering my question? 1) The product main page which I know, thanks to Google, does not have the word "retina" anywhere, 2) the product download page, which I know, thanks to Google despite the completely useless blurb, does not contain the word "retina" anywhere, 3) instructions for installing on Mac, which thanks to Google, I know does not contain the word "retina" anywhere, or 4) a page discussing a variety of apps available on OS X, including explicitly by name, Deluge, which also talks about the retina support of one or more of the aforementioned apps?
I clicked on number 4. A page that talks about Deluge and other torrent clients that are available on OS X and lambasts an (unknown from the blurb) app for not having retina support would ideally be the page that would contain specific information on whether or not Deluge has retina support. It didn't provide the direct answer I was looking for. But it was a hell of a lot more relevant than the first three results, and Google knew it.
Addendum:
Oh, and about deluge.app not being in quotes: that's a lesson learned the hard way. Mac apps unfortunately do not have "unique" names. Pages. Numbers. Deluge. etc. People often append ".app" to clarify their meaning for SEO purposes, and I know that Google indexes "foo.bar" (sans quotes) as "foo bar" (again, sans quotes). Ironically, the only "word" of the original search query that could have been logically dropped is "app". But odds are that a post discussing Mac apps would contain the word "app" or "apps" somewhere. It's not fair to put "deluge.app" in quotes to provide a counterexample, because I knowingly and deliberately did not place it in quotes in the first place, because that's the one term that I do not require to be present verbatim.
Also, this is just the proverbial "straw that broke the camel's back." I run into this problem many times on a daily basis. This is just the concrete example that triggered the post in question, and for which I was able to obtain screenshots of the different variations so that the situation could be properly documented.
Doesn't that basically describe dating? I found my wife because the pool of "perfect soulmates" for me was basically zero, so I figured I'd take a risk and expand my definition of "perfect", and then discovered that I liked what I found.
Back to the topic at hand - it's a bit strange that a page that gives you the wrong answer is the right page because it answered you. It'd be like if you asked "What city is the capital of Kansas?" and I answered "Kansas City" because it had both the words "Kansas" and "City" and both of them are Capitalized, even though the answer is actually Topeka. I'd think a better answer would be "I don't know, but here's a list of state capitals" even though it's missing the words "Kansas" and "City".
https://www.google.com/webhp?q=goat#safe=on&q=goat
Yes a crude Urban Dictionary definition was above the Wikipedia definition. Maybe they think they know my sense of humor and pop culture, but I really wanted to learn about actual goats, the animals, because my 2 year old daughter enjoys them so much at the petting zoo. Not about how to arrange my junk in a particular way.
Searching for "Apple" used to display just pages about the company, not the fruit (even when the company was almost out of business). Searching for "samba" would display just pages about mounting windows filesystems in Linux, not the most important musical genre of a whole country.
Nowadays they improved their algorithms and this bias isn't so strong, but it always important to remember that what is important in the cyberspace isn't necessary important for the whole world.
However, there's a particularly interesting case:
https://www.google.com/#q=tsla
I don't know about you (Google personalizes results to some degree) but the first result I see is Yahoo Finance. Google Finance is second. Why is Google promoting Yahoo over their own product?
This is the problem. Search isn't a democracy. I lookup the results I need, so not giving me filtering ability makes no sense. An engine that does what google does is an amazing achievement, but no longer makes sense as a model for the exact reason you gave as a defense.
edit: If disagreement could be verbalized it would be super helpful to me. I have been thinking about this issue a lot, and often I see people say the same thing as me:
* single website controls almost 100% of english language results
* limiting of images/video from results
* incorrect/wrong results
* limited respect for boolean and quotation operators
* high ranking sites are not authoritative, e.g. w3schools, wordpress automated blogs.
* no way to filter at all
But then down vote my conclusion:
> searches could be filtered and parameters set by user.
Really would be useful to understand my logic error, or if I have missed something.
Qualify your searches and modify them if needed. You were looking for information about goats. You should qualify "goats" with some form of "information about" statement.
"goat" + "animal" makes the wikipedia page for goats the first search result.
"goat" + "facts" gives you countless trivia pages, information, videos, etc
Alternatively:
goat -"greatest of all time" -"goatse" will search for "goat" without including "greatest of all time" or "goatse".
No concept of time. Best case conception of time seems to be either provided by the site, "article date" or something to this effect, or t0 = when google first learned about the page.
So yes, "goat" + "animal" will return your results. Try:
https://www.google.com/#q=%22nodejs%22+%2B+%222016%22++mongo...
Top anser for: "nodejs" + "2016" mongo api
returns top hit: 2015, 2nd hit 2014.
and that I can't give it context myself:
I am on Mac, but my pc is broken looking for windows info or don't include Alexa1000 links as authoritative. million short (i believe) removes the Alexa1000, but not their link authority.
Also, [neverShowWordpressSite unless traffic >3million unique] some larger news sites are actually built on wordpress like bloomberg. But the point is that I would delist by technology, and filter by time and tweak my authority parameters.
However, if google let you do this it would exponentially compound the difficulty as the algorithm would exist on both sides of equation.
As for filtering by backend technology - I'm not sure that would always be feasible. While it may be possible to filter default Wordpress sites, I'm not so sure about sites whom backbone architecture may not be known or publicly available.
E:
As for your computer issue, try searching for "problem/error name" + "solved" rather than "how to fix" + "problem/error name".
I don't actually want to filter backend technology, but I would like to communicate to the search engine that I do not trust (nor want to have returned as a result) any wordpress, blogger or medium website, and I want their rank to be negative.
That is another extreme example, however to discover new things is hard and to find useful information, when communities of bad actors have spent years incentivized to rank higher but not produce quality, it could be easier to simply delist everything and gradually add websites you trust to have authority.
If most people searching for "goat" want to know what the word means in slang, that should be the top result. Possibly you could argue that you want a personal search profile that knows you value Wikipedia higher than other links, but for the default case it feels like optimising for highest odds of success makes sense.
Everyone wants a "search profile" except they would like to control it and how it is applied as it is, for most people, their most important interaction with a computer, e.g. how they access information.
Currently, in some respects, that is out of a single silo or set of balkanized silos.
This will not be true in the future. One place can not dictate information flow for world. Plus, Alphabet has better things to do
As a next step, privacy issues aside, what if they "profiled" you by the types of things you search, and tried to guess what you need based on other people who "think like you"?
For example, I'm a programmer, and if I search "python", I'm probably searching for something different than a biologist who is researching reptiles. This would be fairly obvious to decide based on the other types of things I typically Google for.
I'm sure Google is probably already researching how to do this, though. It sounds difficult to me though because of the sheer number of models you'd have to train and store, and then figure out how to run a distributed index on. It might be more feasible to create some small set (e.g. ~1000ish) profiles of "types of people" and then match you into one of those types. This could also mildly alleviate the privacy issue as the profiling could be done offline on the client.
Google can never necessarily know what you want and can never truly know you achieved your goal, so you could not train it properly.
Not only would you need to discover what profession I am in, assuming you had fully updated linked profile, etc. you would need to build a comparable universe of like minded people and calibrate.
Then, you would have to assume what inputs are similar in that they have same/similar parameters and expect similar results.
Then you would have to assume which link I clicked was the answer, for every person who did this same thing.
Then you would have to discount your bias as an engine, because you provide the top results to me and (for now) people trust the engine so they typically have a false choice of the first 5-10 things. If those 5-10 things are wrong, whole model is in error to extent it is wrong.
Any one of these would provide error and the cascade leads to larger disparity. Google IS SO AMAZINGLY GOOD, it has actually managed to make this not a problem for a very long time.
They do have some amount of confirmation. All of their search results are redirect links, so they're tracking which links you click on. Based on the timing of those clicks, they can tell if you clicked on a result, left that site a few seconds later, and then clicked on another result further down the page, which probably means the first result didn't give you what you want. It's not perfect but it's still potential training data.
If that site has Google Ads or a Google '+1' icon, they can get slightly more information about how you spend your time on that site. I don't know about the legality of this but it's technically feasible.
Because that's how lawsuits happen. They let the search engine run itself. If more people are using Yahoo Finance over Google Finance, the search results will reflect that.
I genuinely believe they aren't dumb enough to open themselves up to monopolistic behavior lawsuits. The Microsoft lawsuits weren't that long ago and I'm sure Google is wary enough to make some attempt at avoiding a repeat.
I have little reason to suspect Google purposefully toys with their search results to promote their own products. Given the quality of their products - I'm more inclined to believe that when a Google product is the first result for a name/search, it's probably because people actually use/enjoy that product.
As an example, Google Maps vs other "Maps". While Google is certainly trying to join the ranks, I find other "Maps" to be entirely unusable with terrible UI and am not surprised in the slightest when Google Maps is the first result when looking for directions.
It could also be you use Yahoo Finance more often and thus personalized results had it listed first. Google Finance ranks 4th on the page for me for that search result.
For what it's worth, Bing and Yahoo also list Yahoo Finance first, if searched for "TSLA".
No they don't. They're a front-end for Bing search (at least until 2019 and in everywhere but Japan, which IIRC is a front-end for Google), but that's being nitpicky. :)
FWIW I've always had an issue with companies being punished for simply being better than alternatives. That includes Microsoft's advertising IE, even if IE at the time wasn't the best browser. People were free to use IE to download a better browser, so I never saw the issue with providing IE as a default. Linux was free, they were free to buy a computer, set it up themselves, and install Linux on it. The fact that they weren't choosing to do so should not result in Microsoft being punished.
However my beliefs and precedent set by previous law (even if it isn't legal precedent?) is still enough to have Google play it cautiously. Especially since there has been threat of such lawsuits if they were caught playing with their search results to advertise themselves and "kill off" competitors.
So maybe it's based on location? Ranked higher based on browser, IP, or something else?
On a desktop, I see what you mean; UrbanDictionary is indeed at the top of the plain-text search results.
Not going to defend Google at all. But for myself, if I'm looking for a Wikipedia-grade introduction to some topic, I usually just search Wikipedia directly. Wikipedia has become so ridiculously useful that it is the top result of a Google search half the time anyway. Wikipedia is a better human-curated web index than any of the old-school 1990s human-curated web indexes ever were.
Urban Dictionary: goat
www.urbandictionary.com/define.php?term=goat
Urban Dictionary
Tucking back your balls and dick, then bending over thus
resembling the back of a goat...according to the rules of
the game, the person who looks gets 4 kicks in ...Generally, I think that the number of people who want the Urban Dictionary result is higher than the number of people looking for the animal, so it's hard to fault Google for this.
What would be ideal is to have something like DuckDuckGo's disambiguation bar that included the Urban Dictionary definition of goat:
https://duckduckgo.com/?q=goat&ia=meanings
Unfortunately, it's not there right now.
I think someone is buying or gaming search results. When I'm in Paris some results take me back to a tour company in central Paris even if they have nothing to do with what I'm looking for.
This has forced me to go back to search aggregators like DDG. I miss the old Google. DDG can be a little too broad with the results and then I have to load it with filters where old Google would kind of get it right away.
Good thing is that with DDG, I can simply use the !g command, and Google won't know I'm making the query from Brazil, and won't be able to localize it at all.
Another that I've more recently noticed is conversions on mobile. I went from being able to type any approximation of "ounces" "to/in/-" "pounds" and getting a number right away to having to click on one of the results to get the number. It seems like backward from Google's normal MO so I don't really understand why they'd do it but I definitely have consistently had problems with that especially more recently.
Not sure that it's the right decision, even if it helps the common case, but it may be. And I say this as somebody who's not particularly fond of Google as a company or its policies.
Does deluge bittorrent support retina display?
The 7th link result for that question has the text "Deluge has updated their program to support the Retina Display" in it's text summary. That's good enough for me.
It would still be nice to have an 'advanced' search mode that was more strict, allowed advanced features, and still took advantage of Google's talent and infrastructure.
https://www.google.com/search?q=deluge.app+%2Bretina
Or in verbatim mode:
https://www.google.com/search?q=deluge.app+retina&tbs=li:1
His article says that the answer is in the 4th result, the first that actually includes both of his search terms. But if you read that article, it's actually saying that Twitter's app has no retina support, and Deluge is mentioned elsewhere on the page, with nothing about retina support.
Edit: Wait, you work for Google, right? Am I the one that's lost? Did they switch back to supporting +term?
But the odd part here (and in line with original post) is that there are pages in Google's index that include both terms: https://discussions.apple.com/thread/7169461?start=0&tstart=.... It's just that they don't start until 40 or so results in.
Searching in Verbatim mode without other modifiers seems to create the response that the original post wants: https://www.google.com/search?q=deluge.app+retina&tbs=li:1
The autocorrection and removal of terms accomplishes the same thing.
While I think these measures may be partially to help users, I think they're actually mostly cost-saving measures on Google's part.
Now this is where a bit of guess work comes in but I'd say Google correctly deduces there is no such thing. Even the best result for both terms just talks about some app not supporting retina, and from the looks of it does help with what other client you might want to use if switching to OSX when you were previously a deluge user on Linux (possibly when dealing with retina). But that's not deluge. That's a useless result considering your primary intent was finding deluge. So Google chooses to give you results that might get you what you want over results that (correctly) only disappoint explaining there is no such thing. So in the first three results Google correctly decides that, to get you a relevant result at all it needs to omit the "retina" to get you results that might possibly get you what you want - deluge despite not having retina support - over just giving you the results that are relevant but definitely won't get you what you want. Your query had no results that would get you what you wanted so Google tried to alter the query and see if it could get you something useful anyway.
I think Google trying to give you results that might be hits instead of giving you disappointment in the first place is very sensible behavior.
One of the reasons why Google has looked so smart is that it has leaned on Wikipedia. I have a kid in middle school and definitely one thing you learn there is you can look like an expert on any subject by consulting Wikipedia (i.e. "what calibre ammunition is used by a tommy gun?") To an extent teachers encourage it, but it can tend towards plagiarism.
The Wikipedia page is generally a safe bet for relevance but it may or may not be a quality answer.
I think today they may be like the middle school student who is learning the tricks to not look like a plagiarist.
Remember that Google search and advertising is actually in big trouble -- there is a reason why they renamed the company to Alphabet. The 90%+ market share they have in many countries is unsustainable for many cultural reasons and they have to diversify.
If someone just wants to know where birds go when it rains, an article on bird habits is less relevant to them than an article that directly answers that question.
Maybe google has become like robocop receiving 100+ prime directives. All these low quality, noisy inputs are simply driving the engine bonkers?
https://www.google.com/search?q=%s&num=100&tbs=li:1&filter=0
aliased to "v" for "verbatim". Here's instructions for adding a custom search engine to Chrome: http://www.swestwood.com/blog/view/fast-searching-chrome* Search can never be decoupled from the browser.
* search is worse and discovery is hard.
* Brave Software needs to focus on building security by design which means a search engine AND browser.
The memex-explorer model is the design of the google killer. Likely why it was abandoned.
Browser wars began ~6 months ago.
The next larry page & brin, WILL write a search engine.
Edit; It is impossible for google to use human language and (questionably respected) boolean operators as the only filtering mechanism For hundreds of billions of pages.
Staggering how well they do with just 1 textbox and image & video tabs, but new players wont have to follow the same roadmap.
============================
Please Explain Downvotes
============================
I am trying to ascertain how I have been thinking about this incorrectly. Often, downvoted but no one explains why. How am I thinking about this incorrectly, what am I missing?
Trying to find time to continue work on something a bit more polished but it is not ready. This is my response in another thread, either ctrl+f my un or "-2" and you will see response to similar query.
Idea is "so obviously crazy", mentioning it publicly for several months and begging people to consider it and work on it has been met with down votes or completely ignoring it. Which means, if you realize it now and start working on it, you will have ~6 (I think less now though) head start as people catch up with it.
edit: Because I don't think this makes sense as a model for search. http://imgur.com/Gz7hXY7
The image (really confusingly) illustrates the searchflow model:
Broweser ==> Google
Goolge ==> Results
results ===> Hacker news or another aggregator
aggregator ===> your fav. sub community (because gooigle discovery sucks)
subreddit helps you find links you want
in those links you find information
==============================
that is what a manual crawl feels like, and how many people use the web.
Thanks for checking out my site. If you have any questions or have any trouble please let me know at help@solveforall.com. And maybe we can explore brainstorming/collaborating since you've clearly put a lot of thought into search.
You've not explained the relationship between security and search either (unless you mean privacy?)