2013 Founders’ Letter
investor.google.com
investor.google.com
I find the most effective way to search is to already know where the most knowledgeable people gather around a specific topic and use google to search those sites specifically. I will often search within the context of content aggregators like hacker news, reddit, stackexchange, and other various forums before I rely on a naked Google search but it's a kludge at best and not something less tech savvy people are going to know how to do. And if I'm completely new to a topic it's often a chore just to even find the place where the experts actually hang out.
I think the approach that Stackexchange is taking is going to be far more valuable in terms of search over the long run than Google. Time should be spent on figuring out how to reward domain knowledge experts and making sure they stay untainted and motivated to share their knowledge rather than tweaking a search algorithm in an endless cat-and-mouse game played with content farms.
For any search where the first result is Wikipeida, the following results don't have to be any good, and there are plenty of terms where the next 20 results are just domains with lots of in-links hosting glossary and dictionary definitions.
Tech was the first area to conduct large amounts of our work online and in the public's eye. As a result, almost all our meaningful literature, documentation, q&a (was mailing list archives, now is just stackoverflow) and "here's how I solved X" blog posts are available for google to scoop up.
> The search engine is working on being able to provide direct answers to questions rather than just a list of results
I dont want what they think the answer is to my question. They don't know my thought process or how I like to research information. I want raw data returned so I can sort and process it how I see fit, not them.
* It only works on chrome, obviously, because it's a chrome extension. That means you will continue to get crappy results on all your other browsers and devices.
* It modifies the links after you load the page so you see them disappear which can be disorienting. I really don't want those search results even making it into the DOM.
* As of now I'm not able to even edit my list (the options menu item is grayed out) and I have no idea why. The quality of the extension is poor at best.
* It has zero intelligence. What would be ideal is that I could subscribe to a blocklist of well known content farms and allow crowdsourcing to maintain and add to that list.
The clear message from Google is that it doesn't really want to support or encourage wholesale blocking of any domains. Google would prefer to be the entity that makes that decision for you, and since these content farms make Google a lot of money don't hold your breath that they're going to start cracking down in a serious way anytime soon.
It would get harder for Google to crawl new content on the internet because they'd be filtering through progressively more clever and aggressive farms to find real pages. I'm sure is already the case, but aggressive blocking would provoke farms into getting worse.
When Google created Adsense it built an ecosystem. While other search engines lost users as soon as they clicked on a link, Google had a fairly high percentage chance that user still made them money.
I don't know if the future looks much better. What Larry Page is saying that instead of display a link of Stack Overflow results he would like to scrape the answer to the query in a way the user never has to go there. The way they will make money from advertising with that model looks even worse.
[1] http://userscripts.org/scripts/show/95205 (looks userscripts.org is down right now)
You might be on to something here with censoring as a service
It would actually be pretty simple. A search bar. If you search doesn't exist, you're prompted to create a new page and add in whatever links you've found. Users could also vote to merge searches, etc.
You'd need some major fraud detection of course but I think it would be possible.
Anyone interested in working on that?
However google actually rewards that.
http://webmasters.stackexchange.com/questions/53703/choosing...
Check several competitive markets in bigger cities. Examples: City Dentist, City Personal Injury Attorney, City Chiropractor. For many the top 3 - 5 are EMDs.
Google just reduced the weight of EMDs, so if thin or spammy content and all they have going for them is the EMD, they won't rank. But an EMD with good content and marketing still ranks well in local.
Adwords is also excellent, but the skill to use it is not well distributed. Advertisers have more obvious incentives and measurement metrics (return on total ad spent / return on investment) than hobbyists do.
It's really basic, but I end up starting most of my searches there, and fallback to Google with a DDG-style !g if I need to.
- [1] http://blog.databigbang.com/letters-from-the-future-challeng...
Let's say I'm in a strange town on a business trip. I've got to eat dinner. I ask Google for a good restaurant in my area. It can just give me a list of all the restaurants nearby, and that's quite useful. But it's tedious sorting through all the ones that I would never consider eating at.
Or, Google could know something about my eating habits, and start off with the restaurants that might actually interest me. That's a more useful result. But to get there, Google had to know some things about me.
Now, may want to be able to choose that tradeoff; certainly Google could do better at giving you controls for this. But the point remains: If you want results that are more useful to you, Google needs to know some things about you.
- Over 63% of the top 100k web sites embed their web bug ( http://smerity.com/cs205_ga/ )
- I have no data, but a huge proportion of web sites use Recapcha in their sign up process. It's impossible to use some domain registrars without revealing the IP address of the probable domain registrant
- I cannot host a web site without Chrome or Google DNS donating some huge percentage of data about my visitors to Google. I'm not sure the mechanism employed here, but I've set up 2 private web sites in the past year where upon having a Chrome user visit, Googlebot turned up within a day or so, despite all efforts to ensure the URL remained private
- I cannot visit Reddit and many similar sites with a cold web browser cache without revealing my use of the web site to Google, thanks to their AJAX library hosting
- A technical user has no difficulty blocking Urchin, but e-mail is a different issue entirely. I cannot find the article now, but one user reported 40% of all mail in his inbox passed through Google's servers at some point, even though he wasn't a Google user.
On the average day, whether I like it or not (and really I don't), I generate tens to hundreds of new records in Google's databases.
Not disagreeing with anything, but in case you didn't know, you can actually disable this in your reddit preferences: https://ssl.reddit.com/prefs/
Choose "load core JS libraries from reddit servers" under Privacy Options.
I submitted a Firefox addon for approval a few years back that introduced a separate cache for all the AJAX libs, but it was rejected for silly reasons
I have actually been looking for such an addon. I too clear cookies/cache/etc on exit in addition to using NoScript and a strict Adblock Edge (block social networks/tracking/not just ads). The only area remaining to be fixed would be the AJAX libs to be hosted locally.
Would you consider making a github page or something so you can provide the addon without being in the firefox addon explorer?
One of the major complaints in the rejection mail (included) was that the add-on bundled copies of all those libraries. So it seems reworking it to have a dynamic list of URLs, and fetching these once on first boot would go some way towards solving that. Unfortunately this further complicates how the redirection works, so I never got around to investigating it
Consider vague long tail queries. Google knocks these out of the park compared to Bing, DDG or others. Let's say there's some old movie you vaguely remember details about:
"movie where evil is in microwave" => Google "Time Bandits". Bing: Nada. DDG: Nada
"dystopian movie with a motorcycle named einstein" => Google "Warrior of the Lost World" Bing: wrong, DDG: wrong
"song about cocaine in california" => Google: Hotel California, Bing, DDG: bottom of page
"star trek episode where spock is possessed" => Google: Return to Tomorrow, Bing/DDG: Spock's Brain (wrong), but nice try"
"previous name of java language" => Google: link to Oak as second item, Bing/DDG: worse result (must click through links to find answer)
"cartoon about plants and a kid with a ring" => Google: Jayce and the Wheeled Warriors, Bing/DDG: Nope
"tv show with scientist named blackwood" => This one is EASY, Google: War of the Worlds, Bing/DDG: Nope
"guy who tried to bomb parliament" => Google: Direct Answer, Others: links
"tv show with alien who has necklace powered by sun" => Google: 1st result, The Phoenix, Others: 3rd or lower
"cartoon car that has auto jacks" => Google: Speed Racer 2nd result, Others: nope
"flying characters similar to thundercats" => Google: Silverhawks, Others: nope
"movie with guy who owns last car" => Google: The Last Race, Others: nope
"commodore 64 game where alien knocks on window" => Google: Rescue on Fractalus, Others: nope
"toy where you program trailer to dump" => Google: Big Trak, Others: nope
You can play this game all day, trying to come up with the vaguest possible query about old things you remember bits and pieces of from decades ago, and I'm often amazed at how vague and obtuse I can get. Google does these queries much better than others. Yes, it fails a lot too, but the instances where it fails, and others succeed are much more rare than the instances where it succeeds and others fail.
But Google isn't an AI yet. And what Larry is speaking about is achieving Star Trek: a computer that can read the meaning of words, understand, rather than just index characters.
That one just comes to mind because it seems like once a month I search for something related to houseplants, and I almost never get good information without having to extensively refine the search, even though the internet is loaded with high-quality sites about plant care, botany, etc.
Google apparently has a "Sambisa Forest Reserve" in it's database, but no "Sambisa Forest"
Submitters: please do your due diligence and read what you post. An article of cherry-picked excerpts from a more original source is rarely a good HN submission. You should submit the original source instead.
1. http://www.zdnet.com/google-a-million-miles-away-from-creati...
2. ‘Google "a million miles away from creating the search engine of my dreams"’
What intrigues me about the Android-equation and the way it entered Googles' playbook is that, indeed, the game is not over. Any other young upstart can arise, almost immediately, and challenge the hegemony .. all you have to do is be willing to commit to a real platform strategy.
I believe the Android vs. iOS challenge for the mainstream is a very vibrant industrial action; echoes of this gameplay can be definitely observed, elsewhere in the F/OSS eco-system. A fundamental difference between iOS/Android: one is closed, the other is open. A new, open challenge to Android could indeed eat its lunch, if the open challenge were won in such a way that it would attract the hunger-economy of the giants.
I personally would like to just have a phone that is 100% open source. I've had both Android/iOS in my life, equally, since inception, and ultimately I find the industrial-baggage load to be pretty high. In the new NSA era, what the hell is the point of trying to hide anything any more .. the new cool is completely open.
I bet it would solve the immense-cruft problem we can easily witness in heft with iOS/Android, as both platforms NIH/NIMBY their way across the 'developer mind-control' table.
Imagine this: someone builds a challenger OS, for free, that runs on everyones' phone, regardless of manufacturer-lockout/lockin. Methinks there are ways to do this under the radar right now ..
Then you have to think that in 3rd world countries acquiring a smartphone means success... and of course that smartphone is going to run on android.
Stop thinking that Europe + America + Australia have more population than the rest of the world :)
I don't disagree with any of the statements you have just made, and nothing I said in my previous comment contradicts that - so why the attitude?
The perfect search engine would always take you to exactly what you were looking for immediately.
Which means, no amount of advertising would ever be useful, since you wouldn't need to be advertised to - you'd type what you wanted and it would show you exactly that.