What If IBM’s Watson Dethroned Google as the King of Search?
wired.com
wired.com
This sounds like it was written in 2007 or something.
Google has been working on question answering for years now, but only exposes it when it is sure of the answers.
https://www.google.com.au/search?q=when+did+jfk+die
(I get a "one box" answer saying "November 22, 1963 John F. Kennedy, Date of death"). You'll note that Google had to derive that I meant "John F Kennedy" by JFK, that I wanted a date, and then retrieve the answer.
https://www.google.com.au/search?q=how+old+was+jfk+when+he+d...
"46 (1917–1963) John F. Kennedy, Age at death"
https://www.google.com.au/search?q=how+did+jfk+die
"Assassination John F. Kennedy, Cause of death"
It's worth clicking the "More info" button under one of these "One boxes". Google will tell you what it regards as "facts" and where it is deriving them from.
(Also, http://www.wolframalpha.com/input/?i=how+old+was+jfk+when+he... is just as impressive)
horror thriller where woman goes to russia meets brother
http://www.imdb.com/title/tt0475937/ - The Abandoned
^The result I was looking for appears first.^
I think it is ludicrous to think that Google can be displace from it position as the global leader of web search in the short or mid term. To displace Google a new search engine wouldn't just have to be better it would have to make Google appear like last century tech. With this said I have to make clear that I am no Google fan or Google hater but I am glad there are other search engines that can at least compete with Google on their national markets (Yandex, Baidu, Naver, Seznam & Yahoo Japan).
I find this only works for very particular, seemingly whitelisted, cases. For example, it does date of death well, but try to ask it the date of any other sort of historical event. It doesn't do the date of surrenders or famous battles, the day Columbus discovered America (it won't even tell me that date if I straight up search for "when is Columbus Day"). "When was the declaration of independence signed" gets me nada, and it can't tell me what year the Magna Carta was issued. It can tell me the day that Hitler died, but can't tell me when victory in Europe was declared. It can't even tell me when Enron went bankrupt, despite the first words in snippet under the first result for "when did enron go bankrupt" being "Before its bankruptcy on December 2, 2001"
I'm not picking and choosing examples here, everything that I tried besides birth and death dates did not work for me.
What Google is doing really seems like nothing but a parlor trick compared to what Watson can do. Maybe they really are doing all the same fancy math behind the scenes, but what is ultimately surfaced to me, the end user, is not impressive.
(Note that although I can't get coherent responses from Wolfram Alpha with natural language questions, Wolfram Alpha can answer most of these things if you don't ask it questions that way. "Battle of Hastings date"? Wolfram Alpha can do that, Google can't.)
It sounds like the author is suggesting dynamic web pages built by Watson that would answer questions (summarise court cases etc). The underlying problem is, where does all this data come from? It sounds like Watson would intepret multiple data points from around the web to compile the information. How can they say this information is correct? Google at least points you in the direction of the information then you make an informed decision yourself if it is correct.
This sounds like something we already have, we have Google and Wolfram Alpha. Problem sorted.
If anything, IBM could replace / merge with Wolfram
I don't see it as replacing search either, people search for websites. What is the end result, that Watson replaces all sites on the internet with it's own dynamic page filled with its own information?
How would Watson choose how to display the information it returns to me? Am I getting the full story?
This article breezes over so many specifics I cannot take it seriously.
Information <--> Actual User Questions <--> Actual User Clicks
It is the continuous feedback loop between the three that makes Google what it is. Without the latter two, Watson is seriously disadvantaged.
On the Google Search side, its algorithms are a lot more than just "the PageRank algorithm". The knowledge graph and its ability to remember and take advantage of context from previous searches are examples of this.
On the Watson side, a huge amount of what it was doing involved identifying keywords and finding relevant information from the clue words. It did not really reason about questions or have a deep understanding of the semantic content of the query. It was hand optimized for the sort of questions that tend to be asked on Jeopardy, which was an impressive feat, but in terms of being able to create new knowledge, as the author suggests? The state of the art in AI is a long way away from that.
The question is -- does IBM have access to the same amount of data?
That being said, I wonder if they're more interested in how their D-wave quantum computer will turn out.
Actually, I needed the answer, since I was reaching into my desk drawer for some wedding gifts and needed to buy some gift boxes (Google helped with that, too).
Watson may have a head start on the AI stuff, but Google is a quick study.
Also, Google is really good at server farms.
†https://www.google.com/search?q=What+is+the+diameter+of+a+kr...
The smallest machines Cray sells cost about $500,000. If you want scalability you gotta pay. 8 nodes isn't a real machine.
Natural language understanding is quite relevant for the crawling and indexing part of information retrieval systems and Google is very good at that. Just look at their quite formidable automatic translation software, which is a by-product of their ability to correctly map natural language concepts to strings.
The thing is: People just don't want to converse with a search engine as if it was a human being. Some library / scientific information retrieval systems tried to go in that direction, which resulted in retrieval systems that were just cumbersome to use.
Google nailed the search engine user interface quite some time ago and Peter Norvig is absolutely right when he says that users simply don't want to ask questions when searching. They're much faster at entering keywords relevant to their search intent because they've learned how to efficiently converse with search engines in their 'native' keyword language.
Hence, passing the Turing test is completely irrelevant for search and information retrieval in general. Even in mobile environments where due to the device's constraints a natural language user interface makes a lot more sense than on the desktop, software like Siri more or less is just some gimmick that in most cases is easily outperformed by more traditional input methods. Sure, asking Siri to 'Show me the way to the next whisky bar' might be fun at first but simply entering 'pub' and the name of the town you're in right now is still a lot more efficient. Again, I think Google nailed the user interface part with Google Now for mobile information retrieval as well. I don't want intelligent machines to pretend they're human. I want them to take a back seat and present me with the right information once it becomes relevant.
If there is one area of search technology where Google has contributed almost nothing, it is the search interface. The subjective user experience is virtually unchanged since the days of Alta Vista circa 1996.
(http://insidesearch.blogspot.co.uk/2011/11/search-using-your...)
> we found that users typed the “+” operator in less than half a percent of all searches, and two thirds of the time, it was used incorrectly.
Check my arithmetic but that means only 1 in 600 searches used + correctly; 2 in 600 used it wrong, and the other 597 searches didn't use it at all.
I think we shall soon see the end to the idea one organisation (based in SV) will be able to service all the worlds needs on a march towards AI - I suspect Google is as an organisation already straining. why not let markets flourish ?
Google works well when there is no clear source of authority and the algorithm can make a judgement. Some searches have a definitive thing that you are looking for and the algorithm would need to know that source to give the right answers. We need more human input which can make search engines that give results that are shamelessly partial to particular sources of information (like patent searches).
Google owns Chrome and Android which is basically 50% of web traffic.
And Google is getting much better at understanding syntax and will continue to improve.
"they could still resort to a simple PageRank-like algorithm as a last resort"