Google acquires Metaweb (Freebase)
googleblog.blogspot.com
googleblog.blogspot.com
The problem with the semantic web is that many need to embrace it. Many people need to tag text with these "bar codes" (uniquely identified entities). That can take a big effort and there has to be a ROI for this big undertaking. The other is that there is no standard. Well, Google just solved those. With a dominant market share, you don't need someone to agree on a standard, you just force them to--or else they lose out to competition. And as far as the ROI in tagging web pages? Well, what's the ROI on SEO? This will bring about a new form of SEO, except that Google can now undercut many of the search results and answer many of the queries directly--so that'll get interesting... and I'm sure Wolphram Alpha certainly agrees.
Google was also smart to buy Metaweb in order to give web app developers a good reason to use their entities and just FB's open graph entities.
Congrats to the Metaweb team! Freebase + Wikipedia are two of the best gifts to humanity.
Of course, that leads to the question : is that ontology actually relevant or is it just important that the data is structured?
Sure, if you weren't concerned exactness and lack of ambiguity, you could expand the world of triples into a giant, poorly organized collection of information. It would be kind of like the web. The approach "works" but we, uh, already have the web.
Also, the Doctorow document excellent. Anyone expected naive metadata to be extensible should have a reply to it.
Inferencing solves this problem.
However, if you think of it "vocabularies for people to tag their own stuff in a structured way", which Google can then index and traverse, then it's more realistic.
That said, the Semantic Web people are indeed guilty of hyping this technology as being able to "reason" on its own.
What about Facebook's new metadata? What are people in the semantic web area saying about that?
This sort of thinking seems to be common in people who like the idea of the semantic Web but who are pessimistic about its implementation. I'm not sure it's going to be the case.
As we've seen happen with other technologies, I suspect we'll see a MetaWeb style approach of "deriving the barcodes" from existing and unformatted content. This will not be a 100% accurate process, but will be "good enough" to make the semantic Web a realistic and large scale underpinning to the next generation of search systems.
For example (for those that are unfamiliar with the richness of this data), visit: http://www.freebase.com/view/en/y_combinator
And there you'll note that Y Combinator is mapped to the official site, it's Wikipedia page, it can tell you who the founders are, etc. If you link to Paul Grahm, you can then find out what his personal site is, and so on.
I do hope Google will do something cool with this acquisition.
There is a lot of cruft in Freebase, but with some manual effort and some automation, it is a good source of a wide variety of information. Depending on application, DBpedia and GeoNames are other good resources for structured data.
I myself learned about it from Stephenson himself during a presentation he did for the book in the now defunct Cody's Books in Berkeley. After 2-3 years of active growth the Quicksilver wiki disappeared from metaweb's site. I was wondering if Google will restore the wiki too?
Metaweb briefly mentioned at the end of this article in Wikipedia: http://en.wikipedia.org/wiki/Quicksilver_(novel)
Glad you got your payout, guys. Hopefully now the full power of Google's infrastructure can make Freebase fast and enormous.
MQL is sweet (e.g. give me all of Tom Cruise's movies since 1995 that have cost over 10million dollars to make), but what good is the data dump if all I can do is resort to simplistic SQL queries? MQL support is just as important as the data itself.
I'm sure that I'm overlooking something simple as usual (shameless self-deprecation reference in an attempt to hold myself less accountable in the event that there's an easy solution available but I was just too lazy to find it).
We hope to be able to provide rdf/rdfa links at some point, because, hey, the semantic web is a darn cool thing :)