Google Search By Image
techcrunch.com
techcrunch.com
What will be even more interesting is if they release an API for it in the future. Sites like imgur and reddit could then suggest if you're uploading or submitting a similar image to one that already exists.
I wonder if Google's will work the same, or if it will be possible to find other images that are similar, but not based on the same image? That's what I really want.
As a rough analogy, consider a textual example: Find a sentence similar to "It is a truth universally acknowledged, that a single man in possession of a good fortune, must be in want of a wife." Now, if you enter this sentence in Google, it retrieves documents that contain it, in TinyEye fashion. What other similarity is desired? Should it retrieve essays on Austen, on marriage, 17th century English literature...?
If you are interested in image similarity search, check out the Pascal challenge (http://pascallin.ecs.soton.ac.uk/challenges/VOC/voc2011/inde...). The advances in the last 5-6 years on object detection and visual feature extraction (which image similarity relies on) is amazing.
- somebody put text over a nice image and I want the unmodified version
- searching for bigger, better quality versions of an image (e.g. wallpapers)
- finding other images from the same author/gallery (since it links to the sites that hosts the copies)
- finding the name of the movie, person or object pictured (because copies will be hosted with different, probably meaningful names and in pages with subtitles)
Often this would happen: Client or user finds an image through Google Image Search on another blog or website, with unclear copyright. Then the client or user would upload this to the server. Then the client, or me, would receive a letter from a (GettyImages) lawyer: If we would please pay for the full licensing right of that 160x160 pixel image.
Thinking about it, I find accommodating to these services can lead to nothing but trouble: Either legal trouble, or hit-and-run users stealing your images, because you paid for higher resolution.
I hope I can separately block this Google service from Google Image Search. Although Google Image Search isn't as good to webmasters as it used to be (especially for those that rely on advertisement clicks) and the users it can send can be negligible: Ranking higher in Google Image Search seems to correlate to ranking higher in Google Web Search.
If it is part and parcel of Google Image Search, I might reconsider my robots.txt directive for Googlebot-Image. They just now opened this up for the public, but it is likely they are already using this internally to gauge (media) quality factors on-page.
P.S.: It would be interesting to see what happens when Google doesn't partner up with GettyImages or iStockPhoto, like in the early days of TinEye you could abuse that service to find the same stock images without watermarks, on the sites of people that already paid for that image.
P.P.S.: Now you can add RDFa or Microdata to mark up your images with a copyright statement, what would happen to sites that host copyrighted images, tagged "not for reproduction"? Google should be able to find the canonical image and "punish" those that don't comply with its copyright.
So they probably don't have much to worry about.
The best results are under 90% accuracy, which sounds pretty good, until you realize that random chance is 50%, and for recognition ("who is this person?"), you're essentially exponentiating that 90% by the number of different people you want to recognize.
It does exactly what you ask for.
If that is right, Google will be able to retrieve different photos in which the same object appears, whereas TinEye only retrieves the same image, with or without some changes. So, they're quite different beasts.
Can't wait to see if I am right...