Facebook’s automatic alt-text for images
facebook.com
facebook.com
Not really. From a technical point of view there is nothing impressive about this, except that they spend the time and money to collect all the training data. Also that it says "this image may contain:" tells you something about how accurate this actually is.
Google Photos is a lot smarter in that it can identify what a set of photos contains. As an example I noticed all pictures of my graduation (which had no metadata saying it was that) have been grouped into a 'Graduation' album. It's also kind of scary...
If this is about the general privacy issue then it's back to the same old thing: don't share things that you want to keep private. Of course people's actions prove that they get more value out of sharing this content then is worth the privacy of not doing so.
It's not the tech that is bad, it is the fact it is in the hands of for-profit companies, who are willing to sell your data just to make their bottom line.
The description might be "two people, outdoors, at beach" while the content could be "Awesome beach day today with my best buddy!" with X tagged in the image, geolocated to "Y beach" (either through your phone/camera's geolocation feature, or the poster specifically tagging a location which is something Facebook encourages), posted at Z time. Which do you honestly believe is going to provide more information for targeting? Hell, even if the content is "It's a beautiful day out!", you'd still have far more information from the metadata (tags, location, time) than the vague description the auto-generator could possibly provide.
Tech keeps getting better and Facebook keeps getting more data and putting it together results in better insights. There are several other major companies and government agencies who can derive just as much, if not more, insights about people through data. This is just natural evolution, not a major revolution.
> but this is surely just a cover for being able to take targeted ads to the next level.
As someone who has built 3 adtech companies, this image info is nowhere near as good as the current posts, relationships and likes that users have already entered. Maybe someday it will be able to accurately know what's actually in the image and also derive some intention and relations out of it, but today it's nothing special at all.
> You could get an unprecedented view of almost anyone's life from that, seeing as users have past the 1 billion mark, and it's all in the hands of a company who sells your data for profit.
> It's not the tech that is bad, it is the fact it is in the hands of for-profit companies, who are willing to sell your data just to make their bottom line.
So are they not allowed to use the data? They are a business and will use everything they can to make their system as good as possible. Whether users care or not is up to them but it's clear that people find plenty of value in Facebook to continue using it.
If you're just stating that the possible surveillance capabilities are increasing then I think this is pretty well known...?
> So are they not allowed to use the data?
Being able to pull semantic knowledge from data changes the game. I expect in the next 5-10 years laws are going to have to come out to stop companies going too far with this. Sadly, laws are only really introduced after something goes horribly wrong, so we've got that to look forward to too.
> If you're just stating that the possible surveillance capabilities are increasing then I think this is pretty well known...?
Privacy advocates are a minority. My friends are almost all non technical and don't even think about it. That's because this is new ground, and there haven't been severe repercussions yet. But they are coming.
I'm not ruling out that they might use it for targeting eventually, but if this was solely done as a cover it would be the equivalent of a terrorist entering an airport shouting out "My suitcase is just really heavy, I do not have a bomb in it at all!"
People would find out within a few days and out would come the pitchforks. Here facebook preempts the conversation and sets the tone, now worriers will be met with "Do you not want blind people to enjoy the web? Why do you hate them so?"
There are times they could give you that button but already knowing you are human they use you to train their AI.
I wouldn't be surprised if they also score you against how well you train AI, and give better trainers more captchas.
Google has so much data, any time you think "I wish we could do X, but we'd never get enough data", it's an opportunity that Google can move on but the rest of the world can't. That's another argument for breaking up monopolies.
Wouldn't this mean that it would be an opportunity that /nobody/ could move on?
The problem I see here is that now they are giving this data to spammers.
I don't use their services but apparently I'm still tagged in Facebook photos. I never accepted ToS permitting that.
Put another way, with absolutely no snark or malice intended: Information about you is not necessarily owned by you.
Take, for instance, license plate scanning. Of course, anyone can see your license plate when you are driving around town, it is public information.
So clearly it is fine if machines scan license plates and keep records of every license plate that crosses a bridge. A team of people could sit on the bridge and do the same thing, so no big deal.
So clearly it is fine if we install the license plate scanners on every street corner in the city, and scan every plate they see.
And it should be totally fine to put that data on a publicly searchable real-time database, so that I can input any license plate number and see immediately where that car is.
These different scenarios might seem like merely matters of degree on the same bit of info, but what it means for us changes immensely as it becomes easier and faster to do it.
Not to say I don't think there should be license plates; obviously there's a public safety argument for having them, and that argument has to be counterbalanced against the privacy concerns (and perhaps rebalanced as technologies advance and change the relative difficulty of some kinds of privacy-invading tasks). But focusing on the scanner is pointless. More broadly: acknowledging that people trivially have access to tons of data that, when taken together, damages privacy, and trying to fix the privacy problem by telling them they're not allowed to look at it, is never going to end well. If the problem is solvable at all, it's by controlling access to the data in the first place.
What I was trying to say is that being able to do something fast and at scale is not just a progressive improvement from being able to do them slowly and by hand; it is actually fundamentally changing what the thing means. What we decide to do with this fact is another question.
In your above scenario, I'd start calling for the metaphorical heads of politicians somewhere between "installing scanners on every corner" and "public searchable real time database".
The example we're speaking of here, Facebook has collected the information that the GP poster exists and looks like this. Already publicly and digitally accessible information. Nobody's privacy has been violated by saying that person X exists and looks like Y, nor could that information be used against them. I certainly hope you understand that a live database of car positions is fundamentally different...
On top of that, the genie was unbottled ages ago. You may be able to make some kind of privacy argument re: the plate scanners now, but as the tech advances more and becomes cheaper and it becomes trivial for anyone to do the same thing (I recall an article about someone who set up a scanner in his front yard with some open source toolkits and relatively cheap gear), the arguments that "anyone can do it, except these people, because reasons" start sounding more and more arbitrary and out of touch.
We haven't seen the full force of what this means yet as it's still early days, but with for-profit companies behind the wheel like Facebook who make money from selling your information, this is an unprecedented level of big brother-esque invasion.
For one thing, you are leaving out the part that says "and doing Z". There have been lots of cases of people having tagged photos be used against them. While you might be fine with the cases of someone being tagged at a baseball game when they said they were sick from work, but what about when someone is tagged at a gay pride parade and they live in a homophobic area? Or being tagged in a picture of a protest when they live in an oppressive country?
I just tried to tag the fictional "Jimmy Andersefsdfsdfsd" ad it doesn't work....
The only new thing here is that Facebook released it to the public.
Just quickly looking at the top few posts in my feed, I see someone celebrating their two year friendship with someone I don't know, one person sharing a link to a new airplane, four people sharing videos, one person liking a sponsored video, and finally one person updating their profile picture.
I wish I could see actual updates from people instead of being kept abreast as to what piece of third party content they've liked at some point in time, what third party content they're sharing, etc.
Your feed is what you make it.
This nicely captures the problem. Facebook (and probably most social media systems - I'm looking at you LinkedIn) are primarily interested in keeping you up to date with Facebook via the medium that is your human relationships. So Facebook wants you to know what your friends did on Facebook today so that you might do that same Facebook thing. This increases Facebook interactions which in turn become more information to propagate to others on Facebook. In the limit there is no need for Facebook because all anyone is ever doing is Facebook.
Sadly, it also brings another complete set of cases to the oh-so-anoying "but Facebook/Google/Twitter/Amazon does it!" clichés that we'll now have to deal with...
The classes that it can detect are from ImageNet, so that might be limiting.
Maybe someone should create a website that lets volunteers label images for this purpose.
* http://i.imgur.com/Wex6pSR.png
Which I say is pretty much what you'd need to train a net to detect people smiling (amongst other things). Of course, there are some refinements you can make to improve accuracy and presentation of the results. My point was: you should begin with datasets that are readily available, and then improve on need (and if resources are available to justify the investment).
https://research.facebook.com/blog/how-blind-people-interact...
Including a link to the publication that was written on the technology here.
https://research.facebook.com/publications/how-blind-people-...
I think its exciting and an honest attempt to make peoples lives better.
If I were blind, I really wouldn't care that this is an image of "two people, smiling". Facebook has facial recognition, tagging, and locations. It would be much more valuable to me to say "Peter and Laura smiling at Channel Islands State Park."
Demo: http://googleresearch.blogspot.ro/2014/11/a-picture-is-worth...
But it's way too expensive.
All I wanted was keywords for alt-text, dimensions for placeholder, and the dominant colour for placeholder background.
https://cloud.google.com/vision/
The price for that would be $7.50 per 1,000 images for the first million images.
I have some 60,000 images on the site I run and don't happen to have $450 in loose change laying around (the whole site costs less than that to run each month).
I guess I don't care about alt tags that much.
Obvious next step: build this into the OS/browser/screen reader.
It's impressive that people don't really connect the dots and see that as a huge threat to their freedom.