Its hard to compete with Google on a task like image classification when Google has immense computational resources, tons of data, and hoards of top researchers.
Its hard to compete with Google on a task like image classification when Google has immense computational resources, tons of data, and hoards of top researchers.
But that is not the case. Which suggest Google isn't that good at productivizing stuff.
When you open an email with images, Google will proxy the request for the tracking image and then cache it. If each user has a unique tracking image, you know when it was opened. Google is not caching the images before the open so you do know that it was opened.
What you possibly lose is repeat opens which might end up with the cached image.
That seems like a loss, but with this change, Google turned on images by default. So you get loads of Gmail users loading images by default rather than the old way where many more people would be loading your message with images off.
MailChimp has a little write-up about it: https://blog.mailchimp.com/how-gmails-image-caching-affects-...
tl;dr: Google's change didn't stop marketers from knowing you opened the email, but did potentially block cookie sending, IP address information, referrer information, etc.
Unfortunately, it's looking more and more like it's going to be a competition on training data and raw computational power, and it's hard to compete with Google's corpora from the web, gmail, captchas, maps, etc. -- not to mention Google's tremendous number crunching resources.
And possibly, though for privacy reasons in a much more filtered down form (see the word2vec Google News dataset). I find it unlikely for private data though.
Could be any kind of niche, from specific industries (utilities, media, transport...) to specific use cases (I don't have many in mind, one could be https://sightengine.com). Ideas welcome :)