Confessions of a location data exec
digiday.com
digiday.com
Contrariwise, if location data from later subscribers were used to provide solid information to earlier subscribers that would be an effective and useful service.
If there is collusion to both sell fake data and detect fake data that is indeed fraud, but it's a different kind of fraud.
Thus, like in a (complete-knowledge) Ponzi scheme, the early subscribers have an incentive to get later subscribers to sign up for the service, so that it can be more useful for them.
But, unlike in a Ponzi scheme, the later subscribers don't have to wait for subscribers even later than them, before they can benefit. The later you subscribe, the more immediate value the system has to you.
"There are ad tech companies that promise ad buyers they will find fake data. Those vendors will usually eradicate that data at a cost-per-thousand, as a data fee. Similar to what happened with viewability and ad fraud. I’m sure there are companies looking to solve this problem, but you’ve got to wonder: Why would I want to pay more to validate the data that I already paid for? If this happened in the financial industry, then people would’ve been locked up for it — it’s like Bernie Madoff’s Ponzi scheme. Companies that detect fraud would not need a reason to exist if the market didn’t pay for the fraud to take place."
The bit they explictly compare to Ponzi/Madoff is having to pay an extra fee to detect fakes in data you bought because the "market" is paying for the fraud to take place.
This is almost, but not quite, entirely unlike a Ponzi scheme.
Also, this is terrible journalism. It's almost impossible to tell which statements constitute the "confession" and which constitute editorial comment.
It's kind of like how dating services buy and sell portfolios of user accounts so that they can pre-seed their services with apparent potential matches, thus attracting real users who will (hopefully) become the real userbase.
I hold a party and charge an entrance fee.
The early arrivals are in an empty room. If nobody comes after them, they won't have much fun. If you arrive after a critical mass of people is already present, though, the value of the party is instant.
I guess the difference here is that people know how parties work in this respect; this business may not have been so clear.
It's about as much of a Ponzi scheme as, say, a day at the track.
IMO, equating it to betting at the track isn't very accurate. Most people that bet at the track know they are taking on risk of no return, and what they gain for that is increased payout. Where's the knowledge here that you might not get what you paid for? What's the benefit of taking on this risk? Is it still betting if you buy something at a store and the shopkeeper turns around and puts the money into a slot machine before you've been given what you paid for and before know what's going on?
The author makes no mention of any factual similarities.
The author is suggesting there are legal similarities: both are criminal fraud. He suggests that those committing fraud on investors are "locked up" and insinuates those committing fraud on advertisers are not.
The author were "[i]f this happened in the financial industry..."
If what happened? Sales of ad data? No. Fraud.
The author is not suggesting that deceiving purchasers of location data is factually similar to deceiving purchasers of investments.
The author is suggesting they are both fraud, i.e. legally similar. Both are intentional deception by a seller on which a purchaser relies. Fraud.
The difference the author sees is that those who commit fraud on investors are incarcerated and those who commit fraud on advertisers are not.
Maybe the author was just trying to cite a well-known, real world example of fraud (intentional deception by seller upon which buyer relies) where the perpetrator was incarcerated. In that case, the example he picked makes sense.
As for the analogy, it's possible this exec is just that bad at language, sure. "My darling, you are like a fine Stradivarius violin... it is possible to make you useless by lighting you on fire."
And yet, people build million dollar businesses on this stuff. Blows my mind.
Edit: And yes, sometimes people play podcasts via players on the web. This makes up a tiny fraction of overall listens. It's not statistically significant.
As it damn well should be.
That said, it's only a matter of time before carriers start tracking and selling this data.
That's one way to get that data.
We do account-based marketing (advertise to specific companies you want to sell to) but we're upfront on the minimum size that companies need to be to have a internet footprint. There are other methods to find people working at a company but one thing we see all the time is companies claiming to let you reach the CFO or an exact job title. That is 100% fake and it would be far more effective to just send a letter to that person instead, but people just love to believe it works.
There is no need to have the phone in hand with the application open in order to collect location data - we get billions of signals from first party data from background usage on both iOS and Android.
These SDK's require legitimate usage of GPS in order to deliver the functions that we provide - our permission requests are overt and informational - we let users know exactly what data they will be sending and how we use it (one of our products actually rewards users for doing this!).
True, Apple and Android are both cracking down on un-authorised usage and collection of this data, it just means you have to follow some more rules in order to collect this data. Either way it is good for user transparency.
It sounds like (for the most part) he is talking about bid stream data, which everyone in the industry knows is the sewerage of the location data industry. Sure you can get a few valuable insights from it - but should it be classed as location data? No way. Even with significant cleansing it is still mostly garbage. If that is what their entire business model is based upon then no wonder it feels like a ponzi scheme to him.
For the rest of the location world who are using quality data - we see things differently!
https://blog.safegraph.com/less-than-10-of-bid-stream-locati...
> we let users know exactly what data they will be sending and how we use it
Many apps will just say “see privacy policy” or similar when requesting access. See [1] for examples. That is an approach I would consider dishonest as most users will not try to find and read through it. Does your method use a more explicit consent flow? I think that is a big problem in this space, so you could really differentiate yourselves if informing users in a truly explicit manner unlike others.
[1] https://guardianapp.com/ios-app-location-report-sep2018.html
One of the things we wanted to achieve with our Reward product is that users not only know that we are collecting data but can actively monitor what data we are receiving from their device. Most importantly they can see a independently verifiable 1-1 relationship of rewards for each days worth of data sent.
It isn't much in the grand scheme of things but your data has value, you could argue that this data is payment for such "Free" services such as Facebook and Google - but transparency of this is important.
I really like what Brave and BAT is doing for turning the advertising paradigm around - part of me thinks that it is against human nature and it won't work but we literally have an entire generation of people who do things for "Likes" or "Upvotes" so maybe I am just wrong (I hope I am!)
Just out of technical curiosity, does your dataset contain PII or just UUIDs and location data?
If no PII, how to buyers of the data join it to their own datasets?
We don't collect any PII data (in fact we actively avoid it!), we don't need to know 'who' a device is, only the location history of that device!
There are companies out there who will attribute a UUID to a physical email / telephone number, I think it is a bit of a grey area and with the current privacy landscape something I think that will soon end.
Sounds more like textbook fraud than a Ponzi scheme (though, honestly I did not read the whole article).
It's true that there are many junk vendors. Heavy politics and misaligned incentives for ad agencies usually means that the bad companies do better than the trusted vendors because fake data can obviously be shaped to look better than the real results.
That being said, location data does work. There are many ways to collect it from visual scanners in doorways, to open wifi networks that ping phones, to ISPs enriching data feeds. There are also 1000s of analytics SDKs embedded in apps that send pings constantly, so having a specific app open is not a necessity and never actually used by any serious network. Pretending that's the only way it works is just misleading.
Those intermediaries are the main worries, because brands or marketers may not trust every random Android app, but they will trust Yelp or Foursquare or something, and frankly given additional scale from data bartering and SDK pings that send back location from many other apps, places like Foursquare or Yelp really can provide targeting services that come close to what adtech has always promised.
To be clear, I’d consider this a bad thing, because places like Foursquare or Yelp, whatever they may say to put a PR spin on it, are quite literally preying on people who unwittingly share their location data and don’t really understand the terms, especially not when it’s some tertiary SDK traffic logs agreement causing some music app or hotel app or weather app to send them data tied to your IP address or device ID.
If those businesses can’t monetize their basic value proposition, like a app to search reviews of restaurants, it’s a signal to delete & shutdown the app... but unfortunately it’s become the reverse: a signal to abandon investment into the actual user and diversify all kinds of deceitful behind the scenes ways of getting user data and making users into the product.
Reminds me of this story about why whales don't die of cancer: because they're so large, that by the time their cancer grows to be dangerous, it gets its own cancer and dies.
The BBC article you posted looks like it has results that are 1) derived from actual animals, and 2) more directly useful, but I'm just a little disappointed that tumor cannibalism isn't their explanation, if I'm honest.
I don't think this hypothesis is widely accepted, but I remembered it because it's neat, and fits the game-theoretic model of cooperation/defection perfectly - which also makes it extremely useful for drawing analogies.
It is my belief that advertising is a cancer on the society; it's exploiting - and in the process, destroying - every vulnerable individual and social heuristic. It involves uncooperative behaviors like manipulating people and lying to them.
The analogy here is that since entities making up the advertising industry eschew cooperation and embrace exploiting others for short-term gains, they're not going to magically start playing fair and cooperating within the industry. Therefore, to the extent you expect advertisers (including adtech) to scam you, they'll scam each other just the same - as seen in this article.
This, fortunately, somewhat limits the effectiveness of that industry.
I came up with this analogy few years ago, when I read accusations that Optimizely designed their A/B testing suite's UI in a way that promotes drawing statistically unsound conclusions from A/B tests, misleading you to believe that the tested intervention worked - and thus making you think Optimizely is successfully helping you learn things. My own personal observations from working alongside one social marketing team also confirmed the soundness of this analogy.
EDIT: that Optimizely debacle I'm talking about:
https://blog.sumall.com/journal/optimizely-got-me-fired.html
To boot, Foursquare at least, and probably Yelp too, does data swapping and SDK agreements, like some thing recently announced with Accuweather (gross) and with Hilton Hotels.
I think people sincerely fail to imagine the real scope of this type of egregious trust violation and surveillance business model.
You can dress it up with whatever language you want about providing value to the user that makes them agreeable to the data collection terms, but I’m sure half of Foursquare users or Yelp users don’t actually know if they have the background location tracking disabled or not, or what other innocuous-seeming apps are silently feeding location pings to build a Foursquare data history about you to make you targetable for ads.
Frankly, I’d personally advise anyone to absolutely delete these apps or anything like them, and essentially vote with your wallet / vote by boycott and just refuse any apps that rely on this kind of business model.
[0]: https://www.adweek.com/digital/foursquare-will-fuel-accuweat...
facebook and google do, but they are not selling it. and the rest of ad-tech loves to boast about their "data" on every conference while all they have is /dev/random output, more or less.
And there were like 2 ads shown to you during this 50 seconds period, while you were in the dark and not looking at your phone but using it as a duh, flashlight. And then you stopped the app to save your battery.
Do you really think some advertiser benefited by using this data?
And what "independent" ad-tech is dealing with is millions of cases above. Consistency is important. FB and GOOG have consistent data, collected over days and years, properly matched, correctly aggregated. All others are dealing with flashlight crap.
Mobile apps are 1000x more invasive than anything on the web and these SDKs constantly send data in the background. There are also 100s of signals that get correlated to build single device profiles across time. FB/GOOG do have an advantage but that doesn't mean everyone else is useless and if you have ISP partnerships then you can actually get even more consistent data down to a device and household.
A lot of those arguing your parent comment's point in other threads here and elsewhere seem to magically forget background data. Out of sight, out of mind.
Use a simple app to track background connections on your phone (e.g. NetGuard or the like) and you'd be amazed at how much gets sent back to the mothership.
The Twitter PWA Chrome app sadly doesn't block Promoted Tweets, and it's the same thing there; low-quality ads from garbage companies I'd never consider buying from.
But the disconnect here is the at the ad VIEWER is not the customer, the purchaser of the ad is the customer. It might be whoever has a relevant ad for you won't bid enough for you to see it.
Whether the ad is well targeted or not, the host still makes a profit (both monetarily and in terms of their own tracking data)
That's solid economics. If a war breaks out, be the one that sells guns. I seem to recall a quote along these lines but it escapes me...
Good luck doing it with independent ad-tech and their "data".
I have a great deal of experience, and I've seen that you can generate final sale for some non-brand obscure SaaS platform (so no branding effect here at all) for like $30 per sale in Facebook, $90 in a good independent ad-tech platform, and $600 in a bad platform.
Facebook still totally kicks ass and absolutely can make money for you if your SaaS brings you $100 in a first year. (100-30=70 in incremental profit for you). Hell, even really good ad-tech platform (yes, they exist) can make you money, though less so (100-90=10 in a first year). Shitty platforms are money sink, though
Verizon bought AOl, HuffPost, Techcrunch and plenty of other media to build an advertising giant by pairing data with sites audience -- and took 4.6 BILLION WRITEDOWN on the whole venture [1] because failed to make it work.
[1] https://www.recode.net/2018/12/11/18136127/verizon-aol-yahoo...
EDIT: oh and AT&T will learn the same lesson with their recent 800M AppNexus aquisition. Just give it time.
https://www.engadget.com/2019/02/07/carriers-were-selling-yo...
> Carriers sell information about you to data aggregators, which normally require the user to consent before selling it on further. But some third parties opted to sell information, like people's whereabouts, on to bodies such as bail bond companies, bounty hunters and landlords. In its original report, Motherboard paid an bounty hunter $300 to get the location of a phone to within a few hundred meters.
It has been illegal for them to do so until relatively recently.
In fact, this is almost tautologically true. If you have:
-- good data
-- on large enough scale
you are not going to sell it. You going to take advertiser dollars and match their ads against right audience, while safeguarding this data as your main know-how. Go and buy "barbell training fans in FL, US" and "women interested in fashion in CA, US" from facebook. Sorry, not for sale.
If instead you decide to "sell the data" it means that you can't really make it work for advertisers. Or, in other words, said data is mostly useless.
I don't know why someone would posit this falsehood.
You are going to sell it to EVERYONE you can (legally and sometimes not). There is no ad-tech company that does not take this stance. You get information about Honda buyers, you're going to sell it to Ford, then take what you sold to Ford (exclusively) and sell a repackaged format of the data (summary) you sold to Ford and sell that to Subaru ALONG with the data Ford asked for in the future.
Data management platforms (DMPs) are an entire concept built around selling the data. Every adtech company (that has any tech) sells the data that is collected, outside of using it for internal targeting. The number of demand partners and amount of useful targeting data to leverage for them, is necessarily a smaller set than what you collect.
Edit: This is experience from multiple senior ad-tech positions.
lol
>You are going to sell it to EVERYONE.
Exactly. And you know what you end up with? I checked lingerie website yesterday to buy a present for my girlfriend, this data was sold, now I'm a women in some DMP. I've bought a bottle of Talisker, now i'm wealthy man in the same DMP (even if I can only afford Talisker one time per year!!!). I've also bought some drugs for my grandfather, and now i'm +70 audience and I'm getting "senior discounts" ads while not being even 35 years old.
This is the problem with data platforms. Take data from everyone, mix it, and you have /dev/random. It's totally fine to milk advertisers for their money (which original article is all about), and not fine for anything else.
Advertisers finally start to understand it. And this is why GOOG and FB are getting all the money and ad-tech is slowly dying.
EDIT: And since -- as you correctly noted -- everyone is incentivised to sell ALL data to EVERYONE -- this is how you end up with a pool full of piss instead of water. Because everyone sold everything to this pool.
In practice, the level and quality could vary significantly, and the difficulty to normalize the feeds could be much harder in practice. If you get data in aggregate from Foursquare and Yelp, potentially it can tell you enough to gauge foot traffic in restaurants, but to your point, it probably can't be used to build an advertising platform.
Insurance companies are a great example. The data providers for pharmacy and automotive are another.
Ad-tech is different by boasting that they do know a lot about you, while in fact they have no clue.
Isn't 80% fake data (i.e. bot clicks, auto play video counting, etc...) on par for the ad industry?
5 out of 10 "surveys" go like this: Which of the following places did you visit recently? Yes? How did you pay? (Credit card, debit card, cash, made no payment, …).
Retargeting/rebrokering it for ever smaller margins doesn't seem to make sense (except for those selling that snake oil).
Volume and campaign quality count more than just getting the exact words and the exact public (and remember you're paying more for a more targeted ad), targeted ads of course work for a targeted segment/audience, for most products that you would find in a high street/main street store, not really.