This adorable little infographic on "the journey of a voice request" conveniently leaves out that it gets used for advertisement[2]. They have also made public statements that outright state that voice data doesn't get used for ad-targeting[3]
[1] https://www.amazon.com/Alexa-Privacy-Hub/b?ie=UTF8&node=1914...
[2] https://www.amazon.com/b/?node=23608618011
[3] In a statement, Amazon said the company took “privacy seriously” and did “not use customers’ voice recordings for targeted advertising.” https://www.nytimes.com/2018/03/31/business/media/amazon-goo...
“not use customers’ voice recordings for targeted advertising.”
I guess it depends on how one reads that quote. A trusting sort could read that to mean, we don't use anything we learn from voice recordings for targeted advertising.
A skeptic might read that quote and determine:
Well we generate metadata from the recording, and we then use the metedata for targeted advertising, but we don't use the actual recording for advertising.
Which makes sense, if I was to implement something like this, I wouldn't use the actual recording, I'd process the recording(which I have to do anyway to answer the request) and if I happen to save some useful for advertising data along the way, well, more $$'s for me!
Which one is true? I guess it mostly depends on how hungry Amazon is to make a buck and what they think they can get away with. As a privacy snob, I'd prefer the trusting version to be true.
The paper does talk about “voice data” which I think is a bit misleading. “Voice data” to me would imply an analysis of the sounds your voice makes directly for ad targeting purposes but it’s clear enough from the context what they actually mean.
Amazon does plenty of rubish things already, no need for us to make up extra things!
1. "Smart speaker vendors or third-parties may infer users’ sensitive physical (e.g., age, health) and psychological (e.g., mood, confidence) traits from their voice."
2. "The set of questions and commands issued to a smart speaker can reveal sensitive information about users’ states of mind, interests, and concerns."
They also mention that "smart speaker platforms host malicious third-party apps", and "record users’ private conversations without their knowledge", but that's mentioned as examples of prior research and thus seems to serve more as background than something this paper is trying to prove.
Point 2 is the one you're focusing on, and yeah, that's not surprising. You'd expect Amazon to build a profile on you based on the stuff you ask Echo to do (though the ethics of this certainly warrants discussion).
Point 1 would be the surprising thing, that smart speakers infer information about people from their voice, rather than from the commands themselves.
Their methodology seems to be to create multiple personas and compare the sorts of ads they get. In order to prove that information is inferred from traits of the voice rather than the words in the commands, they would need two personas which are identical in which commands they send but with different voices (female vs male voice, healthy vs smoker voice, something like that). From skimming section 3, it doesn't seem like they did that, so I'm forced to agree that the thing they prove in this paper (if their statistical methods are valid) is that Amazon builds an advertisement profile based on your interests as expressed in terms of which commands you're sending the device.