> It is modifying what the artist intended you to hear in a destructive way. It is destroying the original performances.
also applies to lossy encoding.
The whole point of audio watermarking is that it survives lossy compression. Inaudibility is merely a secondary concern.
This is an adversarial relationship, and one that the audio watermarking is doomed to lose. By and large, lossy codecs succeed in transparency, and therefore as an inevitable consequence, watermarking fails at it. After decades of development, lossy codecs are just too good - there's nowhere left to hide information.
(It's also worth noting that the consumer benefits from the quality tradeoff that compression makes, in the form of decreased storage and bandwidth. They don't benefit from the audio watermarking at all.)
I'm unable to load the article and haven't heard one in real life, but in theory it could be done in a way that is imperceptible to human ears but detectable by a program. E.g. imagine masking specific frequencies within a noisy section of the video.
There should be less destructive ways for studios to do this, but unless it's in the video or audio signal it doesn't stand a chance of working reliably.
To be clear, I'm also against any kind of watermark, but I'd rather it be done with an imperceptible (to me) audio signal than with a constant graphical watermark as is usually the case. Actually, graphical watermarks could be smarter and less obtrusive as well.
> The difference is that the graphical watermark on e.g. a photograph usually isn't there when you've paid for it.
The same could be the case for these audio ones. It's user hostile if this is done on already paid content (but again, I can't read the original article to confirm).
[1] https://creativepro.com/major-movie-studios-specify-digital-...
[2] https://partnerhelp.netflixstudios.com/hc/en-us/articles/440...
As for more subtle changes, what constitutes a "plausible sound" in any given context is an AI-complete problem. Sure on a particular song, you might get away with e.g. tinkering with the reverb a bit, or changing the tempo by 0.5% - but how can you possibly do that in the general case?
Remember - all such changes have to be audibly salient to humans, because that's the only thing guaranteed to survive.
You’re saying it’s ok to decimate it - as long it’s replaced with some “plausible” proxy? Plausible decided by who, the artist or business requirements?
We tolerate Netflix releasing films in less than 4k. They do that because of a practical business benefit (cost). Similarly, Universal is making a compromise in the artist's vision to meet a business requirement. I think if the compromise is sufficiently small and the business benefit is sufficiently large, it's ok to do this.
When you're passive listening, you're only perceiving the "outlines" of the music, but listening actively reveals a lot of details, and this kind of wobbling becomes both more noticeable and disturbing.
When listening the music actively, anyone can distinguish between natural and unnatural sounds, and this will bother many people if they listen more actively.