Google Drive scans files for copyright infringement
twitter.com
twitter.com
What it's scanning for in this case is material it believes to be copyrighted, and restricts features (notably sharing) for content that matches.
Given that copyright law exists, and that Google doesn't like wasting engineering time on legal stuff unless refusing to do it would result in lawsuits, this scanning policy is a fairly low-impact solution that has probably been deemed legally necessary to avoid media company lawsuits. I don't like it, but the alternative is for Google (and Microsoft, and any other cloud storage that allows sharing) to mount an expensive legal effort to try to overturn decades of digital copyright precedent, which is likely to fail.
All "material" created by any human is copyrighted, and has been for decades. The question is who owns what rights, which Google can't know.
> decades of digital copyright precedent
What precedent?
There can even be the case where you have private conteacts that give you rights that Google cannot know.
I can inly thing that they will look for the low hanging fruit: a folder full of movies, music, and other content.
What a hosting service needs to provide is a way for the user to flag content + provide an "acceptable" response time.
Another approach, a more proactive approach, is for the content owners to share the digital "finger print" of the IP (movies, music, etc.) with the hosting service, so that the hosting service can scan uploaded files and compare them.
Ah, the old "deeming things" trick [0].
Almost everything is copyrighted. Like most of us here I've given original writing, code or music to people who've shared it on Google drive. That material is copyrighted no more or less than anything by Disney or Sony.
Google doesn't just "scan for copyright violations", it specifically acts out of fear or leverage to be an unpaid policeman for special interests, rich and powerful media companies.
I haven't said anything new here, but I do think it's important that we see arguments built on false standards. Google isn't championing the law or anything noble and we would do well to be very precise about choosing words to describe what is happening.
[0] deem: acting by fiat and art without necessary recourse to logic, law, evidence or consistency
Encrypt everything and you, as the company offering services, no longer need to worry about whats being shared because you cant see it.
If you just keep files private on your Drive, or share them with specific other accounts but not to anyone with the link, there's no copyright scanning.
All of this makes perfect sense because Google doesn't want people using their Drive accounts for mass sharing pirated content.
While if you're just backing up your CD's and DVD's for personal private use, everything works fine, zero problems.
head -c 20 /dev/urandom >> movie.mp4
won't affect playback, will affect Google finding your pirated films.They already have to implement such a thing for finding copyrighted material in Youtube videos, so they know how to deal with mixed signals.
Not endorsing this, btw, I much prefer dumb pipes.
... so it has no utility at all, since you definitely shouldn't be storing any unencrypted files on it.
As another data point, I have several manually-uploaded zip and 7z files of my own documents and content, both encrypted and not, and never had a problem.
The majority of my Drive usage by bytes is encrypted backup pack files uploaded via API by Arq Backup or rclone. That's all worked fine for many years.
GMail also complained about direct-attaching a locally-built, unsigned exe file to an address that I'd never corresponded with. To be fair this probably should have caused suspicion :)
I still have a handful of files which are books in PDF format in a RAR file, and simultaneously the book's cover as a jpg.
I d assume that they check some hashes of the file against a database to check for copyright infringement. If only specific actions are not permitted on the file e.g. sharing it widely, this could seem reasonable?
Curious to learn more, what could be other actions the service provider could take to avoid getting a lawsuit?
Note: I'm in a process to completely "de-googled, de-microsoft, ..." all my stuff (big self-hosted TrueNAS server with BlackBlaze backup).
- "so, google has scanned my recently filed scanned files and said it's a copyright infringement"
- "Bro, tell me your Gemini datasplit?"
> "Google"
> "Your file may violate Google Drive's Terms of Service"
> ""05 - You are always choosing.mp3" contains content that may violate Google Drive's Copyright Infringement policy. Some features related to this file may have been restricted. "
> "Restricted file 05 - You are always choosing.mp3" - ""
Glad I don't use google drive.
People laugh when I suggest iCloud but Apple isn’t pulling this shit and has mostly the same functionality a non-business user needs.
I'm not even talking about Gemini's training data.
Recall that it was Google that utterly nonchalantly scanned & stored some 40 million books, without seeking and obtaining permission from a single one of their authors. [ https://en.wikipedia.org/wiki/Authors_Guild,_Inc._v._Google,....]
Is the internet as we knew it drawing to a close?
Were the interactions on the internet as you knew it on completely centralized commercial platforms?