If Google has the copyright on "1", they only need to get "0" as well and they'll have everything.
So, "someone" has them copyrighted.
Statistics don't lie, which is presumably why google employs so many of them, to calculate these efficiencies.
However statistics can be used to confuse.
If a file is 50% 1's, then a 1-digit file has a 50% chance of infringing. More than that, the chance of infringement grows pretty fast.
Also, if "1" is copyright, then a file with a "1" in it is infringing; there's no chance there. It's certainty.