It's rare for all of the human drivers on a road to make the same mistake at the same time under the same conditions, but that's a likely failure condition for automated driving software if the software and sensors are standardized. This could lead to much worse accidents then you'd get with human drivers, because all of the cars will do the same thing.
An all-human example: driving too fast in foggy/whiteout conditions and not having time to stop when the road is blocked. When this happens today, you get 100 car pileups because everyone is doing the same thing: driving along until they suddenly come upon the pileup, and smashing into it because they don't have time to stop.
What if a software error causes this same kind of behavior in more common conditions? Google's testing wouldn't expose this kind of failure mode, because (afaik) they haven't done extensive tests driving fleets of automated cars together. All of their cars have been surrounded by human drivers, and for all we know the varied reactions of the human drivers have prevented them from hitting the Google car when it does something odd.