Found it here, which also goes into some testing the Guild Wars guys did on their population of gamer PCs: http://www.codeofhonor.com/blog/whose-bug-is-this-anyway (scroll down to "Your computer is broken", around 1% of the systems they tested failed a CPU-to-RAM consistency stress test)
Both of them indicate intermittently defective components in running systems are way more common than anybody assumes.
But when they say "Memory Error", even though it's something detected/corrected by ECC I'm not sure we can say 'the memory is defective'
It may be a combination of the conditions of power/load/data/time since last refresh and variance between modules.
Since Google appears not to show all the data they have, we probably are not going to get that from them though :/
I know Intel has been working for some time on the idea of high temperature data centres, this will impact the MTBF of all components but you can always calculate the cost of the losses vs the cost of the cooling: http://www.datacenterdynamics.com/focus/archive/2012/08/inte...
Table 3 suggests that there are data sets that include all components (CPU, memory, power supplies, etc.).
Any large company that makes things employs a bunch of reliability engineers, who are usually EE's or ME's who make Weibull plots and bathtub curves all day (to set the warranty duration, mostly). These guys have all the data you could ever want on this topic, but they're not sharing. Especially at Intel.
As a consumer it's hard enough to keep up with what's reliable in hard drives. Keeping the manufacturers honest with good stats for the most common parts would be great.
https://static.googleusercontent.com/external_content/untrus...
The backblaze guys also have a big data set:
http://blog.backblaze.com/2013/11/12/how-long-do-disk-drives...