Yev from Backblaze here -> at a conference today (JNUC2017 for those curious) but I'll be checking in here to answer questions should any arise!
Love to see some performance stat to see if read /write IOP, MB/second, etc change over time on per model base.
Other than complete failure, do you collect check the any other early warning failure data such as ECC change over the life time of the drive? --- Would be a fun big data exercise for an intern to work on.
Also what's the usage patterns for your HDDs? read/write/spin down/sleep time in term of life time of the HDD.
I assume mostly write, some read. But also very low percentage is active?
Yes, we collect temperature data. If you download the dataset, you'll find a daily sample of the SMART data (including temperature) for every drive.