Backblaze Hard Drive Stats Q3 2019
backblaze.com
backblaze.com
But I'm stuck on one thing. Does Backblaze offer a solution for Linux backup? I've got an NFS server running that I use for home storage that I want to back up - but looks like Backblaze is only offering a Windows or Mac client.
Maybe the business version would work, since it claims to support NAS backup. But then the pricing seems lower than the personal edition (60$/computer/year = 5$/month < 6$/month) - unless that's implying that every computer that accesses the NAS is part of the fee?
So I guess: is there a reasonable Linux offering for home users from Backblaze? If not, what service do folks suggest?
Gotta decide between taking the hit on adding another "except for" to your unlimited marketing claim...versus taking the hit on people pushing their multi-terabyte torrent collection into your service.
It's a lesser of two evils & the two providers picked different ones.
> So then why doesn't Backblaze have a Linux client
There are several choices for Linux including: Duplicacy, Duplicati, GoodSync, HashBackup. All of those will backup your Linux data to the Backblaze datacenter. Is there some feature those are missing? Are those not good solutions for Linux?
> again and just have like a 1TB limit
I believe you can set a 1 TByte limit in all of the above, and then the cost will be a very reliable $5/month or less.
Marketing. You can quantify/understand 1TB limit. Grandma can't.
That's what I meant by you want as few "except fors" on your unlimited claim as possible.
The second you do a tiered by GB model you destroyed your company principle / core selling point - "don't worry".
My rough numbers are 450GB stored monthly, 4GB downloaded monthly, 90k stored files, 385,000 individual transactions, which ends up costing about $2.25 in storage fees, $0.25 for transactions, and $0.10 for download bandwidth.
I've been very happily using Restic and B2 for a long time. It's cheaper than the unlimited service given I have ~600GB stored. Plus, it's a "real" backup: on-going series of snapshots, older files are not replaced when a new back up is made, I can restore from as many points in time that I want to store, and I don't have to worry about Backblaze deleting anything due to inactivity.
One underrated feature of Restic is tagging, which lets you identify a collection of snapshots. I can point multiple separate backups and devices to the same repository tracked with tags; thus, I de-dupe across all of them.
If I'm reading this correctly, Restic + B2 sounds like an absolutely godsend!
Some more (informal) analysis on restic's crypto was done here:
I made a script to backup nested zfs volumes to B2 using duplicacy if anyone is interested : https://gist.github.com/icefo/07aab2789e5cfa71045343953aaf88... It makes a snapshot, backup that and handle unexpected network, power loss or backup that span longer than the cron interval gracefully.
If you're looking for a backup service to handle point in time recovery, differential deduplication, or other features of a backup service (vs. a second copy of your data archive) those also exist though I don't have a clear recommendation for home users.
On Backblaze's business pricing my understanding is they require minimum of 5 users so that's where you'll see the difference solved.
I have used restic and rclone with very crappy platforms (Google Drive and Microsoft OneDrive) for testing purposes and that worked fine enough so I can imagine it works even better on something like B2.
But it just occurred to me, that for home use, if you're a special kind of masochist, you might expose your DAS to a vm running ReactOS and use the windows client?
I don't recommend it, and have no idea if it would work.. But would love to see a write-up if someone wanted to try it...
I was on Crashplan for a long time, moved to RClone to Google and AWS and likely Duplicacy to those systems.
O365 comes with 5tb of space that can be addressed by a diff backup tool like duplicati.
Only gotcha I can see is that it's 5x1tb
Well I run a couple dozen Synology NAS in professionnal setup, as well as two in personnal setup (mine and my parents'), and ever since that post I made the experiment of having almost 50% of all drives be Toshibas, and I have to say they do seem to be much more reliable (on the scale of "why do every other drive from Seagate and WD keep dying first, and often their replacement dies first too").
It is still a scale of use where it's mostly anecdotical rather than verifiable data, so don't take this fun comment for more than that. But I suspect a lot of people reading these posts are not interested for some large scale setup or anything like that but rather to know which drives to put in their home computer or NAS, and honestly I can highly recommend the Toshiba for that. They do tend to be a bit more expensive (around 10% more ? I buy them from ldlc.com and grosbill.com , french IT stores, no bulk buying or anything like that)
Of course no matter the brand never expect no failure and a Toshiba drive may just as much die in the first ten minutes so always plan for it.
So the reliability may be "worth more" to someone who pays a lot for remote-hands in a server farm somewhere, but not in this case.
I suspect when buying in bulk like BB does the difference becomes even larger ? Also they seem to build so as not to care of the reliability of any drive (as long as it stays within acceptable range), so a gain that may seem massive for us low scale user may not be of much impact to them compared to the cost benefit.
I appreciate your diplomatic language, but the time to failure does matter, and consumers don’t have the same cost structure that incentivizes replacing working drives the way you do.
FWIW I share GP’s experience with Seagate. I had quite a few of them, ranging in size from 500 gigs to 2TB. Every last one of them died relatively quickly, while most of my Hitachis, Toshibas, and WDs from that era still work.
Seagate earned a permanent boycott from this customer.
That being said, I occasionally use these data to break ties or see variations between generations of drives.
Good riddance.
Turns out there is a publicly available KB article that mentions known bugs with transfer rate negotiation on some of their SATA3 drives. Of course WD won't say which drives. There's even a utility referenced in the article that you can use to disable SATA3 support. But WD won't make it publicly utility.
Meanwhile their support is stuck in a "have you tried power cycling the computer" loop.
If memory serves Western Digital / HGST / SanDisk were among the first to jack up prices after the Thailand disasters. Fuck em.
Meanwhile is the difference in failure rate between Seagate and Western Digital even statistically significant? For most comparisons you're looking at a fraction of a percent.
Thankfully it was just stressful - we had backups, but drives in our array started failing one by one, with about a weeks interval; unfortunately it took something like 4 days to rebuild the array each time, so we wasted a lot of time shifting writes elsewhere so the last backup we had + writes going elsewhere combined remained recent enough in case a second drive would fail. We got off easy, but the time we spent probably cost us more than having set up a more redundant system in the first time would have.
Also would be nice to have a similar warning to "Your data has not been backed up in [x] days".
I've had friends / relatives experience drive failure a few time in the past, and the look of horror they have when there is very little I can do to help them recover their photos etc. is something that I hate seeing.
And having the OS give a simple warning (that can be dismissed, with a "don't show me this any more" checkbox) would not over complicate things, and may end up saving some people's data.
Really no different than the current warning that Windows has, when you don't have antivirus installed. Or the fasten seatbelt signal that comes on the dashboard of your car when you start it. Or the flashing red light that gets added to some stop sign controlled intersections.
Also it would be nice if there was a standard way that backup software could inform the OS of backup status (this way it would serve as a secondary check in case the backup software's internal reporting fails to notify the user of bad backups). Just a little nice-to-have.
(I'm not advocating for this to be a legal requirement, just a nice feature if any OS vendor wants to add it).
Another thing I hate about the industry today: being hostile to the user and trying to force them to use their computer how you want them to use it.
> I've had friends / relatives experience drive failure a few time in the past, and the look of horror they have when there is very little I can do to help them recover their photos etc. is something that I hate seeing.
Then teach them about proper backups. It doesn't take a rocket scientist to understand "don't keep all your eggs in one basket". Or if you're going to implement some stupid forced user-hostile scheme at least use something that actually qualifies as a backup.
> And having the OS give a simple warning (that can be dismissed, with a "don't show me this any more" checkbox) would not over complicate things, and may end up saving some people's data.
Here's what will happen: the user will dismiss the dialog without even reading it. We have decades of experience showing us this. Users have learned that "warning" dialogs are meaningless precisely because of crap like this. Oh yeah, and they're super annoying.
> Really no different than the current warning that Windows has, when you don't have antivirus installed.
Exactly my point.
I don't really use Windows but you can do something similar--though in my experience it's not as simple.
I don't really care if I avoid any downtime. So long as I have belt and suspenders backups, I'm pretty comfortable.
If we're talking about "protecting" consumers, I think it might be wiser for all OS vendors to provide a free tier of version-controlled cloud storage.
I think Windows does remind to create backup (but not very loudly and it accepts local backups to another drive which is poor solution).
Unfortunately, almost everything is cheapest possible. For example, I would prefer ECC RAM as standard. It's very cheap for production (one additional chip per RAM stick), but Intel wants to force people wanting reliability to pay for Xeon CPU and most people don't care.
I use a Synology drive (using Synology's RAID-like format) for my Time Machine backups, and a Drobo for my less critical stuff.
They have WD drives. I have not liked the Seagates. However, I like the HGST stats.
One of the nice things about both Drobo and Synology, is that I can change the drive type and capacity "on the fly."
I have 10 year old WD black and green drives still kicking. I still have a gen 1 or 2 intel ssd drive that’s still kicking.
Dragster should be seagate. Weird autocorrect.
If you rip them apart they are the same as their laptop drives with seagate.
All drives usually ship with a certain number of bad sectors. An excessively high number of bad sectors can mean there are quality issues with that specific drive. They may be 'binning' those more poorly testing drives for the external USB drives.
I think Seagate is fine if you are willing to overbuild and deal with the refurb process within the warranty period. I'm not, and will not personally buy a Seagate product again unless their reputation improves (similar to what happened with IBM/HGST post-deathstar). It's not worth the time or aggravation, to me.
Seems like WD has similar rates.
I suspect that most of the time after a hard drive failure the drive gets replaced with a higher tier drive (typically from another manufacturer like WD as the person feels burnt by Seagate for having a hdd failure) which may prove more reliable.
That's my wild speculation anyways.
By any metric a thread on an all in all "niche" technical board with almost 5,000 posts and nearly 4,700,000 views should mean that a lot of people experienced the issue:
https://msfn.org/board/topic/128807-the-solution-for-seagate...
- https://en.wikipedia.org/wiki/ST3000DM001
- https://www.backblaze.com/blog/3tb-hard-drive-failure/
Later Seagate drives don't seem to be worse than the competition, but memories of that infamous drive model still linger.
But another fun story: I had a PSU blow up a couple years ago in a machine with three WDs and three HGST. All WDs were dead after that, the others worked flawlessly. Probably not a large enough sample size for any definite conclusions but at least it put a failure mode on my radar that wasn't there before.
My oldest drive is a 500 GB Western Digital from 2008 that's still operational today. I imagine its end is near, but I've thought about that for a couple of years now.
It looks like after stalling in 2017-2018, $/GB has dropped again - https://jcmit.net/diskprice.htm - but JCM doesn't have the large sample sizes Backblaze does.
Wouldn't survival analysis on interval-censored data handle this problem automatically? All of your observations of failure presumably are actually interval data, where all you know is that the drive failed sometime in between the last good check and the first bad check. Then it doesn't matter if some time periods have large intervals and others have small intervals, that just affects the precision of estimates.
https://github.com/gilbertchen/cloud-storage-comparison/blob...
Can anyone from backblaze say anything about their performance compared to other vendors?
The pricing is certainly ahead of others, so I would use if the performance is comparable to some of the leading group tested there.
Consider this: the least reliable drive in this dataset has a 2.7% annualized failure rate at an average age of almost 4 years.
That's low enough that many SOHO users will never see a failure, yet high enough that a rack worth will have seen more than one failure recently. Thus, your question could be answered with complete honesty in both ways: with anecdotes by happy users and with anecdotes by unhappy users.
Therefore, neither positive nor negative answers are useful to you.
This phenomenon underlies the fundamental weakness of self-reported online reviews. You cannot actually get a useful measurement for how reliable a product is solely by self-reported sparse feedback.
I started my NAS in 2012 with two 2 TB Red drives. Later I added two 6 TB Red drives. Some time, around 2016 I think, one of the 2 TB Reds failed and I replaced it with another 6 TB. Then the other 2 TB Red failed, like a year later and I put in another 6 TB so they all matched. I did not get replacements even though one failed during its warranty period (I am pretty sure the 2 TB Reds still had a 5-yr warranty at that time), because I wanted to replace it with a 6 TB anyway.
Currently all four 6 TB drives are still running, plus a couple of 4 TB Toshibas I grabbed at some point.
So, I don't think that drive failure after 4 or 5 years is so bad. I got my money's worth anyway, and that's why a NAS has redundant drives.
I've had a couple of WD reds fail. You pop in another and rebuild the volume. It's normal as long as it's not too often.
I'd rather buy WD than Seagate, but that's just me. I don't really have a choice. Maybe the only people who have a choice are either in the enterprise or buying one drive every few years for a PC build or something.
The rule is to avoid all of your drive being from the same factory date/batch, even if they're all from the same brand order from different stores or whatever to help make sure in case of a defect they're not all affected.
Also for home NAS ensure you have redundancy (a proper raid level, avoid JBOD and Raid-0), and have backups. Raid is not a backup.
ZFS is great in this regard: redundancy, data scrubbing to ensure data integrity and built-in snapshotting and data replication features (zfs send-receive). Can't believe how I got by without ZFS in my earlier years of data hoarding.
I have data that is important to me (family photos, etc.) on my hard drive. I have a backups of that data running on a Raspberry Pi with an attached drive. If my house gets broken into, or burns down, or hit by a tornado, basically something Really Bad, Backblaze has a copy of my data offsite.
It is tempting to use Backblaze as my only backup, but like they describe on their site, their primary value is as a backup of your backups. Normally you should never have to use them, and if you do, it will be slow. Now if you are in a hurry they offer a service to ship you your data on a thumb drive or hard drive, but that gives you an idea of their primary use.
https://www.backblaze.com/blog/why-backblaze-bought-a-porn-s...