Hard Drive Stats for Q2 2016
backblaze.com
backblaze.com
When you're using this as the sole drive in a desktop machine or something like that, paying the extra for a HGST is pretty straightforward, but I see too many people passing over Seagate or WD when they intend to put them in a RAID, for "reliability"'s sake, without (it seems to me) much thought into whether it's really worth it.
Sorting by $/GB for Hitachi, Seagate and WD on drives >= 4TB (http://pcpartpicker.com/products/internal-hard-drive/#m=19,3...), you get:
- 4TB WD: $129
- 4TB Seagate: $114
- 4TB Hitachi (HGST): $158
So in order to get the reliability of a HGST, you're paying a 38% premium, is the extra reliability helpful in a RAID?
If you look at say a RAID6 array, aiming for usable storage of say 16TB, HGST's longest running drive (at 0.11% failure rate) would have a roughly 0.000016% chance of failing in a given year (0.0011^3 * 6 * 5 * 4) (chance of first drive failing = 6x(single), second=5x(single) etc.), assuming nobody replaced a failed disk, since you need 3 drives to fail.
Now an array of Seagate's longest running drive (2.66% failure rate) of the same size would have a chance of failure of ~0.225853% in a given year.
But, the cost of the disks in the Hitachi array would be 6x158 = $948 while the price of the disks in the Seagate array would be 6x114 = $684. For the price of the Hitachi disks you can get 8 Seagate disks.
So what happens if you just add the extra Seagate disks to the array as extra parity? Now you need 5 drive failures, giving an equation that looks like ((8 * 7 * 6 * 5 * 4 * (0.0266^5)) * 100), and a chance of failure of 0.008949%, still way higher than the Hitachi.
In the end buying Hitachi/HGST seems like the right choice anyway but I thought it interesting since I hadn't seen anyone else look at things this way.
If anyone has any problems with my math, please feel free to point it out, my stats background is pretty limited.
So if we assume that we can safely operate at some fixed level of redundancy (say 3 copies) across all hard drive vendors, the only question is (how many drives you need to replace per year) × (the price of those drives).
Obviously the type of redundancy used will depend on your specific application requirements, including performance.
As for paying for the power+drive bays, that's true of the consumer NAS devices but if you buy used servers instead it isn't such a problem. For the price of a 12 bay Synology disk station you can buy a 24 bay, 4U, used Supermciro 846E16-R1200B with a couple of CPUs and 144GB of RAM.
Not a great choice if you're short on space but can be decent otherwise.
I saw this happen once. A drive died, it was replaced and while the array was rebuilding several more failed. They ended up losing the entire SAN.
It wasn't a good day.
Of course, that doesn't work at a large scale, but then you have other methods of increasing redundancy.
Looks like it's this one: http://www.newegg.com/Product/Product.aspx?Item=N82E16822145...
Would anyone advise against getting a http://www.seagate.com/consumer/backup/expansion-portable/?s... drive?
If you want to use more bandwidth in Backblaze, after you install find the Backblaze Control Panel, open the "Settings..." dialog, go to the "Performance" tab and UNSELECT "Automatic Throttle". Slide the manual slider all the way to the far far right. Then dial up the number of threads to 4 or 6. Seriously, you don't need to go any higher on the threads, after we use 100% of your bandwidth you can't get it to backup any faster. :-)
CrashPlan has versioning that is set (by default) to back up changed files in 15 minute increments. If a file is corrupted, CrashPlan will attempt to back up the changes as it is not a corruption detection utility.
However, it's versioning allows for a user to easily restore the uncorrupted file by selecting a date prior to the corruption occurring.
This process is explained in greater detail here:
https://support.code42.com/CrashPlan/4/Troubleshooting/Recov...
The default versioning can be adjusted, so ultimately it's the responsibility of the individual user to confirm the health of their backup and verify the versioning frequency that best fits their needs.
How can that not be the case?
I'm pretty sure the CrashPlan client a .jar, so you just need a jvm, thus broad OS compatibility.
I'd expect a backup hierarchy to just be a directory of read-only snapshots of my files - I mean, this is how I organize backups locally either way (e.g. /snapshot/YYYY-MM-DD/etc/whatever.conf).
The advantage of using a standard API like filesystems and FUSE is that you can use it easily and from the terminal without needing to deal with bloated GUIs and JVMs that are no doubt hard to install, hard to use and hard to rely on.
A JVM is also a pretty hefty requirement to have on a machine. (I personally avoid installing it all costs, mostly because everything that targets JVM seems to be crud to begin with)
https://help.backblaze.com/hc/en-us/articles/217664628-Is-Ba...
Spoiler alert, the answer is a surprisingly firm no. Not even an optimistic "we're working on it!".
Doesn't even have to be "Linux support", just open an API, we'll do the rest! Use quotas for API users if the concern is around abuse, but please help us Linux folks use your service! Please!
We support Linux now with our command line tool found here: https://www.backblaze.com/b2/docs/quick_command_line.html There is no GUI, but the command line tool can "sync" your files with an arbitrary roll back time (the Mac/Windows client are hard coded to 30 days).
For good or bad, syncing to B2 charges "per GByte". If you have less than 1 TByte this will be cheaper than the $5/month for the Mac or Windows client. But if you have more than 1 TByte on Linux, it will cost you about $5/TByte/Month to back it up.
The failure in that picture is that somebody deleted the file found at C:\Program Files (x86)\Backblaze\bzbui_interface.xml That extremely simple file contains all the strings for the Backblaze GUI in all 11 languages. What your screenshot shows is what happens when that file is missing.
If you run a Backblaze installer over the top of that installation it should repair that by placing the latest version in place as it always does. If you STILL cannot get that file installed, something is going entirely, totally hog-wild wrong with your system. Hundreds of people install every day and that file is always present.
If you were running along and suddenly that file disappeared from your computer, I would stop and prepare a full restore from 15 days ago from Backblaze. This is totally free. When files randomly disappear from the hard drive, one cause might be your hard drive is starting to die.
The failure in that picture is that somebody deleted the file found at C:\Program Files (x86)\Backblaze\bzbui_interface.xml
That extremely simple file contains all the strings for the Backblaze GUI in all 11 languages. What your screenshot shows is what happens when that file is missing.
If you run a Backblaze installer over the top of that installation it should repair that by placing the latest version in place as it always does. If you STILL cannot get that file installed, something is going entirely, totally hog-wild wrong with your system. Hundreds of people install every day and that file is always present.
If you were running along and suddenly that file disappeared from your computer, I would stop and prepare a full restore from 15 days ago from Backblaze. This is totally free. When files randomly disappear from the hard drive, one cause might be your hard drive is starting to die.
Looking at it, its clear it had some problems. The issue is the installer said it finished successfully.
I'd love to see some attempts at calculating mean time to failure (which is admittedly difficult to estimate until you have had a disk long enough to see a whole generation of them fail).
Maybe DigitalOcean can publish their SSD failure rates if any.
http://0b4af6cdc2f0c5998459-c0245c5c937c5dedcca3f1764ecc9b2f...
Intel, for example, advertises a 0.2% AFR on its data center SSDs which is much lower than observed hard drive failure rates which is 3-5%.
Does anyone know why some of the confidence intervals around the ARF are asymmetric?
I don't think Backblaze uses any of these so likely they've just pulled a stock photo of a hard drive from somewhere.
Not to mention that a device being solid state by definition precludes it from being a hard disk. Logically it just makes no sense to me.
shakes fist get off my lawn!
'sshd' is the reference command to the 'ssh daemon'.
To dknoll, SSHD for the Solid State Hybrid Disk, means that it has components of both drive systems, spinning platters and a functional amount of NAND for feequently used applications.
You could use it purely for it's SSD assuming the drive controller continues to function. Though that is a ridiculous proposal.
Which when regarded under the conventional drives at the time it was announced being primarily HDD's it makes it more clear.
You have an HDD that is a hybrid with solid state components. Logically clean cut, imo.
Hybrid disk would be the generic term.
There has been work on liquid-state storage.
So Solid State becomes the adjective modifier for providing the classification of the type of Hybrid Disk.
You could in theory have a Liquid State Hybrid Disk, or a Liquid Solid State Hybrid Drive.
"Can you bounce the SSHD?"