A Loud Sound Shut Down a Bank's Data Center for 10 Hours
motherboard.vice.com
motherboard.vice.com
Shades of trying to do laser holography where everything has to remain perfectly inert for seconds on end.
Such a shame that Sun was acquired by Oracle.
For many reasons, but what makes you mention it here; what's the relevance?
Now the nice part: next day, the bank issues a press release stating that malfunction appeared after a programmed [fire suppression] test.
While accidents can happen, this looks more like "fail to plan is plan to fail" issue, as well as very bad communication.
The irony is that these systems are maintained by IBM :)
The issue in the article is that as spinners have become more dense, the track tolerance has decreased. Those small tolerance high density disks must already be expensive, and if they further can't reliably survive the environment (which includes fire suppression), then replacement with SSDs starts to look mandatory on the near horizon.
It's only a matter of time before SSD prices cut below HDD and all that spinning rust is thrown in the garbage.
Also I'm not sure what datacentres you work in but I've never heard of one being turned off for six months. You can spin down drives when they're not in use, but that's mostly to save energy. An idle SSD uses almost no power, there's no reason to spin it down.
> Digital evidence storage for legal matters is a common practice. As the use of Solid State Drives (SSD) in consumer and enterprise computers has increased, so too has the number of SSDs in storage increased. When most, if not all, of the drives in storage were mechanical, there was little chance of silent data corruption as long as the environment in the storage enclosure maintained reasonable thresholds. The same is not true for SSDs.
...
> For client application SSDs, the powered-off retention period standard is one year while enterprise application SSDs have a powered-off retention period of three months. These retention periods can vary greatly depending on the temperature of the storage area that houses SSDs.
...
> The standards change dramatically when you consider JEDEC's standards for enterprise class drives. The storage standard for this class of drive at the same operating temperature as the consumer class drive drops from 2 years under optimal conditions to 20 weeks. Five degrees of temperature rise in the storage environment drops the data retention period to 10 weeks. Overall, JEDEC lists a 3-month period of data retention as the standard for enterprise class drives.
> A check of various drive manufacturers, in this case Samsung, Intel, and Seagate, shows that their ratings for data retention of their consumer class drives are what would be expected for JEDEC's enterprise class drive standards. All three quote a nominal 3-month retention time period. Most likely, the manufacturers are being conservative; however, it demonstrates the potential variability the manufacturers associate with data retention on any SSD in storage.
I know there's a ridiculous amount of offline data that must be maintained and tapes aren't always a practical solution.
Manufacturers are being very conservative when it comes to retention times. I haven't heard of anyone losing data because they left their SSD powered off for two long, but if you have any anecdotes or reports to share, by all means.
Early SSDs were temperamental, flaky, and would burn out quickly. The current generation is durable, almost impossible to burn out, and seems to hold data for years even when powered off. It's like the bad rap that plasma screens got for burning in even when that problem was addressed by the manufacturers.
> “Moreover, to ensure full integrity of the data, we’ve made an additional copy of our database before restoring the system,” ING’s press release reads.
What? Am I to interpret this to mean that ING has a single backup of its data under normal conditions?
The last halon datacentre I was in had breathing apparatus within easy reach no matter where you were standing in the place. Two paces, at most, to grab a mask. You do have seconds if it goes off, and that mask is only to give you enough time to run out of the room.
The newer stuff is less hostile to organic life, so there's no reason for a respirator.
http://webcache.googleusercontent.com/search?q=cache%3Awww.i...
If so it's quite a big co-incidence, and goes to show how hard it is to mitigate against every event.
According to people familiar with the system, the
pressure at ING Bank's data center was higher than
expected, and produced a loud sound when rapidly
expelled through tiny holes.
The bank monitored the sound and it was very loud,
a source familiar with the system told us. “It was
as high as their equipment could monitor, over
130dB”.
Sound means vibration, and this is what damaged the
hard drives.
Those are the words of the author, in that last sentence. It's simply journalism at this point. It's not something worth construing as a technical assessment.They (the staff at Vice) just want to scoop the article, and get Vice into the action. Was it a siren that was too loud? Was it pressure differential, triggering a head crash in a bernoulli box?
Since we're at least two degrees removed from the actual events, and we'll probably never get direct information from a postmortem report, this article, to me, reads as: Data Center Outage in Eastern Europe, Reason Unknown