Redditor acquires decommissioned Netflix cache server with 262TB of storage
arstechnica.com
arstechnica.com
However, we knew someone in the building opposite who was responsible for the company's streaming TV service. They were retiring a whole bunch of CDN machines, similar to these, which had a crazy amount of storage per node.
While we weren't allowed to buy servers, we were allowed to buy components, so I got a bunch of the retired servers and bought enough RAM to max them out.
Made a lovely ES cluster :)
I'm curious. Could you have bought all the components to build your own servers under that policy? :-D
Not so similar after all
Is that what happened? I can imagine all sorts of scenarios where that's not true. You got some better info than the above comment on this?
One piece at a time, Johnny Cash.
I noticed the display too, you can still buy them https://www.crystalfontz.com/product/cfa635tmlku-display-mod...
I remember these units too. Back in 2002-2005, I worked for a small system integrator. We built turn-key appliances for software vendors. Firewalls, security appliances, etc. We'd work with chassis manufacturers to add cutouts for these units on the front of 1U and 2U servers to provide for an interface without needing to hook up a console cart. Same as you might see on a Dell or HP.
Really neat seeing these still in use!
But I definitely do remember the era of cold cathodes:)
https://old.reddit.com/r/homelab/comments/ydollm/so_i_got_a_...
Followup Post
https://old.reddit.com/r/homelab/comments/ydollm/so_i_got_a_...
Darn those were nice times.
Reading at -1 was a bit of a chore, but also sometimes fun.
Whereas a Beowulf cluster is grungy, full of bailing twine, chewing gum, and possessing high levels of Bodge. And dirt cheap, built with what you can get, often used.
"Beowulf" was just the name of an early machine built that way at NASA in the 90s. The approach is not unusual now; many "low-end" supercomputers are just a few racks of commodity servers. But it was very unconventional in the early 90s, which is why the Beowulf term existed, to distinguish it from "real" clusters.
Hardware Ethernet switches -- able to stream from port to port at full speed -- were the revolutionary component. Before Ethernet, clusters required specialized hardware implementing a crossbar-like switched fabric, so that all the nodes could communicate at high speed with each other, with a minimum of hops. These were solutions for clustering mainframes and minicomputers, horribly expensive and proprietary.
At the same time, it demonstrated you could build a useful and inexpensive cluster using white-box linux desktop machines sitting on wire shelving. I did exactly the same thing at the time, except for the channel bonding, which would have helped my cluster scaling significantly, but I didn't have any budget left over for more NICs and since I was using a hub, there would have been lots of collisions anyway.
my last /. comment was 2014, jfc
https://pastebin.com/raw/E39HuK59
Run SMART, check the stats, make sure the grown defect list is 0, make sure it's never been over-temp, run a SMART long test, run bad blocks, run another SMART long test. If it passes all those I'd be fine continuing to use it.
The HUH728080ALE600 drives in that server[1] idle at 5.1 watts[2], so it's 184 watts just on the drives. I'm guessing at idle the server runs around 300 watts. Which is not great, but it's not terrible. I pay ~ $1/watt/year in NC, so running that server 24x7 would cost me $300 annually.
If it were me I'd disconnect at least half the drives to keep as spares.
1. https://imgur.com/a/zb6Mqty
2. https://documents.westerndigital.com/content/dam/doc-library...
My backup server is still running almost 17-18 years old 500GB hard drives. They will be retired soon though, as they consume too much power and lack spindown when not in use.
Thats very good for longevity of the disks
And of course, 36x the drives means 36x the drive failures - and even if you avoid losing data, you've still got the chore of swapping each failed drive.
I think you must mean 36x the chance of drive failure. Drive failure probability per year is published at around 1% though practically as high as 10%. 36 drives means there's somewhere between a 30.35% (1 - 0.99^36) and 97.7% (1 - 0.9^36) chance per year of a single drive failing.
... No.
When you hit hundreds of concurrent operations there is no such thing as sequential. Also there is no need for sequential reads because a stripe of 7-15 disks would give you much more than streaming bitrate needs.
Which begs the question of how old the drives actually are. If intensive usage would result in lots of disk failures after 5 years, then a 9-year-old server would surely have a bunch of new(er) disks inside it? Or do they just reduce the amount of cache space they have with every failure?
https://old.reddit.com/r/homelab/comments/ydollm/so_i_got_a_...
I was under the impression that constant on-off was what put the most wear on a drive, and that if it is constantly on it would last longer.
Maybe my old age is showing, but this just frightens me to no end. 1 50TB failure, and you've lost 50TB. Build an array of smaller drives to get to 50TB, and much less catastrophe if 1 drive dies.
Old disks are always risky. Avoid.
Though I assume that's at typical Netflix workload levels.
- https://www.theverge.com/22787426/netflix-cdn-open-connect
HR later sends out an email regarding the printing of PDFs unnecessarily. Think of the trees that could be saved by not printing.
https://news.ycombinator.com/item?id=33350152 ("I got a Netflix cache server. 262TB", 27 comments)
HBO also uses multiple CDNs, here is a reference from 2020 about GoT: https://www.prnewswire.com/news-releases/hbo-streams-game-of....
They were also mad that one of their distribution partners last week leaked the finale of House of the Dragon early: https://www.theguardian.com/tv-and-radio/2022/oct/21/house-o....
I don't have a link handy, but even Amazon has been known to use CDNs outside of CloudFront for their live events.
I used to work for one of the companies supporting studios getting their content to streamers. This specific story is related to iTunes, but it's easy to do for any of them. The early days of "ramping up" meant putting butts in seats to do the work before automation was in place. The push to automate shot to highest priority when on of the butts in seats incorrectly copy/paste from one spreadsheet to the next which allowed an episode to be downloaded before it aired to anyone that subscribed to that season on iTunes.
The fallout from that was incredible, but credit to the company they did not fire the employee for making a human mistake.
From there they have a "fill" window each day during the ISP's low traffic period where Netflix pushes new content to it.
> Each Open Connect Appliance (OCA) stores a portion of the Netflix catalog, which in general is less than the complete content library for a given region. Popularity changes, new titles that are added to the service, re-encoded movies, and routine software enhancements are all part of the nightly updates, or fill, that each appliance must download to remain current.
-- https://openconnect.netflix.com/en/deployment-guide/
I mean honestly in 2022 their catalogue might fit entirely ( :P ) but yeah looks like it's just a portion, comprised of whatever's popular in the region
That and they had different tiers of hardware I believe.
From my own poor recollection, their dimensions were about 2' x 3' x 1.5' - and obviously quite heavy.
They stored 1GB each.
But that is not the caseless DIY shelf-servers they've displayed in various PR videos.
According to Wikipedia, the 2U models were indeed based on Dell PowerEdge servers, and the 1U models were apparently based on Super Micro servers.[^1]
They figured out that it was more efficient to have hardware techs spend 100% of their time provisioning new hardware, and just let their software detect the broken servers, power them down, and route the tasks to working servers.
Depending on what specifically you're after, there's a number of them on ebay: https://www.ebay.com/sch/i.html?_nkw=google+search+appliance
They want the spyware to run on your hardware. They don't need it on their hardware. If there is spyware on Goog's hardware, I'd be looking at it coming from certain three letter agencies or other nation states.
Kind of a letdown
> Interestingly, the now-defunct dial-up online service Prodigy used a local caching system to distribute data more efficiently using the same basic principle as Open Connect in the 1980s and '90s.
Odd 'fact' to bring up.