Microsoft shelves its underwater data center
tomshardware.com
tomshardware.com
It doesn't sound like they really discontinued the project. It sounds like it got moved to another team.
They mention liquid immersion, which is a less extreme take on the idea of trading lower failure rates for lower maintainability.
1 https://www.datacenterdynamics.com/en/news/microsoft-confirm...
That therefore begs the question: what are the narrowest points of failure in a system like this? For example it would be naive to have 140 fungible compute nodes if they all went through a single networking switch. Node workloads can be moved but a switch failure makes the entire container useless.
This works physically as well. Marine vessels have sealed bulkheads such that a hull breach in one end doesn’t cause flooding in the other end.
How do these vessels fan out all the other points of failure and what are the error rates in those components?
Their biggest point of failure is probably the land connection. Looks like it's one big cable bundle. Those get damaged all the time from sea life, fishing, people throwing anchors where they shouldn't etc. Land-based datacenters have the same issue with excavators damaging power or network lines, but it's easy to run multiple lines far apart from each other when you are on land.
I wonder what factors led them to discontinue the project.
On the other hand 0.7% dead hardware sounds like a marginal cost. What likely killed it was the cost to deploy and recover them, plus all the hassle with sea cables for data and power. Building glorified warehouses with redundant power and AC is likely a lot cheaper than deploying these metal tanks in the ocean.
Having the servers more easy to access and repair also makes them more likely to fail. Also spaces that humans can enter require a bunch of regulation that takes up space and adds to the expense. So the math could easily work out in favor of a sealed box of computers that no-one is allowed to access and any broken servers just get switched off.