Man that's got to suck.
Man that's got to suck.
I HAVE NO TOOLS BECAUSE I’VE DESTROYED MY TOOLS WITH MY TOOLS.A very simple example, you do something stupid on a remote machine (either high network usage or CPU usage) over SSH then you can't undo it because SSH becomes unresponsive
Bohr bugs generate an alert and happily meander through normal support channels.
Heisenbugs go through phases -
1. Probation. On continued failure,
2. Restart. If the app or service fails after a restart,
3. Reboot. If the app or service fails after a reboot,
4. Re-image. If the app or service fails after re-imaging,
5. Remove/elimate the node.
The article states:
> The defense in depth philosophy means we have robust backup plans for handling failure of such tools, but use of these backup plans (including engineers travelling to secure facilities designed to withstand the most catastrophic failures, and a reduction in priority of less critical network traffic classes to reduce congestion) added to the time spent debugging.
Another alternative is low bandwidth flag based roll-backs (for instances such as this where the network is congested but not completely lost).
My point being, it seems that modems are becoming less-and-less viable for out-of-band management.