If you can think of every possible failure and create monitoring and reporting for it before it happens, then you're the best dev on the planet.
I kinda lost count of how many times Nagios barfed itself and reported an error while the application was fine