[08:43 AM PDT] We have narrowed down the source of the network connectivity issues that impacted AWS Services...
[08:04 AM PDT] We continue to investigate the root cause for the network connectivity issues...
[12:11 AM PDT] <declared outage>
They claim not to have known the root cause for ~8hr
The initial cause appears to be a a bad DNS entry that they rolled back at 2:22am PDT. They started seeing recovery with services but as reports of EC2 failures kept rolling in they found a network issue with a load balancer that was causing the issue at 8:43am.
Their 14 updates did not bring my stuff back up.
My nines are not their nines. https://rachelbythebay.com/w/2019/07/15/giant/
Or an NLB could also be load balancing by managing DNS records--it's not really clear what a NLB means in this context
Or there was an overload condition because of the NLB malfunctioning that caused UDP traffic to get dropped
Obviously a lot of reading between the lines is required without a detailed RCA--hopefully they release more info