143 karma · joined September 27, 2020
But...
“Can you check /var/log/messages and see if there’s messages every 30 minutes about ENA going down and then back up?”
Isn't this "sysadmin 101" ? Like... the first thing to check on any server exhibiting weird behaviour ? :-) A message about a NIC going up & down every 30min would have triggered many here instantly.
Interesting journey nevertheless!
On production systems "reclaimPolicy: Retain" on the storageClass feels like a no brainer to (mostly) avoid such disaster.
The before/after graphs are impressive.
Thanks Cloudflare falks
This project makes is super easy to setup a resilient and highly-available postgresql cluster.
And since the postgresql client lib handle connection to multiple replica... no need for some kind of load-balancer (pg_bouncer, pgpool...) in front of it anymore (even if they can still be useful sometimes).