Most teams only see the internal logic of their High Availability (HA) stack during a production crash. These slides from Mydbops MyWebinar 51 provide a technical look at the interaction between etcd, Patroni, and HAProxy.
We focus on the critical 30-second window between a primary database failure and a successful standby promotion. The session breaks down the mechanics of leader election and traffic routing, showing exactly why some failovers take 15 seconds while others turn into 15-minute outages.
Technical topics included in this deck:
The Full Stack Logic: How etcd quorum and Patroni leader election sync with HAProxy health checks.
The Math of Quorum: Why etcd stops accepting writes during quorum loss and how that affects Patroni.
The Leader Lock: A detailed walkthrough of how Patroni heartbeats manage the leader key.
Log Correlation: Learning to read and sync logs across all three layers to find root causes faster.
Configuration Traps: How settings like primary_start_timeout determine your real-world recovery time.
RTO vs RPO: The architectural trade-offs between a fast failover and a safe one.
Essential reading for DBAs, SREs, and architects who need to move past basic setup and engineer a truly reliable PostgreSQL cluster.
Speaker: Manosh Malai, CTO, Mydbops
For PostgreSQL audits, consulting, and 24x7 managed services, visit www.mydbops.com.