Strategies to Mitigate Outages and Maximize Uptime (O'Reilly Architecture Superstream)
2024-06-27
Description
🔴 About this webinar If a database outage were to cost your business a million dollars an hour (or minute), how much should you invest to prevent it? For industries like banking and finance, which face growing resilience regulations with strict expectations of infrastructure hardiness, the stakes are even higher. Sean Chittenden and Chris Casano, experienced former owners of massive applications, will share zero downtime strategies that organizations can adopt to mitigate outage risks and maximize uptime. 👥 Meet our speakers Sean Chittenden, Senior Director of Engineering at Cockroach Labs, is an expert in databases and high-availability distributed systems, renowned for proficiency in PostgreSQL and CockroachDB, and dedicated to delivering practical insights into contemporary database technologies. Chris Casano is the Senior Manager for Sales Engineering at Cockroach Labs. Previously, he was a senior sales engineer at Cloudera Hortonworks and Oracle. Strategies to Mitigate Outages and Maximize Uptime (O'Reilly Architecture Superstream) 00:00 Introduction 00:38 Today's resiliency challenges 01:16 What is typical maintenance for databases? 02:18 How to maintain CockroachDB 03:16 What does unplanned downtime for databases look like? 04:11 Unplanned downtime with CockroachDB 04:39 Zero downtime: why bother? 05:55 Architecting for resiliency 07:27 Best practices for CockroachDB on AWS 08:22 What does necessary infrastructure complexity look like? 08:55 How CockroachDB can help you simplify your tech stack 09:14 An architectural deep-dive 10:56 What makes Distributed SQL different? 11:37 CockroachDB's layered architecture 12:01 Distributed SQL enabled by a KV data store 12:33 KV DB built on ranges 13:11 Ship faster: Decouple use from operations 14:41 Everyone's an SQL gateway! 15:35 Distributed SQL: Zero SPOFs 16:16 Tip 1: How to teach CockroachDB to distribute data (cockroach start locality) 19:53 Unlock a zero downtime journey 20:49 A look at the operational math 21:39 Operational layering 22:32 Tip 2: Deploy with redundant EBS