How we systematically improved our reliability
Earlier this year, we had two incidents in one week. After those incidents, we doubled down on reliability. This post summarizes the reliability work we've been rolling out since then to keep uptime high as our fleet, product, and customer base grow.
Sources
- T1How we systematically improved our reliabilityDatabricks / ClickHouse / DuckDB / Supabase / Neon / PlanetScale