SLA & SLO articles
Uptime checks, incident response, SLA tracking, and the parts of vendor monitoring that actually matter when something breaks.
5 articles in SLA & SLO
Your Uptime Is Not Your Uptime: Availability Across a Dependency Chain
Serial dependencies multiply. Ten vendors at 99.9% each give you 99.0%, or 87 hours a year. How to compute your real ceiling and what redundancy buys.
Scheduled Maintenance and the SLA Math Nobody Checks
Almost every SLA excludes planned maintenance. A 99.9% deal with a four-hour monthly window really means 99.34%, or 4.7 hours down instead of 43 minutes.
SLA Monitoring for APIs: How to Track Vendor Uptime and Prove Impact
Learn how SLA monitoring for APIs helps engineering and operations teams track vendor uptime, document downtime, calculate impact, and prepare contract evidence.
SLA vs SLO vs SLI: The Definitive Guide for Platform Engineers
Confused by the acronym soup of reliability engineering? This guide demystifies SLAs, SLOs, and SLIs with real-world examples, calculation formulas, and a calculator to find your error budget before it finds you.
Calculating Your Error Budget: A Step-by-Step Workbook
An error budget is the difference between 100% uptime and your SLO target, and it's the key to balancing feature velocity with reliability. This workbook provides formulas, real examples, and a decision framework.