Notes on API and vendor monitoring
Uptime checks, incident response, SLA tracking, and the parts of vendor monitoring that actually matter when something breaks.
91 articles · page 5 of 10
Understanding SLA Metrics: MTTR, Uptime, and Incident Response
What do 99.9% and 99.99% uptime actually mean? A practical guide to SLA metrics every engineering team should track.
How PulsAPI Tracks 2463+ Cloud Services in Real-Time
A look under the hood at how PulsAPI's crawler aggregates status data from hundreds of cloud providers every 60 seconds.
Introducing Community Outage Reports: Real User Signals Before the Vendor Knows
PulsAPI's new community reports feature lets engineers submit outage signals in real time, so your team sees emerging incidents minutes before official status pages update.
How Community Reports Catch Outages 15 Minutes Before Official Status Pages Update
We analyzed 90 days of outage data across 2463 cloud services. Community user reports consistently outpace vendor status page updates, often by 10 to 20 minutes.
PulsAPI vs. StatusPage.io: Which Does Your Engineering Team Actually Need?
StatusPage.io helps you communicate outages to your customers. PulsAPI monitors your upstream vendors. These tools solve opposite problems, here's how to choose.
How to Set Up Real-Time Status Monitoring for Your Entire AWS Infrastructure
A step-by-step guide to monitoring every AWS service your stack depends on, with component-level alerts for specific regions and services, not just generic AWS health.
Cloud Outage Report: Which Services Had the Most Downtime in Q1 2026
PulsAPI analyzed 1,240 incidents across 2463 cloud services in Q1 2026. Here's which services had the most outages, the longest MTTR, and the worst SLA compliance.
How to Build an Incident Response Runbook for Third-Party Cloud Outages
Most incident runbooks only cover outages you cause. Here's a template for handling third-party vendor outages, from detection to customer communication to postmortem.
Stripe Is Down: What to Do When Your Payment Processor Has an Outage
A practical guide for engineering and product teams, how to detect Stripe outages early, minimize customer impact, and communicate transparently while you wait for recovery.