Datadog Outage History

60 incidents recorded over the last 365 days, 60 of them resolved. Durations are measured from when PulsAPI first saw the problem to when it cleared, which is usually longer than the vendor's own figure.

Every Recorded Datadog Incident

Datadog incidents, newest first, with date, duration and status.
DateIncidentDurationStatus
Aug 6, 2026Delayed Processes dataThis incident has been resolved.46 minResolved
Jul 30, 2026Delayed Evaluation of Service Check monitorsThis incident has been resolved.42 minResolved
Jul 29, 2026Real user monitoring, distribution and custom processes metrics delayedThis incident has been resolved.1h 46mResolved
Jul 27, 2026Synthetics Test Results DelayedThis incident has been resolved.46 minResolved
Jul 23, 2026Latency with Pages and Monitor state changesThis incident has been resolved.1h 4mResolved
Jul 10, 2026Delayed data streams dataThis incident has been resolved.1h 50mResolved
Jul 3, 2026Errors Pulling Container ImagesThis incident has been resolved.1h 3mResolved
Jul 1, 2026Elevated Error RatesThis incident has been resolved.2h 5mResolved
Jun 30, 2026Delayed data processing and errors in multiple productsThis incident has been resolved.2h 40mResolved
Jun 30, 2026Delayed Monitors NotificationsThis incident has been resolved.43 minResolved
Jun 26, 2026Delayed APM Trace MetricsThis incident has been resolved.6 minResolved
Jun 25, 2026Delayed CI Visibility dataThis incident has been resolved.9 minResolved
Jun 22, 2026Metrics QueriesThis incident has been resolved.1h 13mResolved
Jun 18, 2026Delayed EventsThis incident has been resolved.20 minResolved
Jun 16, 2026Delayed AWS/GCP/Azure MetricsA fix has been implemented and we are monitoring the results.1154h 17mResolved
Jun 12, 2026Delayed EvaluationThis incident has been resolved.2h 27mResolved
Jun 8, 2026Increased latency across multiple productsThis incident has been resolved.1h 54mResolved
May 16, 2026Azure Metrics ReportingThis incident has been resolved and Azure metrics are reporting as expected.1h 25mResolved
May 14, 2026Delayed Metric LoadingThis incident has been resolved.1h 23mResolved
May 13, 2026Degraded Web Application PerformanceThis incident has been resolved.36 minResolved
May 8, 2026Delayed Monitors NotificationsThis incident has been resolved.14h 0mResolved
May 5, 2026Delays in APM-based Monitor NotificationsThis incident has been resolved.1h 38mResolved
May 1, 2026Elevated Error Rates for Metrics QueriesThis incident has been resolved.34 minResolved
Apr 24, 2026Delays in AWS and Azure cloud integration metrics ingestionThis incident has been resolved.23 minResolved
Apr 15, 2026Delayed MetricsThis incident has been resolved.1h 59mResolved
Apr 10, 2026Notifications to Pagerduty, VictorOps, Slack, Webhook, Microsoft Teams and OpsGenie are not deliveredBetween approximately 9:42 AM and 10:38 AM ET on April 10, we observed delivery failures for customers in the US1, AP1, and AP2 regions. During this window, notifications to the following integrations were delayed or…6h 5mResolved
Mar 27, 2026Delayed Traces in APM Trace SearchThis incident has been resolved.32 minResolved
Mar 26, 2026Delayed Monitors NotificationsThis incident has been resolved.5h 32mResolved
Mar 19, 2026Intermittent "No Data" Status for MonitorsThis incident has been resolved.1h 59mResolved
Mar 13, 2026Web Application Not LoadingThis incident has been resolved.1h 3mResolved
Feb 24, 2026Delayed RUM dataThis incident has been resolved.2h 36mResolved
Feb 18, 2026Monitors - Delayed Evaluation for Multiple ProductsThis incident has been resolved.1h 6mResolved
Feb 12, 2026Delayed Traces in APMThis incident has been resolved.1h 10mResolved
Feb 11, 2026Delayed Traces in APM Trace SearchThis incident has been resolved.2h 48mResolved
Feb 5, 2026Delays in Monitor EvaluationsThis incident has been resolved.2h 18mResolved
Jan 29, 2026Delayed Distribution Monitors EvaluationsThis incident has been resolved.2h 19mResolved
Jan 28, 2026Monitors - Delayed EvaluationThis incident has been resolved.3h 5mResolved
Jan 22, 2026Web Application Not LoadingThis incident has been resolved.36 minResolved
Jan 18, 2026Delayed EventsThis incident is resolved. There's no more delay for the processing of Events, nor impact on the event stream, event based widgets and event based monitors.1h 51mResolved
Dec 12, 2025Delayed Processes dataThis incident has been resolved.1h 54mResolved
Dec 12, 2025Delayed APM metric ingestionAll impact related to APM metrics has been resolved. A separate incident has been created to track the remaining impact in live process data.3h 15mResolved
Dec 9, 2025Metrics data ingestion delayed and monitor evaluations degradedThis incident has been resolved. Live data is being processed normally and gaps in distribution metrics on graphs will be backfilled within the next hour.52 minResolved
Nov 19, 2025Web Application Not LoadingThis incident has been resolved as of 2:32PM ET.36 minResolved
Nov 18, 2025Delayed Monitors NotificationsThis incident has been resolved. Notification delays were only affecting our internal monitoring and were due to the ongoing Cloudflare incident: https://www.cloudflarestatus.com/incidents/8gmgl950y3h7/.36 minResolved
Nov 17, 2025Dashboards Not LoadingAll errors stopped as of 12:02ET. This incident has been resolved.56 minResolved
Nov 13, 2025Web Application Not LoadingThis incident has been resolved.9 minResolved
Nov 5, 2025Delayed Metrics for APM and distribution metricsAPM metrics are now processing live.2h 14mResolved
Oct 28, 2025Metrics data ingestion delayed and monitor evaluations degradedThis incident has been resolved.3h 29mResolved
Oct 20, 2025Multiple products impacted with data delaysBackfills for Metrics and Log Management data have completed. All systems are back to normal.49h 54mResolved
Oct 20, 2025Multiple products impacted with data delaysThis incident has been resolved.1h 16mResolved
Oct 14, 2025Delayed AWS, GCP, Azure, SaaS integrations Metrics and LogsThis incident has been resolved.2h 4mResolved
Oct 1, 2025Delayed MetricsWe’ve confirmed that this issue only impacts customers using the OCI integrations feature. The vast majority of customers are not impacted. Impacted customers will see an in-app banner when visiting any Datadog product…1h 32mResolved
Sep 29, 2025Host Tags, Service Checks, and Datadog Events Delayed EvaluationThis incident has been resolved.1h 25mResolved
Sep 18, 2025[SSO] Login Errors from Google SSOThis incident has been resolved.1h 13mResolved
Sep 17, 2025Delayed MetricsThis incident has been resolved.42 minResolved
Sep 5, 2025Delayed Monitors NotificationsThis incident has been resolved.25 minResolved
Sep 2, 2025Delayed RUM dataThis incident has been resolved.1h 3mResolved
Aug 28, 2025Periodic network interruption communicating with multiple Azure regionsOur monitoring has shown Azure’s fix to be stable since our last update. This incident has been resolved.47h 28mResolved
Aug 28, 2025Pagerduty Monitor Notifications DelayedPagerDuty notifications deliveries are back to normal.4h 26mResolved
Aug 27, 2025Partial metrics drop from Datadog Agent in the westus2 azure region to Datadog us1 datacenterWe noticed partial data drop from Datadog Agent in the westus2 azure region to Datadog us1 datacenter. There is no data drop anymore, we are monitoring the situation.2h 54mResolved

How to Read This Datadog Incident Log

Each row is an incident PulsAPI observed, not a summary written afterwards. The duration is wall-clock time between the first failing check and the first clean one, so it includes the window before Datadog acknowledged anything. Vendor post-mortems typically measure from acknowledgement, which is why their numbers are usually shorter.

Incidents still open have no duration yet and are listed as ongoing rather than being given a running total. The archive covers the last 365 days; anything older has aged out of the window rather than never having happened.

The longest single Datadog outage in this window ran 48d 2h, against a mean recovery of 22h 37m. If you depend on Datadog in a customer-facing path, the longest figure is the one to design around. The mean is what happens on a normal bad day; the maximum is what happens on the worst one.