Datadog Outage History
60 incidents recorded over the last 365 days, 60 of them resolved. Durations are measured from when PulsAPI first saw the problem to when it cleared, which is usually longer than the vendor's own figure.
Every Recorded Datadog Incident
| Date | Incident | Duration | Status |
|---|---|---|---|
| Aug 6, 2026 | Delayed Processes dataThis incident has been resolved. | 46 min | Resolved |
| Jul 30, 2026 | Delayed Evaluation of Service Check monitorsThis incident has been resolved. | 42 min | Resolved |
| Jul 29, 2026 | Real user monitoring, distribution and custom processes metrics delayedThis incident has been resolved. | 1h 46m | Resolved |
| Jul 27, 2026 | Synthetics Test Results DelayedThis incident has been resolved. | 46 min | Resolved |
| Jul 23, 2026 | Latency with Pages and Monitor state changesThis incident has been resolved. | 1h 4m | Resolved |
| Jul 10, 2026 | Delayed data streams dataThis incident has been resolved. | 1h 50m | Resolved |
| Jul 3, 2026 | Errors Pulling Container ImagesThis incident has been resolved. | 1h 3m | Resolved |
| Jul 1, 2026 | Elevated Error RatesThis incident has been resolved. | 2h 5m | Resolved |
| Jun 30, 2026 | Delayed data processing and errors in multiple productsThis incident has been resolved. | 2h 40m | Resolved |
| Jun 30, 2026 | Delayed Monitors NotificationsThis incident has been resolved. | 43 min | Resolved |
| Jun 26, 2026 | Delayed APM Trace MetricsThis incident has been resolved. | 6 min | Resolved |
| Jun 25, 2026 | Delayed CI Visibility dataThis incident has been resolved. | 9 min | Resolved |
| Jun 22, 2026 | Metrics QueriesThis incident has been resolved. | 1h 13m | Resolved |
| Jun 18, 2026 | Delayed EventsThis incident has been resolved. | 20 min | Resolved |
| Jun 16, 2026 | Delayed AWS/GCP/Azure MetricsA fix has been implemented and we are monitoring the results. | 1154h 17m | Resolved |
| Jun 12, 2026 | Delayed EvaluationThis incident has been resolved. | 2h 27m | Resolved |
| Jun 8, 2026 | Increased latency across multiple productsThis incident has been resolved. | 1h 54m | Resolved |
| May 16, 2026 | Azure Metrics ReportingThis incident has been resolved and Azure metrics are reporting as expected. | 1h 25m | Resolved |
| May 14, 2026 | Delayed Metric LoadingThis incident has been resolved. | 1h 23m | Resolved |
| May 13, 2026 | Degraded Web Application PerformanceThis incident has been resolved. | 36 min | Resolved |
| May 8, 2026 | Delayed Monitors NotificationsThis incident has been resolved. | 14h 0m | Resolved |
| May 5, 2026 | Delays in APM-based Monitor NotificationsThis incident has been resolved. | 1h 38m | Resolved |
| May 1, 2026 | Elevated Error Rates for Metrics QueriesThis incident has been resolved. | 34 min | Resolved |
| Apr 24, 2026 | Delays in AWS and Azure cloud integration metrics ingestionThis incident has been resolved. | 23 min | Resolved |
| Apr 15, 2026 | Delayed MetricsThis incident has been resolved. | 1h 59m | Resolved |
| Apr 10, 2026 | Notifications to Pagerduty, VictorOps, Slack, Webhook, Microsoft Teams and OpsGenie are not deliveredBetween approximately 9:42 AM and 10:38 AM ET on April 10, we observed delivery failures for customers in the US1, AP1, and AP2 regions. During this window, notifications to the following integrations were delayed or… | 6h 5m | Resolved |
| Mar 27, 2026 | Delayed Traces in APM Trace SearchThis incident has been resolved. | 32 min | Resolved |
| Mar 26, 2026 | Delayed Monitors NotificationsThis incident has been resolved. | 5h 32m | Resolved |
| Mar 19, 2026 | Intermittent "No Data" Status for MonitorsThis incident has been resolved. | 1h 59m | Resolved |
| Mar 13, 2026 | Web Application Not LoadingThis incident has been resolved. | 1h 3m | Resolved |
| Feb 24, 2026 | Delayed RUM dataThis incident has been resolved. | 2h 36m | Resolved |
| Feb 18, 2026 | Monitors - Delayed Evaluation for Multiple ProductsThis incident has been resolved. | 1h 6m | Resolved |
| Feb 12, 2026 | Delayed Traces in APMThis incident has been resolved. | 1h 10m | Resolved |
| Feb 11, 2026 | Delayed Traces in APM Trace SearchThis incident has been resolved. | 2h 48m | Resolved |
| Feb 5, 2026 | Delays in Monitor EvaluationsThis incident has been resolved. | 2h 18m | Resolved |
| Jan 29, 2026 | Delayed Distribution Monitors EvaluationsThis incident has been resolved. | 2h 19m | Resolved |
| Jan 28, 2026 | Monitors - Delayed EvaluationThis incident has been resolved. | 3h 5m | Resolved |
| Jan 22, 2026 | Web Application Not LoadingThis incident has been resolved. | 36 min | Resolved |
| Jan 18, 2026 | Delayed EventsThis incident is resolved. There's no more delay for the processing of Events, nor impact on the event stream, event based widgets and event based monitors. | 1h 51m | Resolved |
| Dec 12, 2025 | Delayed Processes dataThis incident has been resolved. | 1h 54m | Resolved |
| Dec 12, 2025 | Delayed APM metric ingestionAll impact related to APM metrics has been resolved. A separate incident has been created to track the remaining impact in live process data. | 3h 15m | Resolved |
| Dec 9, 2025 | Metrics data ingestion delayed and monitor evaluations degradedThis incident has been resolved. Live data is being processed normally and gaps in distribution metrics on graphs will be backfilled within the next hour. | 52 min | Resolved |
| Nov 19, 2025 | Web Application Not LoadingThis incident has been resolved as of 2:32PM ET. | 36 min | Resolved |
| Nov 18, 2025 | Delayed Monitors NotificationsThis incident has been resolved. Notification delays were only affecting our internal monitoring and were due to the ongoing Cloudflare incident: https://www.cloudflarestatus.com/incidents/8gmgl950y3h7/. | 36 min | Resolved |
| Nov 17, 2025 | Dashboards Not LoadingAll errors stopped as of 12:02ET. This incident has been resolved. | 56 min | Resolved |
| Nov 13, 2025 | Web Application Not LoadingThis incident has been resolved. | 9 min | Resolved |
| Nov 5, 2025 | Delayed Metrics for APM and distribution metricsAPM metrics are now processing live. | 2h 14m | Resolved |
| Oct 28, 2025 | Metrics data ingestion delayed and monitor evaluations degradedThis incident has been resolved. | 3h 29m | Resolved |
| Oct 20, 2025 | Multiple products impacted with data delaysBackfills for Metrics and Log Management data have completed. All systems are back to normal. | 49h 54m | Resolved |
| Oct 20, 2025 | Multiple products impacted with data delaysThis incident has been resolved. | 1h 16m | Resolved |
| Oct 14, 2025 | Delayed AWS, GCP, Azure, SaaS integrations Metrics and LogsThis incident has been resolved. | 2h 4m | Resolved |
| Oct 1, 2025 | Delayed MetricsWe’ve confirmed that this issue only impacts customers using the OCI integrations feature. The vast majority of customers are not impacted. Impacted customers will see an in-app banner when visiting any Datadog product… | 1h 32m | Resolved |
| Sep 29, 2025 | Host Tags, Service Checks, and Datadog Events Delayed EvaluationThis incident has been resolved. | 1h 25m | Resolved |
| Sep 18, 2025 | [SSO] Login Errors from Google SSOThis incident has been resolved. | 1h 13m | Resolved |
| Sep 17, 2025 | Delayed MetricsThis incident has been resolved. | 42 min | Resolved |
| Sep 5, 2025 | Delayed Monitors NotificationsThis incident has been resolved. | 25 min | Resolved |
| Sep 2, 2025 | Delayed RUM dataThis incident has been resolved. | 1h 3m | Resolved |
| Aug 28, 2025 | Periodic network interruption communicating with multiple Azure regionsOur monitoring has shown Azure’s fix to be stable since our last update. This incident has been resolved. | 47h 28m | Resolved |
| Aug 28, 2025 | Pagerduty Monitor Notifications DelayedPagerDuty notifications deliveries are back to normal. | 4h 26m | Resolved |
| Aug 27, 2025 | Partial metrics drop from Datadog Agent in the westus2 azure region to Datadog us1 datacenterWe noticed partial data drop from Datadog Agent in the westus2 azure region to Datadog us1 datacenter. There is no data drop anymore, we are monitoring the situation. | 2h 54m | Resolved |
How to Read This Datadog Incident Log
Each row is an incident PulsAPI observed, not a summary written afterwards. The duration is wall-clock time between the first failing check and the first clean one, so it includes the window before Datadog acknowledged anything. Vendor post-mortems typically measure from acknowledgement, which is why their numbers are usually shorter.
Incidents still open have no duration yet and are listed as ongoing rather than being given a running total. The archive covers the last 365 days; anything older has aged out of the window rather than never having happened.
The longest single Datadog outage in this window ran 48d 2h, against a mean recovery of 22h 37m. If you depend on Datadog in a customer-facing path, the longest figure is the one to design around. The mean is what happens on a normal bad day; the maximum is what happens on the worst one.