Grafana Cloud Outage History
193 incidents recorded over the last 365 days, 162 of them resolved. Durations are measured from when PulsAPI first saw the problem to when it cleared, which is usually longer than the vendor's own figure.
Every Recorded Grafana Cloud Incident
| Date | Incident | Duration | Status |
|---|---|---|---|
| Sep 10, 2026 | Scheduled Database Maintenance – Temporary Grafana Instance UnavailabilityYour Grafana instance will be temporarily unavailable due to scheduled maintenance. Downtime is expected to last between 5 to 15 minutes depending on your instance size. Alerts and scheduled reports may be impacted… | 7h 0m | Scheduled |
| Sep 9, 2026 | Scheduled Database Maintenance – Temporary Grafana Instance UnavailabilityYour Grafana instance will be temporarily unavailable due to scheduled maintenance. Downtime is expected to last between 5 to 15 minutes depending on your instance size. Alerts and scheduled reports may be impacted… | 7h 0m | Scheduled |
| Sep 9, 2026 | Scheduled Database Maintenance – Temporary Grafana Instance UnavailabilityYour Grafana instance will be temporarily unavailable due to scheduled maintenance. Downtime is expected to last between 5 to 15 minutes depending on your instance size. Alerts and scheduled reports may be impacted… | 7h 0m | Scheduled |
| Sep 8, 2026 | Scheduled Database Maintenance – Temporary Grafana Instance UnavailabilityYour Grafana instance will be temporarily unavailable due to scheduled maintenance. Downtime is expected to last between 5 to 15 minutes depending on your instance size. Alerts and scheduled reports may be impacted… | 7h 0m | Scheduled |
| Aug 28, 2026 | Scheduled Database Maintenance – Temporary Grafana Instance UnavailabilityYour Grafana instance will be temporarily unavailable due to scheduled maintenance. Downtime is expected to last between 5 to 15 minutes depending on your instance size. Alerts and scheduled reports may be impacted… | 4h 0m | Scheduled |
| Aug 27, 2026 | Scheduled Database Maintenance – Temporary Grafana Instance UnavailabilityYour Grafana instance will be temporarily unavailable due to scheduled maintenance. Downtime is expected to last between 5 to 15 minutes depending on your instance size. Alerts and scheduled reports may be impacted… | 4h 0m | Scheduled |
| Aug 27, 2026 | Scheduled Database Maintenance – Temporary Grafana Instance UnavailabilityYour Grafana instance will be temporarily unavailable due to scheduled maintenance. Downtime is expected to last between 5 to 15 minutes depending on your instance size. Alerts and scheduled reports may be impacted… | 4h 0m | Scheduled |
| Aug 26, 2026 | Scheduled Database Maintenance – Temporary Grafana Instance UnavailabilityYour Grafana instance will be temporarily unavailable due to scheduled maintenance. Downtime is expected to last between 5 to 15 minutes depending on your instance size. Alerts and scheduled reports may be impacted… | 4h 0m | Scheduled |
| Aug 25, 2026 | Scheduled Database Maintenance – Temporary Grafana Instance UnavailabilityYour Grafana instance will be temporarily unavailable due to scheduled maintenance. Downtime is expected to last between 5 to 15 minutes depending on your instance size. Alerts and scheduled reports may be impacted… | 4h 0m | Scheduled |
| Aug 24, 2026 | Scheduled Database Maintenance – Temporary Grafana Instance UnavailabilityYour Grafana instance will be temporarily unavailable due to scheduled maintenance. Downtime is expected to last between 5 to 15 minutes depending on your instance size. Alerts and scheduled reports may be impacted… | 4h 0m | Scheduled |
| Aug 24, 2026 | API MaintenanceScheduled maintenance is currently in progress. We will provide updates as necessary. | 5h 0m | Scheduled |
| Aug 20, 2026 | Scheduled Database Maintenance – Temporary Grafana Instance UnavailabilityScheduled maintenance is currently in progress. We will provide updates as necessary. | 4h 0m | Scheduled |
| Aug 20, 2026 | API MaintenanceScheduled maintenance is currently in progress. We will provide updates as necessary. | 5h 0m | Scheduled |
| Aug 19, 2026 | Scheduled Database Maintenance – Temporary Grafana Instance UnavailabilityScheduled maintenance is currently in progress. We will provide updates as necessary. | 4h 0m | Scheduled |
| Aug 19, 2026 | API MaintenanceScheduled maintenance is currently in progress. We will provide updates as necessary. | 5h 0m | Scheduled |
| Aug 18, 2026 | Cloud Logs read path outage on eu-west-2We observed an issue which impacted the following service: reads of Cloud Logs in the eu-west-2 region. Affected services would include Explore, dashboards, alert evaluation when querying Logs. The time of impact… | -53 min | Resolved |
| Aug 18, 2026 | Scheduled Database Maintenance – Temporary Grafana Instance UnavailabilityScheduled maintenance is currently in progress. We will provide updates as necessary. | 4h 0m | Scheduled |
| Aug 18, 2026 | API MaintenanceScheduled maintenance is currently in progress. We will provide updates as necessary. | 5h 0m | Scheduled |
| Aug 17, 2026 | Scheduled Database Maintenance – Temporary Grafana Instance UnavailabilityScheduled maintenance is currently in progress. We will provide updates as necessary. | 4h 0m | Scheduled |
| Aug 14, 2026 | Scheduled Database Maintenance – Temporary Grafana Instance UnavailabilityScheduled maintenance is currently in progress. We will provide updates as necessary. | 4h 0m | Scheduled |
| Aug 13, 2026 | K6 Test OutageThis incident has been resolved. | 1h 22m | Resolved |
| Aug 13, 2026 | Cloudflare MaintenanceMaintenance will begin as scheduled in 60 minutes. | 1h 0m | Scheduled |
| Aug 13, 2026 | Scheduled Database Maintenance – Temporary Grafana Instance UnavailabilityScheduled maintenance is currently in progress. We will provide updates as necessary. | 4h 0m | Scheduled |
| Aug 12, 2026 | Scheduled Database Maintenance – Temporary Grafana Instance UnavailabilityScheduled maintenance is currently in progress. We will provide updates as necessary. | 4h 0m | Scheduled |
| Aug 11, 2026 | Scheduled Database Maintenance – Temporary Grafana Instance UnavailabilityScheduled maintenance is currently in progress. We will provide updates as necessary. | 4h 0m | Scheduled |
| Aug 11, 2026 | Scheduled Database Maintenance – Temporary Grafana Instance UnavailabilityScheduled maintenance is currently in progress. We will provide updates as necessary. | 4h 0m | Scheduled |
| Aug 10, 2026 | Scheduled Database Maintenance – Temporary Grafana Instance UnavailabilityScheduled maintenance is currently in progress. We will provide updates as necessary. | 4h 0m | Scheduled |
| Aug 10, 2026 | Metrics: Elevated Error Rates Reads/WritesBetween 12:20 and 13:15 UTC, we experienced elevated error rates for metrics reads and writes due to a networking issue. Impact was limited to a subset of deployments in the prod-us-central-0 region. A fix has been… | 0 min | Resolved |
| Aug 8, 2026 | Alerting expressions pipeline failing when recovery settingsDue to a software bug the evaluation of some alert rules (primarily ones that have recovery threshold setting) were failing to be evaluated starting 14:00 UTC to 19:30 UTC today. | 0 min | Resolved |
| Aug 7, 2026 | Scheduled Database Maintenance – Temporary Grafana Instance UnavailabilityThe scheduled maintenance has been completed. | 4h 0m | Scheduled |
| Aug 6, 2026 | Scheduled Database Maintenance – Temporary Grafana Instance UnavailabilityThe scheduled maintenance has been completed. | 4h 0m | Scheduled |
| Aug 6, 2026 | Scheduled Database Maintenance – Temporary Grafana Instance UnavailabilityThe scheduled maintenance has been completed. | 4h 0m | Scheduled |
| Aug 5, 2026 | Some Cloud Test Runs TerminatedBetween approximately 13:40 and 15:05 UTC today, a subset of cloud test runs were unexpectedly terminated and marked Aborted (by system). This affected runs that were in progress during three short windows in that… | 0 min | Resolved |
| Aug 5, 2026 | Scheduled Database Maintenance – Temporary Grafana Instance UnavailabilityThe scheduled maintenance has been completed. | 4h 0m | Scheduled |
| Aug 5, 2026 | o11y requests too large for nods in the prod-eu-west-2 region.This incident has been resolved. | 3 min | Resolved |
| Aug 4, 2026 | Some Grafana Instances UnavailableWe are observing a continued period of stability and at this point, we are marking the incident resolved. | 6h 20m | Resolved |
| Aug 4, 2026 | Logs latency increase within prod-eu-west-3This incident has been resolved. | 3h 14m | Resolved |
| Aug 1, 2026 | K6 - Cloud test-run issuesThis incident has been resolved. | 1h 5m | Resolved |
| Jul 31, 2026 | Partial Read Outage for Loki in prod-us-east-4We are investigating a partial read outage affecting Loki in prod-us-east-4. Between 19:41 UTC and 19:50 UTC, a significant portion of read queries may have failed or returned errors. The issue has been identified and… | -63 min | Resolved |
| Jul 31, 2026 | Degraded Performance: Stack Provisioning Failures within certain reigons (PDC Setup)He have identified the cause and applied a fix for this issue and all effected stacks within the affected regions are not working as expected. | 1h 22m | Resolved |
| Jul 30, 2026 | PDC Authentication IssuesThis incident has been resolved. | 2h 54m | Resolved |
| Jul 30, 2026 | Issues with Billing/Usage Dashboard Metrics and Panels.This incident has been resolved. | 2h 1m | Resolved |
| Jul 29, 2026 | IRM Performance Degradation in EU RegionThe issue affecting Grafana IRM in the EU region has been resolved. The IRM UI, public API, and alert notification processing have been restored and are operating normally. | 2h 52m | Resolved |
| Jul 29, 2026 | Grafana Cloud non-billing usage metrics gaps in select regionsImpact period: April 1 – July 29, 2026 Summary: During this period, some customers in a subset of regions may have experienced gaps in select ruler/recording rule metrics within their Grafana Cloud instances. Not all… | 0 min | Resolved |
| Jul 28, 2026 | Partial OTLP Write Outage in prod-us-east-3The issue affecting OTLP ingestion in the prod-us-east-3 region has been resolved. Between 13:30 UTC and 17:45 UTC, some customers experienced intermittent failures when writing telemetry to the OTLP endpoint. Logs were… | 3 min | Resolved |
| Jul 24, 2026 | Write OutageThis incident has been resolved. | 1h 18m | Resolved |
| Jul 24, 2026 | Errors creating new Slack integration for Grafana IRMThis incident has been resolved. | 4h 45m | Resolved |
| Jul 23, 2026 | Intermittent write outage from 21:20-21:26, 21:54-21:56, and 22:22-22:23Writes are looking stable in the last 6h | 6h 31m | Resolved |
| Jul 23, 2026 | Metrics Write Path ErrorsThis incident has been resolved. | 2h 33m | Resolved |
| Jul 23, 2026 | Stacks using SCIM user provisioning are currently unable to log into GrafanaThis incident has been resolved. | 5h 24m | Resolved |
| Jul 22, 2026 | K6 - Cloud output test-runs are failing to fetch the script logsThis incident has been resolved. | 57 min | Resolved |
| Jul 21, 2026 | Cloud Log Exporter UnavailableThis incident has been resolved. | 2h 22m | Resolved |
| Jul 20, 2026 | PDC IssuesThis incident has been resolved. | 1h 15m | Resolved |
| Jul 17, 2026 | Grafana Cloud login issues for users with role of NoneUsers with the None role authenticating via Grafana.com should be able to access their stacks again | 18h 52m | Resolved |
| Jul 16, 2026 | Adaptive Metrics aggregation delay in eu-west-0 region.The incident is now fully resolved. | 2h 2m | Resolved |
| Jul 15, 2026 | Partial Outage in prod-eu-west-2This incident has been resolved. Thank you for your patience. | 1h 12m | Resolved |
| Jul 15, 2026 | Delayed Aggregated Metrics (prod-us-central-0)This incident has been resolved. | 8h 35m | Resolved |
| Jul 15, 2026 | Some Reports of Grafana Not Loading.This incident has been resolved. Thank you for your patience. | 2h 53m | Resolved |
| Jul 15, 2026 | Fleet Managment Interface 404'sBetween 9:45 and 12:00 UTC, we experienced an issue affecting the Fleet Management interface. During this time, the Remote Configuration tab for production stacks returned a 404 error when accessed through the… | 0 min | Resolved |
| Jul 14, 2026 | Mimir Write Performance DegradationThis incident has been resolved. Thank you for your patience. | 1h 32m | Resolved |
| Jul 14, 2026 | Mimir Partial Write OutageThis incident has been resolved. | 1h 53m | Resolved |
| Jul 14, 2026 | Issue with Dashboard Views Being RegisteredA fix has been implemented and dashboard view and error counts in the Dashboards and Folder list are updating as expected. Thank you for your patience while we worked to address this issue. | 146h 5m | Resolved |
| Jul 10, 2026 | Network Degredation in prod-us-central-0From approximately 20:24 UTC - 20:47 UTC a network issue in prod-us-central-0 isolated part of our infrastructure in one availability zone, temporarily cutting off connectivity to a subset of backend services. This… | 0 min | Resolved |
| Jul 10, 2026 | IRM Mobile App forcing some users to logoutThis incident has been resolved. | 64h 33m | Resolved |
| Jul 10, 2026 | Issue with Dashboard Views Being RegisteredAt this stage, we are considering the incident resolved. | 61h 25m | Resolved |
| Jul 9, 2026 | PDC Degraded performanceThis incident has been resolved. Thank you for your patience. | 8h 17m | Resolved |
| Jul 9, 2026 | Delayed ingestion and recording rule evaluation failures for Mimir in prod-ap-south-1This incident has been resolved. | 2h 39m | Resolved |
| Jul 9, 2026 | Loki read path in prod-eu-west-2 was downBetween 5:49 and 5:54 UTC, the read path in prod-eu-west-2 was down. This has completely recovered by 5:57 UTC. Recording rules may have failed to evaluate during this period which may result in gaps. | 0 min | Resolved |
| Jul 8, 2026 | Loki Billing Data LostBetween 13:57 and 14:20 UTC (23 minutes), a data gap occurred affecting billing usage data for Loki. Data for this window was not recorded and cannot be recovered. | 0 min | Resolved |
| Jul 8, 2026 | Grafana rulers crash-looping on prometheusThis incident has been resolved. | 5h 6m | Resolved |
| Jul 8, 2026 | Cannot access dashboard set up as a home pageWe have found that the issues are related to the grafana.unifiedHomepage feature rollout active 16:50 UTC yesterday, to 10:20 UTC today. We have rolled this back and systems are now working as expected. | 32 min | Resolved |
| Jul 7, 2026 | Delayed usage and billing dataWe identified an issue with the internal job that calculates month-to-date usage and cost data, which caused usage attribution and billing dashboards to display stale information. The root cause has been identified and… | 0 min | Resolved |
| Jul 5, 2026 | Grafana Cloud IRM alert groups failing to produce alerts in us-east-3 region.We haven't noticed any further issues in this region for alert group processing since yesterday. This incident i fully resolved. | 19h 59m | Resolved |
| Jul 3, 2026 | Partial Write Outage for Grafana Cloud Logs in prod-eu-north-0.Grafana Cloud Logs in prod-eu-north-0 experienced a 10-minute partial write outage between 13:45 and 13:54 UTC. Impacted users may have experienced 5xx errors during this time. | 0 min | Resolved |
| Jul 2, 2026 | Some Queries FailingThis incident has been resolved. Thank you for your patience. | 3h 25m | Resolved |
| Jul 2, 2026 | Loki and Frontend Observability - Major Outage in prod-us-central-0 regionThis incident has been resolved by restarting the affected services. | 1h 53m | Resolved |
| Jul 1, 2026 | Elevated Loki Query Bytes ReportingWe’ve implemented a fix and can confirm the issue is fully resolved as of 20:25 UTC. Thank you for your patience. | 2h 17m | Resolved |
| Jun 29, 2026 | Confluent API OutageThis incident has been resolved. | 1h 49m | Resolved |
| Jun 29, 2026 | Mimir read errors and high latency in prod-eu-west-0Since the mitigation has been applied, we have not seen the errors return. At this point, we are considering the incident resolved. | 20h 43m | Resolved |
| Jun 29, 2026 | Rule evaluation error on cluster prod-gb-south-0This incident has been resolved. | 2h 38m | Resolved |
| Jun 26, 2026 | K6 - Test run metrics processing is delayedThis incident has been resolved. | 37h 27m | Resolved |
| Jun 24, 2026 | Scheduled maintenance: Grafana Cloud Migration AssistantThe scheduled maintenance has been completed. | 8h 0m | Scheduled |
| Jun 22, 2026 | Elevated Logs Query Usage — Unexpected Billing ImpactStarting June 20, 2026 at approximately 20:00 UTC, some Grafana Cloud customers experienced unexpectedly elevated logs query usage. The issue persisted until it was resolved on June 22, 2026 at approximately 16:00… | -3010 min | Resolved |
| Jun 19, 2026 | Rule Evaluation Outage in prod-us-central-0This incident has been resolved. Thank you for your patience. | 2h 40m | Resolved |
| Jun 18, 2026 | Issues with actions in the Grafana IRM mobile appThis incident has been resolved. Thank you for your patience. | 1h 48m | Resolved |
| Jun 18, 2026 | Potential Issues Loading Grafana for Users in IndiaThis incident has been resolved. Error rates continued to remain near 0 and operations are performing as expected. | 141h 28m | Resolved |
| Jun 18, 2026 | Degraded k6 cloud UI performanceThis incident has been resolved. Thank you for your patience. | 4h 56m | Resolved |
| Jun 18, 2026 | Frontend Observability - Suspected commit feature not working as expectedA fix has been deployed and the issue after monitoring as been fixed. | 1h 24m | Resolved |
| Jun 17, 2026 | Loki data source-managed alert rules not visible in the Grafana Cloud Alerting UIThis incident has been resolved. | 17h 51m | Resolved |
| Jun 12, 2026 | Brief Loki Prod-012-eu-west-2 DisruptionOur team had discovered a read issue around 19:35-20:08 UTC. Impact at the time would have provided errors similar to context deadline exceeded (DatasourceError response). This has since been resolved, and should not… | -204 min | Resolved |
| Jun 10, 2026 | Grafana Dashboards page not displaying when set to ‘View by Folders’This incident has been resolved. | 18h 46m | Resolved |
| Jun 9, 2026 | Investigating Issues with Data Source-Managed AlertingWe continue to observe a continued period of recovery. At this time, we are considering this issue resolved. No further updates. | 14h 0m | Resolved |
| Jun 8, 2026 | IRM Degraded PerformanceThis incident has been resolved. | 9h 44m | Resolved |
| Jun 7, 2026 | Brief Rule Evaluation Failures in prod-eu-west-3This incident has been resolved. Thank you for your patience. | 138h 12m | Resolved |
| Jun 5, 2026 | Permissions Issues with IRMThis incident has been resolved. | 5h 24m | Resolved |
| Jun 5, 2026 | Silences not Working as ExpectedThis incident has been resolved. | 58 min | Resolved |
| Jun 4, 2026 | Grafana Assistant Skills Page BlankThis incident has been resolved. | 1h 44m | Resolved |
| Jun 3, 2026 | K6 Test Runs Degraded PerformanceThis incident has been resolved. | 9h 7m | Resolved |
| Jun 3, 2026 | Synthetic Scripted/Browser checks failureThis incident has been resolved. | 6h 44m | Resolved |
| Jun 2, 2026 | tempo prod-25 write-path-downThis incident has been resolved. | 8h 31m | Resolved |
| Jun 1, 2026 | Alert manager unavailable in prod-us-central-0This incident has been resolved. | 32 min | Resolved |
| May 29, 2026 | Grafana Loki Log Query IssuesThis incident has been resolved. | 1h 56m | Resolved |
| May 28, 2026 | Status Page ImprovementsThe scheduled maintenance has been completed. | 2h 0m | Scheduled |
| May 27, 2026 | Prometheus Datasource Errors/Outage in prod-us-east-0This incident has been resolved. Thank you for your patience. | 2h 36m | Resolved |
| May 18, 2026 | Grafana K6 metrics processing and test runs degradationThis incident has been resolved. | 7h 18m | Resolved |
| May 13, 2026 | Intermittent Errors and High latency Writing to Cloud Metrics, Cloud Logs and Cloud TracesWe continue to observe an extended period of recovery and we're marking the incident as resolved at this point in time. | 22h 16m | Resolved |
| May 11, 2026 | "Failed to Load Dashboard" ErrorsThis incident has been resolved. Thank you for your patience. | 18h 56m | Resolved |
| May 11, 2026 | SSL/TLS Connectivity IssuesThis incident has been resolved. Thank you for your patience. | 1h 51m | Resolved |
| May 8, 2026 | Cloud Metrics -High Write Latency and Errors in prod-us-central-7We have continued to observe stability. This incident is now being considered as resolved. Thank you for your patience. | 1h 13m | Resolved |
| May 8, 2026 | Hardware failure on CSP within prod-us-west-0We observed an underlying hardware failure on our CSP which triggered an automatic live VM migration. The situation caused a degradation in write performance for Grafana Cloud Metrics on prod-us-west-0 between 05:26 UTC… | -3214 min | Resolved |
| May 7, 2026 | Metrics read errors in prod-ap-south-1 regionAt this time, we have confirmed that the query errors have gone and we are considering this issue resolved. | 38 min | Resolved |
| May 6, 2026 | Datasource Query Performance IssuesWe’re currently investigating an issue affecting Datasource query performance in prod-us-east-4. Our team is actively working to identify the cause. Thank you for your patience. | 2135h 51m | Resolved |
| May 5, 2026 | Elevated Error Rate of Browser Checks in PoP OregonThis incident has been resolved. Thank you for your patience. | 4h 2m | Resolved |
| May 4, 2026 | k6 Partial OutageThis incident has been resolved. Thank you for your patience. | 3h 11m | Resolved |
| May 1, 2026 | Ingestion Errors for AWS Cloud Provider Observability Metric Streams in prod-us-central-7This incident has been resolved. | 1h 13m | Resolved |
| Apr 29, 2026 | Performance Testing – Degraded Service (Resolved)We experienced degraded performance affecting Performance Testing from 13:10 UTC to 13:20 UTC. During this time, users may not have been able to start new test runs. The issue has been resolved, and the service is now… | -111 min | Resolved |
| Apr 29, 2026 | Elevated write latency for AWS Metrics Streaming integration in us-east-3 region.We were facing an incident with AWS Metrics Streaming integration in us-east-3 region manifesting in elevated ingestion latency. The incident started at around 10:45 UTC and was resolved at around 12:30 UTC. Some… | -147 min | Resolved |
| Apr 28, 2026 | Investigating Issues Saving SQL Datasource CredentialsThis incident has been resolved. Thank you for your patience. | 18h 51m | Resolved |
| Apr 28, 2026 | Gateway Slowness Detected in Prod (US-East-1)After further review, this was a false alarm and should not have affected any users. This incident has been resolved. Thank you for your patience. | 53h 51m | Resolved |
| Apr 27, 2026 | InfluxDB Datasource - Intermittent FailuresThis incident has been resolved. Thank you for your patience. | 6h 16m | Resolved |
| Apr 23, 2026 | Cloudwatch Datasource OutageThis incident has been resolved. Thank you for your patience. | 5h 35m | Resolved |
| Apr 20, 2026 | Restrictions on Alerts & Reports for Grafana Cloud Free/Trial UsersGrafana Labs has taken steps to safeguard the Grafana Cloud platform against the distribution of unauthorized emails. We have implemented the following changes to new Grafana Cloud Free and Trial accounts, effective… | 89h 51m | Resolved |
| Apr 20, 2026 | Elevated 429 Errors Impacting Metrics Querying Across Multiple RegionsThis incident has been resolved. Thank you for your patience. | 21 min | Resolved |
| Apr 17, 2026 | Query Caching - Degraded PerformanceThis incident has been resolved | 1h 34m | Resolved |
| Apr 16, 2026 | Issues on Stack creationThis incident has been resolved. | 1h 10m | Resolved |
| Apr 15, 2026 | Degraded Ticket Visibility in Support SystemThis incident has been resolved and our ticketing system is fully operational. Thank you for your patience. | 17 min | Resolved |
| Apr 14, 2026 | k6 Cloud Service DisruptionBetween approximately 12:30 UTC and 13:15 UTC, k6 Cloud experienced a service disruption due to issues introduced in a recent API release. During this time, users were unable to access the k6 Cloud application. The… | -134 min | Resolved |
| Apr 14, 2026 | Loki write instability in prod-eu-west-2.loki-prod-012There was a period of write instability yesterday. It was between ~1330 -1730 UTC yesterday. This was due to a scheduled maintenance. | -1472 min | Resolved |
| Apr 14, 2026 | K6 Sporadic DNS IssuesThis incident is now resolved. We had intermediary issues with a flaky DNS server that caused random tests to not start properly. Since the DNS server was fixed, we haven't been seeing the issue anymore. | 27h 36m | Resolved |
| Apr 10, 2026 | Grafana Cloud Logs - Write degradation in us-east-3This incident has been resolved. | 43 min | Resolved |
| Apr 10, 2026 | Tempo Write OutageThis incident has been resolved. Thank you for your patience. | 1h 19m | Resolved |
| Apr 9, 2026 | K6 Browser Testing/Timeline Not AvailableThis incident has been resolved. Thank you for your patience. | 1h 15m | Resolved |
| Apr 8, 2026 | Stability Issues for Some Customers in the prod-gb-south-1 Region.We had a stability issue for a subset of customers in the prod-gb-south-1 region. The impact was between UTC 15:20-16:30 which impacted roughly 30% of queries and rules evaluations. We've applied mitigations, queries… | 0 min | Resolved |
| Apr 7, 2026 | Unable to Edit Notification PoliciesThis incident has been resolved. Thank you for your patience. | 4h 59m | Resolved |
| Apr 6, 2026 | Notification Policies and Contact Points Missing in UI on the Slow Release ChannelThis incident has been resolved. | 21h 38m | Resolved |
| Apr 3, 2026 | Partial K6 Test Run OutageThis incident has been resolved. Thank you for your patience. | 2h 8m | Resolved |
| Apr 1, 2026 | AWS integration Degraded PerformanceThis incident has been resolved. Thank you for your patience. | 46 min | Resolved |
| Apr 1, 2026 | Query degradation and possible rule evaluation failure on prod-eu-west-0.cortex-prod-01This incident has been resolved. | 11h 16m | Resolved |
| Mar 31, 2026 | k6 Cloud DegradationFrom approximately 11:00 UTC - 15:00 UTC we had a degradation that caused test start errors for a large percentage of Cloud runs managed as scripts in the GCK6 app. This has since been resolved. | 0 min | Resolved |
| Mar 31, 2026 | Synthetic Monitoring: Some Check Creations & Updates Might be Blocked.This is a retroactive status page linked to the following incident: https://status.grafana.com/incidents/38wwbz50ggrp This retroactive status page is meant to clarify the time of impact. This issue first started at… | 0 min | Resolved |
| Mar 31, 2026 | Synthetic Monitoring: Some Check Creations & Updates Might be Blocked.This incident has been resolved. | 24 min | Resolved |
| Mar 31, 2026 | Some of the CloudWatch queries are failingThis incident has been resolved. | 35 min | Resolved |
| Mar 30, 2026 | Tempo Reads Outage for Small Subset of CustomersWe encountered an issue impacting only a small subset of customers in the prod-us-central-0 region. The incident occurred between 16:20 and 17:50 UTC on 3/30/26. This incident is now resolved. | -124 min | Resolved |
| Mar 27, 2026 | Some Grafana Instances UnavailableThis incident has been resolved. Thank you for your patience. | 7h 12m | Resolved |
| Mar 25, 2026 | Prometheus writes in prod-eu-west-3 are degradedThis incident has been resolved. Thank you for your patience. | 701h 56m | Resolved |
| Mar 24, 2026 | Service degradation on Dashboard loading in several clusters.An issue affecting Grafana Cloud instances was diagnosed yesterday 24th of March that avoided Dashboards to be loaded correctly. The incident impacted the following clusters: - GCP US Central (us-central-0) between… | 0 min | Resolved |
| Mar 24, 2026 | Prometheus writes, Logs, and Synthetic Monitoring in prod-eu-west-3 are degradedThis incident has been resolved. | 27h 43m | Resolved |
| Mar 23, 2026 | Grafana Assistant Unavailable in prod-us-east-0This incident has been resolved. | 1h 44m | Resolved |
| Mar 20, 2026 | Authentication API Database Down in prod-eu-west-2 and prod-eu-west-4This incident has been resolved. | 40 min | Resolved |
| Mar 19, 2026 | Various Datasource IssuesThis incident has been resolved. | 1h 57m | Resolved |
| Mar 19, 2026 | Degraded performance of Grafana Cloud k6 test runsOur engineering team has deployed a fix and we continue to observe a continued period of recovery. At this time, we are considering this issue resolved. No further updates. | 6h 54m | Resolved |
| Mar 13, 2026 | Grafana Cloud Logs - Write degradation in Azure Netherlands (eu-west-3)We have been observing stability for a period of time and will mark the incident as resolved at this time. | 116h 45m | Resolved |
| Mar 13, 2026 | Increased number of Aborted-by-Systems with a k6 binary building errorsThis incident has been resolved. | 10h 30m | Resolved |
| Mar 11, 2026 | Rule Evaluation Outage in prod-us-west-0This incident has been resolved. | 49h 5m | Resolved |
| Mar 11, 2026 | Grafana Cloud Logs - Write degradation in Azure Netherlands (eu-west-3)This incident has been resolved. | 28h 47m | Resolved |
| Mar 10, 2026 | Various Issues with HG PagesThis incident has been resolved. | 1h 10m | Resolved |
| Mar 10, 2026 | Some Write Failures in prod-eu-west-3.This incident has been resolved. | 27h 47m | Resolved |
| Mar 10, 2026 | Service degradation on Logs Read path in AWS US West (us-west-0)This incident has been resolved. | 5h 12m | Resolved |
| Mar 9, 2026 | Metrics write path outage in prod-us-central-0 and prod-us-central-5This incident has been resolved. | 27h 13m | Resolved |
| Mar 9, 2026 | Fleet Managment Elevated Rate of ErrorsThis incident has been resolved. | 30h 33m | Resolved |
| Mar 8, 2026 | Service degradation on Logs Read path in AWS US West (us-west-0)We continue to observe a continued period of stability since 19:40 UTC. At this time, we are considering this issue resolved | 6h 13m | Resolved |
| Mar 7, 2026 | Outage for prod-eu-central-0 due to AWS S3 outage.This incident has been resolved. | 36h 52m | Resolved |
| Mar 6, 2026 | Some Grafana Instances UnavailableThis incident has been resolved. | 1h 27m | Resolved |
| Mar 5, 2026 | Write failures in prod-eu-west-0We continue to observe a continued period of recovery. At this time, we are considering this issue resolved. No further updates. | 1h 8m | Resolved |
| Mar 4, 2026 | Elevated rate of errors for Fleet Management in prod-us-central-0This incident has been resolved. | 1h 41m | Resolved |
| Mar 3, 2026 | Test Run Browser Screenshot Upload FailingTest run browser screenshot upload experienced failures from 13:12 to 14:51 UTC. The issue has been resolved | -335 min | Resolved |
| Mar 3, 2026 | Grafana Cloud Logs - Write degradation in Azure Netherlands (eu-west-3)This incident has been resolved. | 54h 23m | Resolved |
| Mar 2, 2026 | Write outage for logs in prod-eu-west-3This incident has been resolved. | 8h 11m | Resolved |
| Mar 2, 2026 | Complete outage in prod-me-central-1Following our ongoing communications regarding the complete outage in prod-me-central-1, we are now closing this incident. As noted in the latest AWS update, the Middle East (UAE) region (ME-CENTRAL-1) has suffered… | 2571h 14m | Resolved |
| Feb 27, 2026 | Increased Latency for Small Subset of CustomersA recent rollout caused the AuthZ (RBAC) service to perform many redundant folder-tree fetches for each authorization check. For a small number of tenants in the prod-us-east-0 and prod-eu-west-2 regions with very large… | 0 min | Resolved |
| Feb 27, 2026 | Trace querying issue in all Tempo clustersThis incident has been resolved. | 9h 51m | Resolved |
| Feb 27, 2026 | Incorrect pipeline assignment after custom attributes are assignedThis incident has been resolved. | 2h 26m | Resolved |
| Feb 26, 2026 | Grafana Cloud Faro slowness of listing and uploading sourcemaps in all regions.This incident has been resolved. | 13h 48m | Resolved |
| Feb 25, 2026 | Grafana Cloud Metrics - Intermittent Write Latency in prod-us-central, prod-us-central-5, and prod-eu-west-0This incident is now resolved. During the incident the Cloud Metrics platform experienced intermittent latency spikes communicating with a backend cloud service in the prod-us-central-0 and prod-us-central-5 regions… | 478h 28m | Resolved |
| Feb 25, 2026 | Issues Loading Dashboards and Alert Folders in Hosted GrafanaThis incident has been resolved. | 2h 6m | Resolved |
| Feb 25, 2026 | Partial Write & Rule Evaluation Outage in prod-eu-west-3This incident has been resolved. | 2h 15m | Resolved |
| Feb 25, 2026 | Grafana Cloud Traces prod-eu-west-6 region (AWS Ireland) wrong URL endpoint shown for traces ingestion.This incident has been resolved. | 2h 24m | Resolved |
| Feb 24, 2026 | Some Alert Rule Evaluations FailingThis incident has been resolved. | 2h 38m | Resolved |
| Feb 18, 2026 | Brief Disruption in Azure prod-us-7-centralWe experienced an issue impacting a cell within the Azure prod-us-central-7 region, which occurred between 14:26 and 14:36. Affected users may have noticed increased errors with rule evaluations, as well as a some… | -56 min | Resolved |
| Feb 18, 2026 | Degraded performance of Grafana Cloud k6 test runsThis incident has been resolved. | 12h 49m | Resolved |
| Feb 18, 2026 | Grafana Cloud metrics degredationThis incident has been resolved. | 1h 47m | Resolved |
| Feb 17, 2026 | Maintenance task for Synthetic Monitoring ProbeFailedExecutionsTooHigh alert ruleThis incident has been resolved. | 1h 34m | Resolved |
| Feb 17, 2026 | Degradation of service on Synthetic Monitoring Public Probe AWS Canada (Calgary)There was a service degradation today from ~12:09 UTC until ~12:35 UTC on the Public Probe of Calgary for Synthetic Monitoring. Impact may include SM check fails where the probe was used. | -3 min | Resolved |
| Feb 13, 2026 | Self-Serve Users Unable to Sign UpThis incident has been resolved. | 28 min | Resolved |
| Feb 13, 2026 | Loki writes outage in prod-ca-east-0We continue to observe a continued period of recovery. At this time, we are considering this issue resolved. | 29 min | Resolved |
| Feb 12, 2026 | Loki Delete Endpoint BugThis incident has been resolved. | 17h 46m | Resolved |
| Feb 12, 2026 | Essential Maintenance for Faro ServicesThis incident has been resolved. | 2h 58m | Resolved |
| Feb 12, 2026 | Grafana Cloud Metrics elevated write and rule evaluation latency in prod-eu-west-2 region.We no longer observed any problems with our services - this incident has been resolved. | 1h 57m | Resolved |
| Feb 11, 2026 | Unable to Install Slack IntegrationThis incident has been resolved. | 7h 25m | Resolved |
| Feb 11, 2026 | Loki error response rate spike on prod-ap-southeast-1This incident has been resolved. | 34 min | Resolved |
| Feb 10, 2026 | Write failures in prod-us-central-0We continue to observe a continued period of recovery. At this time, we are considering this issue resolved. | 1h 5m | Resolved |
| Feb 9, 2026 | Athena Queries BrokenThis incident has been resolved. | 3h 32m | Resolved |
| Feb 9, 2026 | Grafana Cloud Logs – Write Ingestion DegradationThis incident has been resolved. | 48 min | Resolved |
How to Read This Grafana Cloud Incident Log
Each row is an incident PulsAPI observed, not a summary written afterwards. The duration is wall-clock time between the first failing check and the first clean one, so it includes the window before Grafana Cloud acknowledged anything. Vendor post-mortems typically measure from acknowledgement, which is why their numbers are usually shorter.
Incidents still open have no duration yet and are listed as ongoing rather than being given a running total. The archive covers the last 365 days; anything older has aged out of the window rather than never having happened.
The longest single Grafana Cloud outage in this window ran 107d 3h, against a mean recovery of 1d 18h. If you depend on Grafana Cloud in a customer-facing path, the longest figure is the one to design around. The mean is what happens on a normal bad day; the maximum is what happens on the worst one.