Groq Outage History
27 incidents recorded over the last 365 days, 26 of them resolved. Durations are measured from when PulsAPI first saw the problem to when it cleared, which is usually longer than the vendor's own figure.
Every Recorded Groq Incident
| Date | Incident | Duration | Status |
|---|---|---|---|
| Jul 1, 2026 | Data Center Failure Impacting CapacityThe **capacity issue has been resolved**. All services are operating at normal capacity. The incident was caused by a **power loss issue that led to a cooling system failure at one of our US Central data centers**. We… | 1h 33m | Resolved |
| Mar 19, 2026 | openai/gpt-oss-120b Performance IssueThe issues affecting openai/gpt-oss-120b have been **resolved**. The model is operating normally. Actions were taken to cancel billing plans and restrict verification status of organizations engaged in coordinated… | 59 min | Resolved |
| Feb 7, 2026 | meta-llama/llama-4-scout-17b-16e-instruct Degraded PerformanceThis incident has been resolved. Thank you for your patience. | 1h 32m | Resolved |
| Feb 5, 2026 | meta-llama/llama-4-scout-17b-16e-instruct Degraded PerformanceThe issues affecting meta-llama/llama-4-scout-17b-16e-instruct have been **resolved**. The model is once again operating normally. We apologize for the disruption and thank you for your patience. | 23 min | Resolved |
| Feb 5, 2026 | meta-llama/llama-4-scout-17b-16e-instruct Degraded PerformanceThe issues affecting meta-llama/llama-4-scout-17b-16e-instruct have been fully **resolved**. The model is operating normally. We apologize for the disruption and thank you for your patience. | 1h 44m | Resolved |
| Jan 26, 2026 | meta-llama/llama-4-scout-17b-16e-instruct Degraded PerformanceThis incident has been resolved. Thank you for your patience. | 1h 53m | Resolved |
| Jan 24, 2026 | Data Center Failure Impacting Model Latency - SYDThis incident is now resolved. | 2h 30m | Resolved |
| Dec 24, 2025 | Model Performance or Availability Issue: llama-3.3-70b-versatileThe issues affecting **llama-3.3-70b-versatile** have been **resolved**. The model is operating normally. We apologize for the disruption and thank you for your patience. | 1h 15m | Resolved |
| Dec 17, 2025 | Data Center Failure Impacting Capacity - DMM1**The data center capacity issue has been resolved. All services are operating at normal capacity. The incident was caused by a failed ARP entry on the network gateway device caused the Salam network link to go down in… | -3941 min | Resolved |
| Dec 6, 2025 | meta-llama/llama-4-scout-17b-16e-instruct Degraded PerformanceThe issues affecting meta-llama/llama-4-scout-17b-16e-instruct , meta-llama/llama-4-maverick-17b-128e-instruct, & groq/compound have been **resolved**. The models are operating normally. We apologize for the disruption… | 1h 25m | Resolved |
| Dec 2, 2025 | meta-llama/llama-4-scout-17b-16e-instruct & groq/compound Degraded PerformanceThe issues affecting meta-llama/llama-4-scout-17b-16e-instruct & groq/compound have been **resolved**. Both of the models are operating normally. We apologize for the disruption and thank you for your patience. | 2h 25m | Resolved |
| Nov 18, 2025 | Console, Website, & API Availability Issues: Global Cloudflare OutageWe are resolving this incident as we have now observed ~30 minutes of normalized service availability and customer traffic. We will continue to closely monitor our systems for any signs of regression as [Cloudflare's… | 3h 33m | Resolved |
| Nov 5, 2025 | Cloud API DegradationThe issues causing degraded performance have been **resolved**. All models are now operating normally. Upon investigation we learned the issue was scoped to an internal feature yet to be released that is still in… | 34 min | Resolved |
| Nov 4, 2025 | Model Performance Issue: llama-3.1-8b-instantThe issues affecting llama-3.1-8b-instant have been **resolved**. The model is operating normally. Root cause: Two recent changes to the inference-engine-instances repository that were contributing to the elevated Orion… | 36 min | Resolved |
| Oct 24, 2025 | Planned Maintenance: meta-llama/llama-4-maverick-17b-128e-instruct (me-central-1) **When:** Monday, Nov 3, 2025 • 00:00 to 03:00 UTC (Sunday, Nov 2, 16:00 to 19:00 PT) **What** • We’re performing planned maintenance on the me-central-1 deployment of… | — | Scheduled |
| Oct 22, 2025 | openai/gpt-oss-120b Degraded PerformanceThe latency issues affecting openai/gpt-oss-120b have been resolved. The model is now operating normally and latency has returned to expected ranges. We apologize for the disruption and thank you for your patience. | 2h 8m | Resolved |
| Oct 21, 2025 | openai/gpt-oss-120b Degraded PerformanceThe issues affecting openai/gpt-oss-120b have been **resolved**. The model is operating normally. We apologize for the disruption and thank you for your patience. | 1h 53m | Resolved |
| Oct 14, 2025 | meta-llama/llama-4-scout-17b-16e-instruct & groq/compound Degraded PerformanceThe issues affecting meta-llama/llama-4-scout-17b-16e-instruct have been **resolved**. The model is operating normally. We apologize for the disruption and thank you for your patience. | 2h 2m | Resolved |
| Oct 6, 2025 | Llama-3.1-8b-instant model Degraded PerformanceThe issues affecting llama-3.1-8b-instant have been resolved. The model is operating normally. Thank you for your patience. | 49 min | Resolved |
| Oct 6, 2025 | llama-3.3-70b-versatile and llama-3.1-8b-instant Degraded PerformanceThe issues affecting llama-3.3-70b-versatile and llama-3.1-8b-instant have been **resolved**. The models are operating normally. We apologize for the disruption and thank you for your patience. | 48 min | Resolved |
| Sep 18, 2025 | Degraded login via VercelVercel‑initiated login is operating normally. Other login methods were unaffected. This incident is now resolved. | 16h 7m | Resolved |
| Sep 10, 2025 | Elevated Latency on Search Tooling This issue has been resolved. | 40 min | Resolved |
| Sep 8, 2025 | llama-3.3-70b-versatile Performance IssueThis incident has been resolved. | 1h 52m | Resolved |
| Sep 5, 2025 | moonshotai/kimi-k2-instruct Unavailable or Degraded PerformanceThe issues affecting moonshotai/kimi-k2-instruct have been **resolved**. The model is operating normally. We apologize for the disruption and thank you for your patience. | 37 min | Resolved |
| Sep 4, 2025 | me-central-1 Data Center Failure Impacting CapacityWe have **fully resolved** the failure in our me-central-1 region. All services are operating at normal capacity. The incident was caused by a **network failure**, which has been corrected. We apologize for the… | 1h 35m | Resolved |
| Sep 3, 2025 | API Errors Affecting Multiple RegionsThe authentication issue impacting multiple regions has been fully resolved. Services are operating normally, and error rates remain stable. | 41 min | Resolved |
| Aug 25, 2025 | Latency Issues Across Various ModelsThis incident has been resolved. | 1h 29m | Resolved |
How to Read This Groq Incident Log
Each row is an incident PulsAPI observed, not a summary written afterwards. The duration is wall-clock time between the first failing check and the first clean one, so it includes the window before Groq acknowledged anything. Vendor post-mortems typically measure from acknowledgement, which is why their numbers are usually shorter.
Incidents still open have no duration yet and are listed as ongoing rather than being given a running total. The archive covers the last 365 days; anything older has aged out of the window rather than never having happened.
The longest single Groq outage in this window ran 16h 7m, against a mean recovery of 2h 3m. If you depend on Groq in a customer-facing path, the longest figure is the one to design around. The mean is what happens on a normal bad day; the maximum is what happens on the worst one.