Tavus Outage History
49 incidents recorded over the last 365 days, 48 of them resolved. Durations are measured from when PulsAPI first saw the problem to when it cleared, which is usually longer than the vendor's own figure.
Every Recorded Tavus Incident
| Date | Incident | Duration | Status |
|---|---|---|---|
| Aug 13, 2026 | Elevated start up times for some conversationsWhen one of our GPU providers performed maintenance, we asked to migrate traffic to a different cluster beforehand to avoid disruptions. During the migration back, the provider encountered an issue that led to some… | -19 min | Resolved |
| Aug 5, 2026 | Intermittent Response Issues with Cartesia TTSCartesia has fully restored service, and conversations using Cartesia TTS are responding normally. | 14 min | Resolved |
| Aug 5, 2026 | Some Custom Replica Creations failing due to upstream Infra Provider IssueThis incident has been resolved. | 1h 15m | Resolved |
| Aug 3, 2026 | Some Conversations Failing to StartAll conversations are starting normally. | 26 min | Resolved |
| Aug 1, 2026 | Cartesia IncidentThis incident has been resolved. | 18 min | Resolved |
| Jul 30, 2026 | Increased conversation latency using the Gemma 4 modelConversation latency for the Gemma 4 model has returned to normal, and all traffic has been routed. | 41 min | Resolved |
| Jul 29, 2026 | Some conversations are fail to startCartesia has restored service, and conversations configured to use Cartesia for TTS are starting successfully again. | 14 min | Resolved |
| Jul 24, 2026 | Some conversations are unable to startConversation creation temporarily failed for approximately 10 minutes. The issue has been resolved, and conversation creation is operating normally. | 13h 10m | Resolved |
| Jul 15, 2026 | OSS LLM (tavus-oss) elevated latencyUncovered and rolled back related change - OSS LLM (tavus-oss) traffic is healthy again. | 2h 29m | Resolved |
| Jul 6, 2026 | Delaying Joining on Google Meet InvitesThis incident has been resolved. | 57 min | Resolved |
| Jun 17, 2026 | Replica Join Failures for New ReplicasA fix has been deployed and this incident is now resolved. A very limited subset of conversations using newly trained replicas were affected by join failures. We identified the root cause and do not expect any further… | 1h 49m | Resolved |
| Jun 16, 2026 | Async video generation and replica training are processing slower than expectedThis issue has been resolved, and all existing replica trainings and video generations are now processing as expected. Our team is continuing to monitor all runs. | 13 min | Resolved |
| May 8, 2026 | Some conversations not startingThis incident has been resolved. | 26 min | Resolved |
| Apr 30, 2026 | Degraded Performance Across Multiple ServicesThis incident has been resolved. | 5h 13m | Resolved |
| Apr 29, 2026 | Increased error rates with perception modelsThe incident has been resolved, and Raven is now working as expected. | 5h 8m | Resolved |
| Apr 17, 2026 | Intermittent Call Connection IssuesCloudflare has implemented a fix and downstream services are returning to normal service | 3h 45m | Resolved |
| Apr 14, 2026 | Replica takes longer to join in some conversationsThis has been resolved. | 36 min | Resolved |
| Apr 10, 2026 | Elevated Error Rate When Joining ConversationsThis incident has been resolved | 14 min | Resolved |
| Apr 5, 2026 | Conversations not properly initiatingThis incident has been resolved. | 7 min | Resolved |
| Apr 4, 2026 | Some conversations slow to bootThis incident has been resolved. | 19 min | Resolved |
| Mar 26, 2026 | Some conversations not initiatingWe experienced a brief disruption due to a configuration change that impacted conversation routing, causing some conversations to stall or take longer to start. The issue has been resolved and all systems are now… | 32 min | Resolved |
| Mar 25, 2026 | Some conversations are encountering increased startup timesThis incident has been resolved. | 14 min | Resolved |
| Mar 10, 2026 | Some conversations took longer time startWe detected several conversations with start times exceeding 10 seconds. We have diverted traffic and are continuing to monitor the situation. | 0 min | Resolved |
| Mar 6, 2026 | Some conversations not starting upWe have identified an issue with one of our GPU providers and shifted traffic away from them | -18 min | Resolved |
| Feb 17, 2026 | List Replicas EndpointAPI is back to operational | 17 min | Resolved |
| Feb 12, 2026 | Conversations slow to joinThis incident has been resolved. | 24 min | Resolved |
| Jan 23, 2026 | Conversations are seeing delayed initiationThis incident has been resolved. | 1h 41m | Resolved |
| Jan 21, 2026 | Some conversations not initiating due to webRTC issueThis incident has been resolved. | 3h 11m | Resolved |
| Jan 15, 2026 | Some conversations not initiatingThis incident has been resolved. | 1h 1m | Resolved |
| Jan 7, 2026 | Increased wait times for replica to join conversationsThis incident has been resolved. | 2h 39m | Resolved |
| Dec 18, 2025 | Temporary Video Generation OutageThis incident has been resolved. | 3 min | Resolved |
| Dec 6, 2025 | Elevated error rates with Llama-4 due to downstream provider errorsThis incident has been resolved. | 53 min | Resolved |
| Dec 2, 2025 | Llama-4 Elevated Error RatesThis incident has been resolved. | 1h 22m | Resolved |
| Dec 2, 2025 | Replica Training and Video Generation Audio IssueThis incident has been resolved | 2h 37m | Resolved |
| Nov 20, 2025 | Intermittent errors on Conversation CreationOur firewall identified anomalous traffic patterns and the systems are now stabilized. | 16 min | Resolved |
| Nov 20, 2025 | LiveKit Service DisruptionThis incident has been resolved. | 4h 17m | Resolved |
| Nov 18, 2025 | Conversation creation failingServices are back to normal operations. Let us know if you still see issues support@tavus.io . | 6h 18m | Resolved |
| Nov 13, 2025 | Upstream LLM Provider DisruptionThis incident has been resolved. | 4h 0m | Resolved |
| Nov 12, 2025 | Experiencing intermittent log in issuesThis incident has been resolved. | 20 min | Resolved |
| Nov 12, 2025 | Increased utterance to utterance latency for tavus-llamaThis incident has been resolved. | 1h 30m | Resolved |
| Nov 11, 2025 | Conversation Start Failures (GPU Provider Degradation)This incident has been resolved. | 3h 7m | Resolved |
| Nov 5, 2025 | Conversation creation failingThis outage was caused by a GPU provider outage, and traffic has now been redirected. All conversations are properly initiating and services are fully restored. Our team will continue to closely monitor. | 10 min | Resolved |
| Oct 26, 2025 | Scheduled Maintenance - Sun, Oct 26, 2025 at 2:00–3:00 AM PDT (09:00–10:00 UTC)The scheduled maintenance has been completed. | 30 min | Scheduled |
| Oct 20, 2025 | Some Conversations Not Joined By Replica due to Ongoing AWS outageThis incident has been resolved. | 13h 5m | Resolved |
| Oct 20, 2025 | AWS outage causing Conversation Creation TimeoutsSome conversation creation requests were impacted by the AWS us-east-1 incident and resulted in a 504 Gateway Timeout. This affected CVI. The incident lasted from 10:00 UTC and lasted until 10:20 UTC. | 2h 30m | Resolved |
| Oct 14, 2025 | Cartesia outage resulting in text to speech failuresThis incident has been resolved. | 19h 46m | Resolved |
| Oct 5, 2025 | Delayed Video Generation CallbacksSome video generation callbacks didn't properly deliver between October 3–5 as the callback jobs remained in an "in-progress" state and did not send as expected. The issue has been fully fixed now, and all affected… | 0 min | Resolved |
| Sep 30, 2025 | Video List RegressionThis incident has been resolved | 34 min | Resolved |
| Sep 20, 2025 | Some CVI Conversations failing to startThis incident has been resolved. | 11 min | Resolved |
How to Read This Tavus Incident Log
Each row is an incident PulsAPI observed, not a summary written afterwards. The duration is wall-clock time between the first failing check and the first clean one, so it includes the window before Tavus acknowledged anything. Vendor post-mortems typically measure from acknowledgement, which is why their numbers are usually shorter.
Incidents still open have no duration yet and are listed as ongoing rather than being given a running total. The archive covers the last 365 days; anything older has aged out of the window rather than never having happened.
The longest single Tavus outage in this window ran 19h 46m, against a mean recovery of 2h 19m. If you depend on Tavus in a customer-facing path, the longest figure is the one to design around. The mean is what happens on a normal bad day; the maximum is what happens on the worst one.