New Multi-region uptime checks and custom-domain status pages

LiveKit Outage History

Daily status observations, past incidents, and reported issue history for LiveKit.

Checking current status...
76.9% of 91 observed days had no reported issue

90-Day Trend

May 28Aug 25

Monthly Status Summary

Month Issue-free days Days Tracked Days with Issues
August 2026 56% 25 11
July 2026 90.3% 31 3
June 2026 80% 30 6
May 2026 80% 5 1

This percentage summarizes normalized provider-status observations by calendar day. It is not duration-based, component-weighted, or contractual uptime. See the methodology and limitations.

Daily Status (Last 91 Days)

May 27 Today
Operational Degraded Partial Outage Major Outage Maintenance No Data

Incident History

August 2026
Agent session analytics displaying incorrect values in EU Central and India regions
minor 4h 3m

Started:

India - Analytics Ingestion Europe Central - Analytics Ingestion
monitoring
We are continuing to monitor while the historical data backfill completes, and will resolve this incident once all analytics for the affected window are fully restored.
monitoring
We identified an issue where agent session analytics in the Cloud dashboard displayed zero concurrent sessions, beginning late on August 21 (UTC). This was a reporting issue only — agent sessions, calls, and all real-time services in these regions operated normally throughout. Ingestion has been restored and we are backfilling historical analytics for the affected window; some charts may show incomplete history until this completes. We are monitoring while the backfill finishes.
Investigating reports of elevated Egress API errors
minor 39m

Started:

Global Egress
investigating
We had a spike of errors between 20:52 and 21:01 UTC, and the errors have subsided to baseline as of 21:01 UTC. We are actively monitoring and will share more updates in 30 minutes.
investigating
We are currently investigating reports of elevated API errors across the Egress service. We will post another update in 5 minutes.
Investigating delayed session data ingestion on Cloud Dashboard
minor 6h 36m

Started:

Cloud Dashboard (cloud.livekit.io)
monitoring
mitigations have been deployed and we are now catching up to realtime.
identified
the amount of data we are ingesting in the sessions view is outpacing our ability to ingest, thus causing a delay in ingestion. we are pursuing a few mitigations right now including moving to a larger database instance to relieve the processing bottleneck.
identified
We are continuing to work on a fix for this issue.
identified
Sessions data ingest remains delayed for about an hour. It's not impacting other dashboards. we've identified the source of the slow ingest and are working on a mitigation. will update again in 30 mins
investigating
We are currently investigating delayed session data ingestion on Cloud Dashboard. No services appear to be impacted and we don't expect any data to be lost.
Investigating reports of failed SIP Transfer calls in US East
minor 55h 58m

Started:

US East - SIP
investigating
No transfer failures has been observed since 18:05 UTC and SIP transfers are currently completing normally. Impact was limited to a small percentage of calls between 15:48-16:08 UTC and 17:24-18:05 UTC. We're actively monitoring while we investigate the cause and put safeguards in place to prevent a recurrence.
investigating
We're investigating an elevated rate of failures when transferring active SIP calls in the US East region. SIP calls themselves remain connected, and inbound and outbound calling are otherwise operating normally.
Hosted agent builds failing in US East
minor 124h 59m

Started:

monitoring
The total build error rate has decreased significantly due to the mitigation, but a small percentage of builds in US East are still failing. We are continuing to monitor and will follow up with another update as soon as possible.
monitoring
We have deployed a mitigation and builds are now successful. We will continue monitoring while we collect any further details to share.
identified
New hosted agent builds are failing in US East. Existing agent workloads are not affected. We are currently deploying a mitigation and will follow up with more information shortly.
Investigating longer than usual dashboard loading times
minor 149h 27m

Started:

monitoring
We are seeing a 5-10 minute delay in loading some sessions, but we are confident that no data is being lost. We are currently monitoring and will update again once ingestion times have returned to baseline.
investigating
We are currently investigating longer than usual dashboard loading times. No services appear to be impacted and we don't expect any data to be lost.
Investigating alerts for increased timeout error rates on ListParticipants APIs
175h 34m

Started:

monitoring
Error rates on ListParticipants API requests returned to normal as of 18:58 UTC. We will follow up with more details as soon possible.
investigating
We are currently investigating alerts for elevated server error rates on ListParticipants APIs across multiple regions. The impact appears to be intermittent and limited to a subset of room-management API requests. We will follow up with more details as soon possible.
Investigating alerts for degraded room service
292h 46m

Started:

monitoring
We have applied a mitigation to the affected regions and are now monitoring to ensure that performance signals return to baseline. We are still working on determining impact (including how users can determine if they were affected) and will share those details as soon as possible.
investigating
We are currently investigating automated alerting which triggered for degraded Room service performance in the US West and US Central regions beginning around 19:45 UTC. We are working to determine if there is any customer-facing impact. If there is impact, we don't currently have reason to believe that it is widespread. We will update again within 30 minutes.
Elevated error rate on LiveKit Inference (google/gemma-4-31b-it)
minor 318h 17m

Started:

monitoring
A burst of requests increased the failure rate. Both primary and secondary model providers failed to fulfill the requests at the time; we are root-causing the issue. We are continuing to monitor.
identified
We are actively investigating elevated error rate on LiveKit Inference (google/gemma-4-31b-it)
identified
Between approximately 18:45 and 19:15 UTC today, a subset of LiveKit Inference requests using the google/gemma-4-31b-it model returned errors. Affected agent sessions would have seen an LLM request error on those requests. Other models and other LiveKit services were not affected. No action is needed on your part. If you continue to see errors, please reach out to support. We apologize for the disruption.
Elevated timeouts on LiveKit Inference (google/gemma-4-31b-it)
minor 365h 23m

Started:

monitoring
We are monitoring elevated timeouts that affected LiveKit Inference between 14:05 and approximately 15:00 UTC today. During that window, roughly 3% of LLM requests using the google/gemma-4-31b-it model timed out. Impacted customers would have seen LLM request timeouts in their agent sessions (for example, APITimeoutError in agent logs). Other models were not affected. We have identified the cause. During a period of increased traffic, response times for this model slowed, and a defect in our...
Sessions view not populating in the Cloud Dashboard
minor 394h 38m

Started:

investigating
We have traced the underlying issue to a backlog in our data pipeline, and we are continuing to investigate the root cause while working to clear the backlog. This data ingest delay results in the Sessions view returning no results for any time range ending at the current time, including all Quick Ranges options in the dashboard. As a temporary workaround, select "Custom range" and set the end time at least 30 minutes in the past to load your sessions. Real-time connectivity is not affected a...
investigating
We are currently investigating an issue where the sessions view does not populate in the Cloud Dashboard. This issue is specific to the dashboard display; all LiveKit services are operating normally and no sessions or sessions data are affected.
July 2026
Investigating longer than usual dashboard loading times
487h 16m

Started:

investigating
We are currently investigating longer than usual dashboard loading times. No services appear to be impacted and we don't expect any data to be lost.
Elevated create room error rates in London
minor 880h 20m

Started:

monitoring
Between 08:54 am and 09:06 am UTC, some requests to create new rooms in our London region failed with errors. Error rates returned to normal by 09:06 am UTC and a fix has been applied. We are continuing to monitor.
Elevated connection failures in Japan
minor 1063h 29m

Started:

investigating
We are currently investigating increased connection failures affecting our Japan region beginning at 18:18 UTC.
June 2026
Connection failures and SIP transfer errors in US Central
major 1405h 3m

Started:

monitoring
Update: As soon as the US Central region became unresponsive around 12:25 UTC, traffic was automatically re-routed to the nearest healthy region. Customers may have noticed failed API requests for approximately 1 minute while the re-route took place. Failed SIP Transfers continued until we finished manually draining the impacted region at 13:09 UTC. We are also investigating impact to Agent Dispatches and will post another update when we know more.
monitoring
We have begun routing traffic away from US Central to mitigate reports of WebRTC connection failures and SIP transfer errors.
Partial Outage in LiveKit Cloud Dashboard
minor 1496h 19m

Started:

investigating
Session data (except Agent Insights) on the dashboard can still be accessed by changing the timeline on the Sessions page to "Past 7 days" or greater. We're still working to restore full access.
investigating
Due to networking issues faced in our US East region, certain features of Cloud Dashboard are not accessible, including Sessions page and Agent Insights. There is no data loss or ongoing session impact associated with this. Our team is actively working on restoring the service.
Investigating connection failure and SIP transfer error reports in US East
minor 1496h 59m

Started:

monitoring
We have routed traffic away from US East region and connection failure errors are back to baseline. We are continuing to investigate and monitor SIP transfer errors.
investigating
We are currently investigating connection failure reports in US East.
Elevated rate of egress ending early in India region
minor 1552h 20m

Started:

monitoring
We identified an issue in the India region where a small number of active sessions may have ended earlier than expected. Affected sessions could experience egress recordings stopping mid-session. We've deployed a mitigation and are monitoring to confirm resolution. Other regions are unaffected.
Intermittent errors for Google Gemini models via Inference
minor 1664h 16m

Started:

identified
We have identified the cause as an upstream limit on our Google Gemini API account. Automatic failover to an alternate Gemini deployment is in place and has reduced the impact, but a subset of requests to Google Gemini models may still intermittently return errors when failover capacity is exceeded. We are actively engaged with Google support to restore full capacity and will share further updates as we have them. Other models and providers remain unaffected.
investigating
We are currently investigating intermittent 429 errors for Google Gemini models routed through LiveKit Inference.
Elevated latency in Hyderabad region
minor 1817h 2m

Started:

investigating
As of 09:21 UTC, we've applied a mitigation by routing traffic away from the Hyderabad region. Latency rates are returning to normal levels. We're continuing to monitor the issue.
investigating
We're investigating elevated latency affecting Ingress and SIP services in our Hyderabad region.
Degraded Connectivity in US Central
minor 1884h 3m

Started:

monitoring
We've identified the root cause as a single problematic media node in US Central, which has been suspended at 14:38 UTC. We are continuing to monitor before marking this resolved.
monitoring
We have routed the traffic away from the US Central region at 14:00 UTC and are seeing the connection failures returning to normal levels. We are continuing to monitor the issue.
investigating
We're investigating elevated SIP call connection failures in our US Central region beginning at ~12:06 UTC. We are working to mitigate this issue.
May 2026
Elevated Reports of Participant Connection Latency and Errors In US East Region
critical 2003h 16m

Started:

monitoring
Service is operating normally with traffic routing through other regions. We are working on bringing US East back online and will share a detailed RCA once complete.
monitoring
Connection errors have fully cleared since the US East drain completed at 15:44 UTC. We'll continue monitoring. We will be sharing a detailed RCA.
monitoring
US East drain is complete and all new traffic is now routing to other regions. Error rate is trending down and we'll continue monitoring.
identified
We are observing database connection timeouts in the US East region and are currently seeing an impact to all services in the region. We are currently draining the region and routing away all traffic to other regions.
investigating
We are currently investigating reports of spikes in participant connection latency and errors
Partial data outage on LiveKit Cloud Dashboard
major 2046h 15m

Started:

monitoring
A fix has been implemented and we are seeing no more data outage for newly created sessions. We are monitoring the fix and are currently working on backfilling the missing data.
identified
We have identified an issue causing partial data unavailability in LiveKit Cloud Dashboard: - For some users, Agent Observability is temporarily unavailable for some sessions. - Some sessions in the 24-hour filter are showing as Active even though they have ended. A fix has been identified and we are backfilling the missing data.
investigating
We are investigating reports of Agent Observability not accessible for some sessions.
Network instability in Frankfurt
2141h 43m

Started:

monitoring
We have re-enabled our Frankfurt cluster and are monitoring for further disruptions.
investigating
We have observed network instability impacting cross-region traffic from Frankfurt beginning at 18:15 UTC. Traffic has been redirected from Frankfurt as of 19:55 UTC. Participants originally connecting via Frankfurt may see slightly increased latency until the region is restored.
Intermittent connectivity disruptions in US regions
minor 2313h 12m

Started:

monitoring
We have not observed further instances of connectivity disruptions since 13:19 UTC. We are continuing to investigate why our reroute mechanism did not activate in these isolated instances and are monitoring for further disruptions.
investigating
We are investigating intermittent connectivity disruptions affecting a subset of sessions, caused by internal network degradation between several of our US clusters (most heavily impacting US East). Affected sessions may see brief interruptions (less than 5 minutes) to media relay. We have observed these interruptions at the following times: - 5/12/26 05:00 UTC - 5/12/26 17:00 UTC - 5/13/26 19:00 UTC - 5/15/26 07:30 UTC We are actively investigating. Further updates to follow.
Outbound TLS calls are failing
major 2526h 22m

Started:

monitoring
We believe this incident was not provider-specific and as such have removed Twilio from the incident title.
monitoring
A fix has been implemented and we are monitoring the results. The issue persisted from 18:10 - 19:58 UTC. We will follow up with a postmortem as soon as possible.
investigating
We have noted that outbound calls using TLS via Twilio are failing. We are currently investigating and will update here as soon as we know more.
Cloud Agents experiencing deployment failures in US East
2542h 7m

Started:

monitoring
US-East recovered as of 08:55 UTC and has been serving traffic normally since. New and existing agent deployments are working as expected. We are moving to monitoring while we confirm sustained stability. A full post-incident review will follow.
investigating
We are continuing to work on restoring the Kubernetes API server in US-east. Starting at 08:15 UTC, we are observing impact to existing agent deployments in addition to new deployments. Mitigation is in progress.
investigating
etcd service in our US East Kubernetes cluster is currently down, resulting in API server unresponsiveness and failures for new deployments and redeployments. We are actively working on mitigating this with our data center provider.
investigating
Our team is actively working to resolve deployment failures in US East. Existing deployed agents should not be affected in any way.
investigating
New agent deployments are temporarily unavailable in the US East region, and our team is actively working on this. Dispatches to previously deployed agents are not affected.
investigating
We are currently investigating intermittent deployment failures on Cloud Agents in US East.
investigating
We are currently investigating degraded performance on Agents Hosted on LiveKit Cloud in US East.
April 2026
Investigating SIP participant timeouts and signalling connection errors
2815h 57m

Started:

identified
A fix is being implemented. We still have not received any further reports and believe the impact to be minor, but we're continuing to monitor for further issues.
investigating
We received a single report of CreateSIPParticipant timeouts in Singapore. While investigating this, we discovered an increased rate of signalling connection errors on a very small minority of requests, mainly in India. We haven’t received any other reports, but are proactively creating this incident in case users encounter increased connection latency or failed inbound SIP calls. We are continuing to investigate and will update here once we know more.
Increased Latency in RoomService APIs in US West Region
minor 3026h 45m

Started:

investigating
While applying a fix for the API latencies, we are temporarily seeing increased failure rates (about 10% of the calls) in RoomServices APIs, including CreateRoom and DeleteRoom APIs. We are actively working on mitigating this.
identified
We continue to see the long Room API latencies are also impacting other regions. The latency increases appear to originate from a specific table in our distributed database. The issue has been escalated with the database vendor and we are working on a workaround for decreasing the API latencies. Other services are not impacted.
investigating
We believe these elevated latencies began around 22:00 UTC. We have confirmed that only API requests in US-West should be impacted. The current list of impacted APIs appears to be CreateRoom, DeleteRoom, and UpdateRoomMetadata. We are working on mitigating the issue to return latencies back to normal.
investigating
We are continuing to investigate this issue.
investigating
We are investigating reports of increased latencies in RoomService APIs in the US West region, specifically on CreateRoom, DeleteRoom, and UpdateRoomMetadata APIs.
LiveKit Cloud Dashboard Down
critical 3320h 55m

Started:

investigating
We are currently investigating this issue and will update as soon as we know more.
Degraded Connectivity Issues – EU (Frankfurt)
minor 3375h 31m

Started:

monitoring
We have identified and implemented a fix for degraded connection errors affecting Real Time Communication and TURN services in our EU (Frankfurt) region. We are actively monitoring to confirm stability.
March 2026
Analytics updates on Cloud Dashboard are delayed
minor 3422h 42m

Started:

monitoring
We have initiated the process of backfilling delayed analytics data.
monitoring
The fix has been applied and real-time processing for the Cloud Dashboard is back online globally. We are monitoring the pipeline.
identified
A fix has been deployed. We expect the service to be restored shortly and will provide another update once it is back online.
identified
The issue has been identified. We are working on a fix.
investigating
We are currently investigating the issue. Only dashboard updates are affected (real time communication is not affected).
Degraded Performance LiveKit Cloud dashboard
minor 3425h 0m

Started:

monitoring
A fix has been implemented and we are monitoring the result
investigating
We are continuing to investigate this issue.
investigating
Users may be unable to access the dashboard. Our team is actively investigating.