Weights & Biases Outage History
Daily status observations, past incidents, and reported issue history for Weights & Biases.
Checking current status...
90-Day Trend
Monthly Status Summary
| Month | Issue-free days | Days Tracked | Days with Issues |
|---|---|---|---|
| August 2026 | 100% | 25 | 0 |
| July 2026 | 80.6% | 31 | 6 |
| June 2026 | 73.3% | 30 | 8 |
| May 2026 | 80% | 5 | 1 |
This percentage summarizes normalized provider-status observations by calendar day. It is not duration-based, component-weighted, or contractual uptime. See the methodology and limitations.
Daily Status (Last 91 Days)
May 27
Today
Operational
Degraded
Partial Outage
Major Outage
Maintenance
No Data
Incident History
July 2026
Errors fetching media files
Started:
investigating
We are currently experiencing elevated errors fetching media files in Weights & Biases.
Metric Ingestion Backup
Started:
identified
We are currently experiencing elevated ingestion delay due to high volume.
Metrics Ingestion Backup
Started:
identified
We are experiencing metric ingestion delay of ~1 hour
Metric Ingestion Backlog
Started:
identified
We're currently experiencing elevated ingestion time of about 30 minutes for some run metrics.
Timeouts and 503s for serverless inference
Started:
monitoring
Impact has been mitigated.
identified
Serverless Inference users could experience timeouts and 503 response codes while attempting to leverage models on the product. We are actively engaged in troubleshooting and will post an update here as soon as we know more.
Metric Ingestion Delay
Started:
identified
We are currently seeing a backup in metric ingestion of up to 20 minutes
June 2026
Logins failing
Started:
monitoring
An Auth0 incident that had earlier prevented logins has recovered, allowing restored login access. We are continuing to monitor the situation.
Auth0 Incident: https://status.auth0.com/incidents/z8sc9m1gzqb7
identified
We have identified an issue with Auth0 that may cause some users to be unable to log in.
investigating
We are investigating an issue where logging in to wandb.ai may fail for some users.
Web App loading issue
Started:
monitoring
A fix has been deployed and we are monitoring.
investigating
We are continuing to investigate this issue.
investigating
We're investigating reports of the W&B web app (app.wandb.ai) showing a blank page on load for some users. The API and SDK data logging remain operational.
A fix is in progress and we'll share an update soon.
Metric ingestion delayed
Started:
monitoring
We identified an infrastructure issue as the root cause and have resolved it. We are rapidly working through the backlog of ingested metrics and should fully catch up soon.
identified
We are currently responding to delay of up to 5 hours for metrics written into wandb. We have identified the issue and are working to process metrics as quickly as possible. There is no data loss and all data will be complete once the backlog has drained. We're very sorry for the disruption.
Metrics ingestion delay
Started:
investigating
We are currently investigating an issue where metric ingestion is delayed for some runs up to 1 hour.
May 2026
Slowness and missing metrics
Started:
monitoring
Starting at 10:48AM PST we experienced an incident causing missing metrics and slowness across the application. As of 12:10 PM metrics should now be restored (there was no data loss) and we are continuing to monitor performance.
Delayed Metrics Processing – Workspace & API Updates Impacted
Started:
identified
We are currently experiencing delays in processing newly logged metrics.
As a result, recent metrics from active runs may take longer than usual to appear in charts, dashboards, reports, and API queries.
No data is lost. All metrics are being processed and will become visible once the backlog clears.
We are actively working to restore normal processing times and will provide updates here.
Run Update Ingest Delays
Started:
monitoring
A fix has been implemented and we are monitoring the results.
identified
We identified an issue where updates to W&B Models runs may be delayed for up to 30 minutes. There is no data loss.
API Errors
Started:
monitoring
We are continuing to monitor for any further issues.
monitoring
A fix has been implemented and we are monitoring the results.
investigating
We are currently investigating this issue.
monitoring
The frontend and backend APIs have fully recovered after we mitigated excess load on the system.
We will continue to actively monitor the situation to ensure stability.
investigating
We are currently investigating this issue.
April 2026
Delayed Metrics Processing – Workspace & API Updates Impacted
Started:
identified
We are currently experiencing delays in processing newly logged metrics.
As a result, recent metrics from active runs may take longer than usual to appear in charts, dashboards, reports, and API queries.
No data is lost. All metrics are being processed and will become visible once the backlog clears.
We are actively working to restore normal processing times and will provide updates here.
Backend API Errors
Started:
investigating
We are currently investigating an issue where API requests may sporadically fail.
Elevated API Errors
Started:
monitoring
We are continuing to monitor for any further issues.
monitoring
The issue has been identified and a fix has been rolled out. We are monitoring for any further issues.
investigating
We are currently investigating an issue where requests are failing.
API and UI outage
Started:
monitoring
We are seeing the site recover and be fully operational - we are continuing to actively monitor the situation to ensure stability.
monitoring
A fix has been implemented and we are monitoring the results.
identified
The issue has been identified and a fix is being implemented.
Delay in files and media getting uploaded
Started:
identified
We are continuing to work on a fix for this issue.
identified
We have identified a delay in files and media getting uploaded.
March 2026
W&B Inference Maintenance
Started:
investigating
The W&B Inference endpoint is currently unavailable. This is due to some maintenance on the gateway that was scheduled but that we missed the announcement step for. We expect intermittent outages between 9am and 11am PDT. We apologize for any inconvenience.
Elevated API Errors
Started:
investigating
We're experiencing an elevated level of API errors and are currently looking into the issue.
investigating
We're experiencing an elevated level of API errors and are currently looking into the issue.
Elevated API request latencies
Started:
monitoring
We've deployed mitigations and are actively monitoring. Request latencies have decreased but are not yet back to baseline.
We appreciate your patience as we work to resolve this issue.
investigating
We are aware of an issue causing elevated API request latencies, resulting in degraded performance in the UI and SDK.
Recent changes in traffic patterns are resulting in increased load, which we urgently are working to address.
Delayed Metrics Processing – Workspace & API Updates Impacted
Started:
identified
We are currently experiencing delays in processing newly logged metrics.
As a result, recent metrics from active runs may take longer than usual to appear in charts, dashboards, reports, and API queries.
No data is lost. All metrics are being processed and will become visible once the backlog clears.
We are actively working to restore normal processing times and will provide updates here.