New Multi-region uptime checks and custom-domain status pages

Harness Outage History

Daily status observations, past incidents, and reported issue history for Harness.

Checking current status...
47.3% of 91 observed days had no reported issue

90-Day Trend

May 28Aug 25

Monthly Status Summary

Month Issue-free days Days Tracked Days with Issues
August 2026 68% 25 8
July 2026 38.7% 31 19
June 2026 36.7% 30 19
May 2026 60% 5 2

This percentage summarizes normalized provider-status observations by calendar day. It is not duration-based, component-weighted, or contractual uptime. See the methodology and limitations.

Daily Status (Last 91 Days)

May 27 Today
Operational Degraded Partial Outage Major Outage Maintenance No Data

Incident History

August 2026
Feature Management & Experimentation (FME) user interface unavailable
major 17m

Started:

FME
monitoring
A fix has been implemented and we are monitoring the results.
investigating
We are currently investigating reports that the Feature Management & Experimentation (FME) user interface is failing to load. Customers attempting to access the FME console may encounter errors or unresponsive pages. Feature flag evaluation and SDK traffic are not believed to be affected. A further update will follow shortly.
investigating
We are currently investigating this issue.
All modules are running slow in Prod1/2/3/4 due to cloud provider incident
major 2h 7m

Started:

Platform Platform Platform Platform
monitoring
We are observing improved latencies across the board following the fix implemented by our cloud provider. We are continuing to monitor the situation closely and will provide further updates as warranted. We noted some stuck executions for CI for couple of customers, which we are investigating
identified
Our cloud provider has confirmed an ongoing incident impacting multiple regions. Harness pipelines have not experienced failures as a result, though some users may continue to experience slowness. We are monitoring the situation closely and will provide updates as more information becomes available.
investigating
We are continuing to investigate this issue.
investigating
Our cloud provider is facing an active incident and we are following up.
investigating
The slowness could cause either of below symptoms: - Pipelines not starting - Delays in execution - Pipelines being cancelled due to timeouts
investigating
We are currently investigating this issue.
FME API write operations started returning 499 errors
minor 32m

Started:

FME
monitoring
A fix has been implemented and we are monitoring the results.
identified
The issue has been identified and a fix is being implemented.
investigating
We are currently investigating this issue.
Data ingestion is delayed on Traceable US production
major 12h 46m

Started:

US - app.traceable.ai / api.traceable.ai
monitoring
We are continuing to monitor for any further issues.
monitoring
A fix has been implemented and we are monitoring the results.
identified
The issue has been identified and a fix is being implemented.
investigating
We are continuing to investigate this issue.
investigating
We are continuing to investigate this issue.
investigating
We are currently investigating this issue.
SEI 2.0 dashboards are not loading
major 323h 38m

Started:

identified
The issue has been identified and a fix is being implemented.
investigating
We are continuing to investigate this issue.
investigating
We are currently investigating a Harness component that is experiencing issues. We are working to identify the cause and restore normal operations as soon as possible.
Monitoring - Pipelines Stuck - Prod2
minor 324h 19m

Started:

monitoring
We are monitoring the stuck pipelines in prod2. For the customers who are still seeing stuck pipelines, we request you to abort and re-trigger.
monitoring
We are monitoring the stuck pipelines in prod2. The new executions are passing as we are continuously monitoring the services.
Editing 'Variable Sets' in the IaCM module is experiencing issue
minor 374h 2m

Started:

investigating
We are continuously rolling out fixes to all Clusters, Prod3 has been resolved.
investigating
We are continuously rolling out fixes to all Clusters, Prod EU1 has been resolved.
investigating
We have identified the issue and started to implement the fix , prod2 is restored.
investigating
We are continuing to investigate this issue.
investigating
We are continuing to investigate this issue.
investigating
We are currently investigating this issue.
Legacy Dashboards - Degraded
minor 407h 1m

Started:

monitoring
A fix has been implemented and we are monitoring the results.
investigating
We are currently investigating this issue.
July 2026
UI dashboards are lagging behind (CI)
461h 46m

Started:

monitoring
A fix has been implemented and we are monitoring the results.
identified
The issue has been identified and a fix is being implemented.
investigating
We are currently investigating this issue.
Harness Artifact Registry upload is failing from pipeline - EU1 region
minor 468h 20m

Started:

monitoring
A fix has been implemented and we are monitoring the results.
identified
The issue has been identified and a fix is being implemented.
investigating
We are currently investigating this issue.
Intermittent External Network Connectivity Issues Affecting Build VMs
minor 500h 10m

Started:

monitoring
We are continuing to monitor for any further issues.
monitoring
A fix has been implemented and we are monitoring the results.
investigating
Summary - We are intermittently facing network connectivity issues with our Build VM's unable to connect to external resources. We are currently investigating the issue.
Prod3 Filestore is failing with HTTP 500 errors
major 569h 0m

Started:

monitoring
A fix has been implemented and we are monitoring the results.
investigating
We are currently investigating this issue.
The Prod3 & Prod1 environment is experiencing intermittent outages. We are currently investigating the issue.
major 571h 26m

Started:

monitoring
We are continuing to monitor for any further issues.
monitoring
We are continuing to monitor for any further issues.
monitoring
A fix has been implemented and we are monitoring the results.
identified
We are continuing to work on a fix for this issue.
identified
The issue has been identified and a fix is being implemented.
investigating
We are continuing to investigate this issue.
investigating
We are currently investigating this issue.
monitoring
A fix has been implemented and we are monitoring the results.
identified
The issue has been identified and a fix is being implemented.
investigating
We are currently investigating this issue.
AIDI Dashboards – Degraded Performance
major 683h 14m

Started:

monitoring
A fix has been implemented and we are monitoring the results.
identified
The issue has been identified and a fix is being implemented.
investigating
We are investigating an issue impacting AIDI dashboards. Users may experience increased load times or intermittent failures when accessing dashboards. Our team is actively working to identify the root cause and restore normal performance. We will provide updates as more information becomes available.
Degraded CI performance
minor 800h 52m

Started:

monitoring
A fix has been implemented and we are monitoring the results.
identified
The issue has been identified and a fix is being implemented.
investigating
We are currently investigating this issue.
CI git clone calls are failing for specific arm builds
minor 807h 38m

Started:

monitoring
We are continuing to monitor for any further issues.
monitoring
A fix has been implemented and we are monitoring the results.
identified
The issue has been identified and a fix is being implemented.
investigating
We are currently investigating this issue.
Code services are degraded, git clone is failing
minor 835h 10m

Started:

investigating
We are currently investigating this issue.
Certain users are unable to see feature flags in prod1 and prod2
minor 878h 43m

Started:

investigating
Certain users are unable to see feature flags in prod1 and prod2
Few customers experiencing intermittent errors during pipeline execution.
minor 989h 47m

Started:

monitoring
A fix has been implemented and we are monitoring the results.
investigating
We are currently investigating this issue.
Testing Dev Status Page 9 jul
minor 1000h 47m

Started:

investigating
Please check dev
Prod3: Slowness in Harness Platform
minor 1026h 19m

Started:

monitoring
A fix has been implemented and we are monitoring the results.
investigating
We are currently investigating this issue.
monitoring
A fix has been implemented and we are monitoring the results.
investigating
We are currently investigating an issue in Prod3 cluster where Harness platform components are loading slower than expected.
IACM module is impacted in EU1
major 1047h 28m

Started:

monitoring
We have implemented a fix on our side which has resolved this issue. We are continuing to monitor this issue.
investigating
We are currently investigating an issue in EU1 where users are having issues using IACM module.
Prod3: Slowness in loading pipeline components
minor 1047h 47m

Started:

monitoring
A fix has been implemented in one of our microservice responsible for dealing with pipeline templates which has resolved the issue. We are continuing to monitor the issue.
investigating
We are currently investigating an issue in Prod3 cluster where pipeline components are loading slower than expected.
FME - Some customers are experiencing delays in scheduled exports of impressions
minor 1162h 21m

Started:

monitoring
We are continuing to monitor results while the backlog is worked through.
monitoring
We are seeing the backlogged items worked through and continuing to monitor results.
monitoring
A fix has been implemented and we are monitoring the results.
June 2026
Dev Statuspage incident
major 1216h 52m

Started:

investigating
Dev Statuspage incident - Testing
Few pipelines in prod4 are running with degraded performance
major 1235h 13m

Started:

monitoring
A fix has been implemented and we are monitoring the results.
identified
We are continuing to work on a fix for this issue.
identified
We are continuing to work on a fix for this issue.
identified
Harness is investigating an issue affecting our Prod 4 environments. Our log-services are affected and harness is currently working on restoring function.  This can affect our CI, Chaos, IACM modules.
investigating
We are currently investigating this issue.
Testing Dev Status Page
critical 1237h 31m

Started:

investigating
We are continuing to investigate this issue.
investigating
Testing CD outage
Login to Traceable Clusters Impacted
minor 1314h 48m

Started:

monitoring
We are continuing to monitor for any further issues.
monitoring
Issue has been fixed and we are closely monitoring
identified
The issue has been identified and a fix is being implemented.
investigating
We are currently experiencing intermittent login issues with the Traceable cluster due to an issue with our login provider. We are actively working with the provider to resolve the issue. This incident does not impact data ingestion, and all ingestion pipelines continue to function normally. We will provide updates as we have more information.
Harness Login Failure in EU1 environment
major 1358h 6m

Started:

monitoring
We are continuing to monitor for any further issues.
monitoring
A fix has been implemented and we are monitoring the results.
investigating
Harness login is failing. We are investigating the problem.
Integration Authentication Issue Affecting Data Ingestion
minor 1387h 22m

Started:

identified
We have identified an issue affecting a small subset of accounts in the EU1 region. As a result, data ingestion for impacted integrations may have stopped.
Testing Dev Status Page 19June
major 1480h 58m

Started:

monitoring
A fix has been implemented and we are monitoring the results.
investigating
Testing Dev in QA
IDP Workflows and Backend failures
minor 1521h 24m

Started:

monitoring
A fix has been implemented and we are monitoring the results.
identified
IDP Systems' communication was interrupted due to an update.  Harness is rolling back changes. Impacted environments - prod0, prod1, prod2
identified
The issue has been identified and a fix is being implemented.
investigating
We are currently investigating this issue.
Cloud builds are failing with 429 Too Many Requests error while downloading artifacts from central sonatype registry.
minor 1524h 26m

Started:

monitoring
We are continuing to monitor for any further issues.
monitoring
Maven Central is rate-limiting requests originating from our us-central1 NAT egress IP. Other regions (us-west1, us-east5) are not affected. The Maven/Sonatype team has been engaged and is actively working on lifting the rate limit for our IP range.
monitoring
A fix has been implemented and we are monitoring the results.
identified
The issue has been identified and a fix is being implemented.
investigating
We are investigating an issue affecting artifact downloads from Maven Central following recent rate-limiting changes implemented by Sonatype. Sonatype has recently implemented rate limits and builds downloading artifacts from central sonatype  https://central.sonatype.org/faq/429-contact-support/. Customers downloading artifacts would be impacted.  We are actively engaging with Sonatype to mitigate the issue here.
Testing Dev Status Page 16 Jun
major 1551h 38m

Started:

investigating
We are currently investigating this issue.
Certain customers may see odd web artifacts on Code UI. Logging out and Logging in resolves this issue.
minor 1736h 36m

Started:

monitoring
We are continually monitoring this issue.
Intermittent 503s on IDP services
minor 1803h 13m

Started:

monitoring
A fix has been implemented and we are monitoring the results.
investigating
We are currently investigating this issue.
FME SDKs - SDKs falling back to polling
minor 1835h 35m

Started:

monitoring
A fix has been implemented and we are monitoring the results.
identified
Instability in streaming connections could cause a delay in propagating updates. All SDKs are designed to poll against the CDN on streaming disconnects and instability scenarios. This could also cause some noise in the logs as SDKs try to reconnect to streaming. But updates are unaffected.
Testing Dev Status Page 3rd June
major 1879h 11m

Started:

identified
The issue has been identified and a fix is being implemented.
investigating
Please ignore this incident.
Hosted CI Builds Failing Intermittently
minor 1912h 7m

Started:

monitoring
A fix has been implemented and we are monitoring the results.
investigating
We are currently investigating this issue.
May 2026
Testing Dev Status Page
minor 2001h 48m

Started:

investigating
Dev Testing
Pipeline logging streams experiencing degraded delivery to console.
minor 2027h 49m

Started:

monitoring
A fix has been implemented and we are monitoring the results.
investigating
We are currently investigating this issue.
Hosted CI Builds Failing Intermittently
major 2029h 33m

Started:

monitoring
A fix has been implemented and we are monitoring the results.
identified
Harness is currently implementing a failover to mitigate the intermittent issue for some customers
identified
Harness has implemented a change and are seeing failures reduced, but are continuing to work on completing mitigation. Customers can expect executions to succeed more frequently. Prod1/Prod2 are back to normal.
identified
Harness is continuing to investigate and implement changes to fully restore functionality. At this time, we are still seeing some CI Builds intermittently fail.
identified
We are continuing to work on a fix for this issue.
identified
The issue has been identified and a fix is being implemented.
investigating
We are currently investigating this issue.
Testing Dev Status Page incident
major 2050h 26m

Started:

investigating
Testing Dev Status Page incident message
Prod3 pipeline execution UI slowness
minor 2172h 13m

Started:

monitoring
A fix has been implemented and we are monitoring the results.
identified
The issue has been identified and a fix is being implemented.
investigating
Our investigation has revealed that the pipeline execution are happening as expected in the backend, The slowness is limited to updates in pipeline UI only.
investigating
We are currently investigating an slowness issue in pipeline execution ui.
Prod3 pipeline slowness
minor 2199h 20m

Started:

monitoring
A fix has been implemented which resolved the pipeline slowness , we are currently monitoring the results.
investigating
We have received reports of slowness in Harness pipelines in Prod3 cluster. We are actively investigating this issue. Please monitor this space for further updates.
Custom Dashboards are failing intermittently in prod3
minor 2309h 34m

Started:

monitoring
A fix has been implemented and we are monitoring the results.
investigating
We are currently investigating this issue.
IACM Pipeline Struck In Prod4
minor 2325h 14m

Started:

monitoring
We are continuing to monitor for any further issues.
monitoring
A fix has been implemented and we are monitoring the results.
investigating
We are currently investigating this issue.
Pipeline Updates in Prod4 is taking time.
minor 2340h 33m

Started:

monitoring
A fix has been implemented and we are monitoring the results.
identified
The issue has been identified and a fix is being implemented.
investigating
We are continuing to investigate this issue.
investigating
We are continuing to investigate this issue.
investigating
We are currently investigating this issue.
IACM Pipeline Struck
minor 2341h 9m

Started:

monitoring
A fix has been implemented and we are monitoring the results.
identified
The issue has been identified and a fix is being implemented.
investigating
We are continuing to investigate this issue.
investigating
We are continuing to investigate this issue.
investigating
We are currently investigating this issue.
Traceable APAC Cluster Slowness
minor 2343h 11m

Started:

identified
We have identified a potential issue causing the service access problem and are working hard to address it. Please continue to monitor this page for updates.
investigating
We are currently investigating this issue.
Hosted CI customers using secure connect are facing connectivity issues
minor 2380h 31m

Started:

monitoring
A fix has been implemented and we are monitoring the results.
identified
The issue has been identified and a fix is being implemented.
investigating
We are currently investigating this issue.
Platform access issues in Prod1/Prod2/Prod3
minor 2384h 35m

Started:

identified
The issue has been identified and a fix is being implemented.
investigating
We are continuing to investigate this issue.
investigating
We are currently investigating this issue.
Testing Dev Status Page incident
minor 2392h 49m

Started:

investigating
incident testing
Vercel Integration for FME is having delays to synchronize changes
minor 2403h 26m

Started:

monitoring
A fix has been implemented and we are monitoring the results.
investigating
We are currently investigating this issue.
AI Summary Feature is not available across all Insights
minor 2416h 18m

Started:

identified
The issue has been identified and a fix is being implemented.
investigating
We are currently investigating this issue.
Delegate Upgrades are failing for some customers
2419h 21m

Started:

monitoring
A fix has been implemented and we are monitoring the results.
identified
The issue has been identified and a fix is being implemented.
investigating
We are currently investigating this issue.
FME SDK are experiencing elevated error rates and delays in response.
minor 2481h 3m

Started:

monitoring
A fix has been implemented and we are monitoring the results.
investigating
We are currently investigating this issue.
Deployment Degradation - Slowness in Prod3
critical 2514h 8m

Started:

monitoring
We are continuing to monitor for any further issues.
monitoring
A fix has been implemented and we are monitoring the system.
monitoring
A fix has been implemented and we are monitoring the results.
identified
Executions are running fine now. However, there is still some slowness in the execution graph generation so we might see some delay in the execution updates to show up in UI. But executions would work without any delay/slowness.
identified
We have identified the root cause as elevated queue depth and throughput saturation on the MongoDB layer, leading to processing latency. A fix is being implemented to optimize handling and restore database stability
investigating
We are continuing to investigate this issue.
investigating
We are currently investigating this issue.
Deployment Degradation – Failures in Prod1,2,3
minor 2514h 14m

Started:

monitoring
A fix has been implemented and we are monitoring the results.
investigating
We are continuing to investigate this issue.
investigating
We are continuing to investigate this issue.
investigating
Deployment Degradation – Failures in Prod1,2,3
AI Summary Feature Unavailable in SEI Insights
minor 2529h 57m

Started:

investigating
We are currently investigating an issue where the AI Summary feature in the Harness SEI module is not available across all Insights. Our engineering team is actively working to restore the functionality as quickly as possible. We will share further updates as more information becomes available.
infracost integration degraded in iacm pipelines
minor 2552h 54m

Started:

investigating
We are currently investigating this issue.
K8s Customer Billing Data will be outdated
minor 2562h 12m

Started:

investigating
We are currently experiencing an issue, K8s customer billing data will be stale.
Prod3 experiences slowness in pipelines
minor 2649h 40m

Started:

monitoring
We are largely mitigated and most pipelines are running normally. We are monitoring all parameters to make sure there are no issues before closing it.
monitoring
We are continuing to monitor for any further issues.
monitoring
We are continuing to monitor for any further issues.
monitoring
A fix has been implemented and we are monitoring the results.
investigating
We are currently investigating this issue.
Intermittent slowness during pipeline executions (Prod1, Prod2)
minor 2651h 6m

Started:

monitoring
We are largely mitigated and most pipelines are running normally. We are monitoring all parameters to make sure there are no issues before closing it.
monitoring
A fix has been implemented and we are monitoring the results.
investigating
We are currently investigating this issue.
April 2026
FF (Classic) some API's have elevated latencies in prod2
minor 2666h 59m

Started:

monitoring
A fix has been implemented and we are monitoring the results.
identified
Issue has been identified and fix is underway
investigating
We are currently investigating this issue.
Platform is experiencing degraded performance for some organizations.
minor 2673h 43m

Started:

identified
Issue has been identified and mitigated
investigating
We are currently investigating this issue.
FME UI errors on creating/updating Flags for subset of customers who have Jira integration.
minor 2674h 12m

Started:

investigating
We are currently investigating this issue.
Testing Dev Status Page Degradation Testing
major 2722h 41m

Started:

investigating
We are currently investigating this issue.
Testing Dev Status Page Adjusting to test the Active incident length
major 2726h 18m

Started:

monitoring
A fix has been implemented and we are monitoring the results.
investigating
We are currently investigating this issue.
CI pipelines using cache intelligence are seeing degraded performance.
minor 2746h 4m

Started:

identified
The issue has been identified and a fix is being implemented.
investigating
We are currently investigating this issue.
Degraded Performance — Feature Flags in PROD2
minor 2818h 3m

Started:

monitoring
A fix has been implemented and we are monitoring the results.
investigating
We are currently investigating this issue.
Testing Dev Status Page
major 2828h 45m

Started:

investigating
We are currently investigating this issue.
Testing Dev Status Page
critical 2885h 57m

Started:

investigating
We are currently investigating this issue.
Testing Dev Status Page
major 2915h 38m

Started:

investigating
We are currently investigating this issue.
IACM infrastructure pipelines using terraform are currently experiencing an outage
major 2940h 0m

Started:

identified
The issue has been identified and a fix is being implemented.
investigating
We are currently investigating this issue.
We are noticing CI hosted build failures
minor 3018h 53m

Started:

monitoring
A fix has been implemented and we are monitoring the results.
identified
We are continuing to work on a fix for this issue.
identified
The issue has been identified and a fix is being implemented.
investigating
We are currently investigating this issue.
Feature Flags unable to update
minor 3026h 13m

Started:

monitoring
A fix has been implemented and we are monitoring the results.
investigating
We are continuing to investigate this issue.
investigating
We are currently investigating this issue.
Identified and fixed feature flag metrics impact calculations not progressing and monitoring
minor 3081h 21m

Started:

monitoring
Feature flag metrics impact calculations were not progressing. This issue does not impact experiment calculation. We expect delays in metrics impact calculations as the queue drains. Currently monitoring progress after a rolling out a fix.
Legacy Run Test step is failing intermittently for all customers in Prod2
minor 3188h 39m

Started:

monitoring
A fix has been implemented and we are monitoring the results.
identified
The issue has been identified and a fix is being implemented.
investigating
Some of the legacy run test step connectivity to test intel service is failing intermittently. We are currently investigating the issue here.
Prod2 is facing login issues
minor 3192h 26m

Started:

monitoring
A fix has been implemented and we are monitoring the results.
identified
The issue has been identified and a fix is being implemented.
investigating
We are currently investigating this issue.
Experiencing issues impacting pipeline executions.
minor 3205h 0m

Started:

monitoring
A fix has been implemented and we are monitoring the results.
investigating
We are currently investigating this issue.
Degraded Performance — Pipeline Insights Dashboards
minor 3321h 16m

Started:

monitoring
We are continuing to monitor for any further issues.
monitoring
A fix has been implemented and we are monitoring the results.
identified
The issue has been identified and a fix is being implemented.
investigating
We are currently investigating this issue.
CI degradation with CI steps using AWS connector with inherited authentication
minor 3346h 56m

Started:

investigating
We are investigating a degradation in CI steps when using AWS connectors and inherited authentication.
Autostopping service degraded in AWS (Middle East South 1)
minor 3379h 24m

Started:

monitoring
Update: AutoStopping functionality for AWS has been restored for all regions except me-south-1. The issue was caused by elevated latency from AWS in the affected region, impacting operations such as warm-up, cool-down, schedule execution, and traffic detection. We have now isolated this region to prevent impact on other customers. Resources in me-south-1 will continue to experience the issue until the region fully recovers. We are actively monitoring the situation and will provide further upd...
identified
Status update : CCM AutoStopping functionality for the AWS cloud provider is currently impacted due to increased latency from AWS in the me-south-1 region. This is affecting multiple operations, including warm-up, cool-down, schedule execution, and traffic detection. In addition, CCM Asset Governance functionality is also impacted for resources in the me-south-1 region. We are actively working on isolating/excluding the affected region to restore functionality for the remaining customers. Re...
identified
The issue has been identified and a fix is being implemented.
March 2026
Feature Flag SDK authentication operations are running slow in Prod2
minor 3559h 54m

Started:

monitoring
We are continuing to monitor for any further issues.
monitoring
A fix has been implemented and we are monitoring the results.
identified
The issue has been identified and a fix is being implemented.
investigating
We are currently investigating this issue.
Degraded Performance for SCIM users during login.
minor 3681h 54m

Started:

monitoring
The issue has been identified and a fix is in place.
monitoring
The issue has been identified and a fix is in place. (Note: It is still in a degraded state)
investigating
The issue has been identified and a fix is in place.
Test Intelligence service is impacted in Prod1
minor 3728h 4m

Started:

monitoring
A fix has been implemented and we are monitoring the results.
identified
The issue has been identified and a fix is being implemented.
investigating
We are currently investigating this issue.
FME SDK is experiencing elevated error rates for Impressions and events
minor 3729h 32m

Started:

monitoring
We are now monitoring the results.
investigating
The issue started around ~6:45AM PT and the team is currently investigating
Degraded performance in CI Steps in Prod 2 and Prod 3
minor 3825h 56m

Started:

monitoring
A fix has been implemented and we are monitoring the results.
identified
The issue has been identified and a fix is being implemented.
investigating
We are noticing degraded performance in CI Steps in Prod 2 and Prod 3 environments The issue is intermittent. We are investigating the cause
Prod 2 - Customers may see some executions from March 11 in a "running" but hung state
3849h 48m

Started:

identified
The issue has been identified and a fix is being implemented.
investigating
Customers may continue to see that some pipeline executions show that they are "running" even though they have completed, aborted, or failed as a result of yesterday's incident. (https://status.harness.io/incidents/4y4dl47v2qhc) This behavior is a UI-only artifact from the incident and should not affect customers' ability to start new executions.  We are working on clearing these artifacts.