New Multi-region uptime checks and custom-domain status pages

Honeycomb Outage History

Daily status observations, past incidents, and reported issue history for Honeycomb.

Checking current status...
84.6% of 91 observed days had no reported issue

90-Day Trend

May 28Aug 25

Monthly Status Summary

Month Issue-free days Days Tracked Days with Issues
August 2026 96% 25 1
July 2026 87.1% 31 4
June 2026 70% 30 9
May 2026 100% 5 0

This percentage summarizes normalized provider-status observations by calendar day. It is not duration-based, component-weighted, or contractual uptime. See the methodology and limitations.

Daily Status (Last 91 Days)

May 27 Today
Operational Degraded Partial Outage Major Outage Maintenance No Data

Incident History

August 2026
Canvas not loading
minor 38m

Started:

ui.honeycomb.io - US1 App Interface ui.eu1.honeycomb.io - EU1 App Interface
monitoring
We have deployed the fix and are seeing recovery of normal Canvas functionality.
identified
Canvas is not loading for customers in the US and EU regions and we have identified what we believe to be the cause and are rolling out a fix.
US1 cold query outage
minor 229h 7m

Started:

monitoring
From 11:36UTC to 12:43UTC, queries in US1 that read older data returned errors or timed out. Queries over recent data were un-affected, and no telemetry is being lost. The issue has been identified, and a fix is in place. We are currently monitoring the situation to confirm resolution.
July 2026
Triggers intermittently failing in EU region
minor 564h 55m

Started:

monitoring
We are continuing to monitor for any further issues.
monitoring
The partial degradation has been resolved for all customers except for those specifically contacted. We are working on mitigation measures so that we can fully restore trigger functionality for everyone.
monitoring
We are seeing a regression in performance with partial degradation of both triggers and querying. We are working to remediate.
monitoring
We have identified the source of the issue, applied a temporary fix, and are working on a permanent fix. We continue to monitor the situation.
investigating
We are continuing to see issues with triggers and are also noticing issues with querying. We are investigating for the cause of both.
investigating
We are seeing triggers intermittently failing in the EU region. We are actively investigating the cause.
honeycomb.io marketing website not working
major 1159h 58m

Started:

monitoring
Clearing the CDN cache appears to have solved the issue.
investigating
We are currently investigating this issue.
June 2026
Activity Log delayed in US
minor 1228h 1m

Started:

monitoring
Starting at 14:24 PT, there was an issue with the Activity Log causing events to stop being processed. You may notice a gap starting at that time. We are slowly backfilling these events and do not expect any data loss to occur. We expect the full data log to be restored around 17:30 PT.
Ingest outage in EU
minor 1327h 59m

Started:

monitoring
We are monitoring the situation. Events may have been delayed between 17:53 and 18:04 UTC.
investigating
We are investigating an issue with delayed ingest in the EU region. SLOs and Trigger evaluations may also be delayed.
Ingest Outage in US and EU
major 1351h 55m

Started:

monitoring
Triggers, SLOs and Service Maps are recovered. We are continuing to monitoring the Ingest Service
monitoring
We have rolled back the deploy and ingest service has recovered. There has been an ingest outage from 17:55 - 18:14 UTC. We have also observed an delay for Service Maps.
investigating
We are investigating an issue with delayed ingest in the US and EU region. Our engineers are rolling back a recent deploy. SLOs and Trigger evaluations may also be delayed.
MCP tool access degraded
minor 1473h 29m

Started:

monitoring
We are aware of an issue affecting Honeycomb MCP. Write tool functionality may be degraded. Read tools are unaffected.
investigating
We are aware of an issue affecting Honeycomb MCP. Write tool functionality may be degraded. Read tools are unaffected.
Querying Issues
minor 1524h 8m

Started:

monitoring
Querying in our Production EU Region is fully operational. We will continue to monitor for any abnormal behavior.
investigating
We are continuing to investigate intermittent slowness and query failures affecting our Production EU Region. We will provide an update as soon as we have more information.
Querying Issues in EU
minor 1526h 7m

Started:

investigating
We are currently investigating the cause of slowness and query failures in our Production EU Region.
Delayed ingest in EU
minor 1685h 10m

Started:

identified
The issue has been identified and a fix is being implemented.
investigating
Queries should no longer be delayed. SLOs and Trigger evaluations remain impacted.
investigating
We are investigating an issue with delayed ingest in the EU region. Received events are still being stored, but there may be a delay in event retrieval in queries. SLOs and Trigger evaluations may also be delayed.
Canvas UI issues
minor 1737h 31m

Started:

investigating
We are currently investigating this issue.
Canvas tools unavailable
minor 1853h 18m

Started:

identified
Canvas agent cannot load honeycomb internal tools.
May 2026
Investigating querying issues
minor 2162h 46m

Started:

investigating
Impact from this issue has concluded. For a period of time (different between US and EU instances, noted below), queries spanning data older than the most recent 2 hours with certain GROUP / WHERE clauses returned inconsistent results. Data was never lost, and reruns of the same queries will show the previously missing data. US impact times: 21:50 UTC - 23:20 UTC EU impact times: 20:50 UTC - 23:20 UTC
investigating
We have identified and are working to resolve an issue that is causing query results to return inconsistent results.
activity log for US instance is delayed
minor 2490h 9m

Started:

monitoring
Activity Log recovered at 07:00 UTC-07:00 today (about 6.5h ago).  All streams are caught up.
monitoring
All activity log streams are now caught up, except for the query runs table, which should have all data since the start of the incident backfilled by 1600 PDT.
monitoring
We are continuing to monitor for any further issues.
monitoring
We believe that the activity log should now be caught up.
monitoring
Backfilling of the past 2h30m of data is now in progress and should complete shortly.
identified
Activity log delay is rising, and is more than an hour behind due to a MySQL replica failover. Estimated recovery by 04:00 PDT.
April 2026
Links in Slack triggers not working
2675h 6m

Started:

monitoring
The fix has been implemented and links on Slack triggers are functional again.
identified
We discovered that links in Slack triggers are not working and have identified the cause and are in process of implementing a fix
Activity Log events interrupted for EU
major 2913h 13m

Started:

identified
Activity Log events are interrupted, starting from around 15:51UTC. Customers may experience some data lost since the beginning of the incident.
Activity Log ingest is delayed
minor 3031h 4m

Started:

identified
Activity Log is experiencing a delay in processing events. We are working on remediating the issue.
Activity Log delayed ingest in EU
minor 3052h 33m

Started:

monitoring
A fix has been implemented and we are monitoring recovery.
identified
Activity Log events may be delayed, starting from around 19:55UTC. We're actively working on remediation.
Ingest issues for some customers
major 3156h 27m

Started:

monitoring
We are continuing to work on the systemic solution and monitor. We'll leave the incident open until we're able to confirm the issue won't recur.
monitoring
The manual fix has been applied to all environments.
monitoring
We've applied the manual fix to nearly all affected datasets (less than 1% remaining). The systemic solution is still in progress.
identified
We've identified a manual fix and are deploying that by hand to each affected environment, while also investigating a systemic solution.
identified
We have further narrowed down that only teams on Classic environments are currently seeing issues. We are figuring out mitigation mechanisms at the moment.
identified
We have confirmed reports of some teams getting issues ingesting their data. We have identified a probable source behind this behavior and are currently trying to correct it.
Activity log delay in production US
minor 3182h 16m

Started:

identified
SLO processing has fully recovered.
identified
SLOs experienced a period of additional processing delays. A fix has been deployed and recovery is ongoing.
identified
We have identified an issue with triggers and SLOs and have implemented a fix. Customers should no longer see degraded performance for triggers or SLOs. A fix for the activity log delay has been implemented and is in the process of recovering. Customers may continue to see activity log unavailability while this completes.
identified
We have identified an issue with triggers and SLOs and are actively investigating. Customers may experience delays when creating or updating triggers and SLOs. A fix for the activity log delay has been implemented and is in the process of recovering. Customers may continue to see activity log unavailability while this completes.
identified
We have identified the cause and are working to resolve the issue.
March 2026
Honeycomb UI displaying an error
critical 3704h 1m

Started:

identified
We have identified the issue and are working on a fix.
investigating
We are currently investigating this issue.
Elevated Querying Errors
minor 3854h 13m

Started:

identified
We have rolled back a code change to our query engine, but our triggers service has not recovered. We are investigating.
identified
We are investigating elevated error rates affecting queries in our US1 region. Our team is actively working on a resolution. We will provide updates as we learn more.