Hi CircleCI Team.
I wanted to open a thread to ask about recent platform reliability and see if there’s any context on why we’ve seen an uptick in outages this past week:
-
October 9, 2026 (Ongoing / Today): Major disruption causing delays in processing pipelines, executing workflows/jobs, sending webhooks/notifications, and a ~20–30 minute lag in UI updates across API and dashboard operations.
-
October 7, 2026: A cluster of three separate incidents occurred back-to-back:
-
Increased task wait times for Docker Gen2 (18:54 – 20:05 UTC)
-
Elevated level of infrastructure failures on customer jobs (14:27 – 15:21 UTC)
-
Delay on starting Machine Job Tasks (13:40 – 14:06 UTC)
-
-
October 1, 2026: Delays starting Docker Gen 2 jobs (~2-hour impact).
These recurring delays are directly blocking our deployments and slowing down our entire engineering team.
Could someone from CircleCI share:
-
Is there a common root cause or infrastructure bottleneck behind these recent queue backlogs?
-
What steps are being taken to prevent these regressions from repeating?
We rely heavily on CircleCI for our core delivery pipeline, so any transparency on platform stability would be much appreciated.
Thanks