DigitalOcean servers unable to pull BuildGrid images
Started August 24, 2026 at 7:08 PM UTC · Ongoing
IssuesMinor incident
Affected components
Google Cloud Platform Google Cloud Storage
identified
DigitalOcean's IPv6 range appears to be blocked from accessing Google Cloud Storage. This means CSv1 and CSv2 customers who have servers with public IPv6 addresses are currently unable to pull new BuildGrid images onto their servers.
identified
GCS confirmed that an internal issue is causing the DigitalOcean IPv6 range to be blocked, and is currently working on a resolution.
identified
We are still waiting on GCS to resolve the issue.
DigitalOcean servers unable to pull BuildGrid images
Started August 24, 2026 at 4:24 PM UTC · Ongoing
IssuesMinor incident
Affected components
Google Cloud Platform Google Cloud StorageBuildGrid
identified
DigitalOcean's IPv6 range appears to be blocked from accessing Google Cloud Storage. This means CSv1 and CSv2 customers who have servers with public IPv6 addresses are currently unable to pull new BuildGrid images onto their servers.
identified
We are continuing to work on a fix for this issue.
Google login issues
Started June 12, 2025 at 7:09 PM UTC · 1h 21m
IssuesMinor incident
Affected components
Google Cloud Platform Google Kubernetes EngineGoogle Cloud Platform Google Cloud StorageCore SystemsGoogle Cloud Platform Google Cloud SQLGoogle Cloud Platform Google Cloud NetworkingGoogle Cloud Platform Google Compute Engine
investigating
Google authentication issues is having issues causing login and access issues for our systems and customers. We are investigating the impact and mitigation strategies.
resolved
The underlying issue with GCP seems to be resolved, and we are not seeing further issues with the logins and other systems.
Kubernetes Cluster Scale/Creation
Started January 2, 2025 at 12:24 PM UTC · 3h 21m
IssuesMinor incident
Affected components
Core Systems
investigating
We are aware of problems when scaling/creating Kubernetes clusters. This appears to be caused by new rate limits imposed by docker hub. The team are working on a solution.
identified
The issue has been identified and a fix is being implemented.
investigating
We are currently investigating this issue.
resolved
This incident has been resolved.
Stuck Deployments
Started June 28, 2024 at 1:35 AM UTC · 12h 12m
Pending
Affected components
Core Systems
investigating
We are currently investigating a number of stuck deployments.
monitoring
We have identified the issue causing the deployments to be stuck and fixed the issue. This was caused by a database failure at our cloud provider which we are investigating and monitoring.
resolved
This incident has now been resolved.
CustomConfig git access
Started May 13, 2024 at 8:27 AM UTC · 7d 6h
IssuesMinor incident
Affected components
Core Systems
investigating
We are aware of issues with our CustomConfig git repositories. This issue impacts direct git access to the CustomConfig git repository. Direct access to CustomConfig pages via the web dashboard and API is not affected. We are investigating the root cause of the problem.
investigating
Applications are able to be deployed, but the CustomConfig backend is not available for the moment.
resolved
The CustomConfig git backend migration has been completed
Production Issues
Started March 19, 2024 at 10:07 AM UTC · 6m
IssuesMinor incident
Affected components
Core SystemsContainerNet
investigating
We are investigating issues affecting some production systems.
investigating
We are continuing to investigate this issue.
resolved
This incident has been resolved.
Github API Outage
Started May 9, 2023 at 11:45 AM UTC · 3h 36m
Pending
Affected components
Core Systems
monitoring
Github is experiencing some API outages, deployments don't currently appear to be affected. We are monitoring (https://www.githubstatus.com/)
resolved
This incident has been resolved.
Multiple Google Cloud services in the europe-west9 region are impacted
Started April 26, 2023 at 3:36 PM UTC · 16h 46m
IssuesMinor incident
Affected components
Core Systems
investigating
Water intrusion in a data center in europe-west9 has caused a multi-cluster failure and has led to a shutdown of multiple zones. We expect general unavailability of the europe-west9 region. There is no current ETA for recovery of operations in the europe-west9 region at this time, but it is expected to be an extended outage. Customers are advised to failover to other regions if they are impacted.
investigating
Due to a major outage in Google Cloud (https://status.cloud.google.com/regional/europe) - Kubernetes installations may fail on installations on servers located in the same region, regardless of the cloud, due to Kubernetes images themselves being hosted on Google Cloud
monitoring
We are waiting on Google resolutions
resolved
Maestro installations are working as expected again
Github Not Listing Repositories
Started October 6, 2022 at 1:07 PM UTC · 1d 9h
IssuesMinor incident
Affected components
Github
investigating
Currently, GitHub is having an issue with listing repositories while using the Cloud 66 Github app. As a workaround while this issue is ongoing.
1. Remove the Github installation.
2. Manually configure SSH key access to Github, please see the link below.
https://help.cloud66.com/rails/how-to-guides/common-tools/access-your-code#manually-configuring-github-access
identified
The issue has been identified and a fix is being implemented and this should resolve the issue for most users. If the error is still persisting, please use the workaround provided:
1. Remove the Github installation.
2. Manually configure SSH key access to Github, please see the link below.
https://help.cloud66.com/rails/how-to-guides/common-tools/access-your-code#manually-configuring-github-access
resolved
Issue confirmed to be fixed by GitHub.
Subset of Buildgrid Image pushes slower than normal
Started September 14, 2022 at 8:15 PM UTC · 4d 18h
IssuesMinor incident
Affected components
BuildGrid
identified
The team is aware that some Buildgrid pushes are taking longer than normal, and are working on a mitigation strategy
identified
Our engineering team is continuing to look into mitigations for this issue.
monitoring
The team has completed the first part of the mitigation strategy for intermittent slow Buildgrid pushes. We are now monitoring ongoing performance.
resolved
This incident has been resolved.
Production Agent Reporting Outage
Started September 9, 2022 at 7:20 AM UTC · 1h 31m
IssuesMinor incident
Affected components
Core SystemsBuildGridGithubContainerNetRealtime Push Systems
identified
The issue has been identified and a fix is being implemented.
resolved
This incident has been resolved.
postmortem
At approximately 07:00 UTC we started to perform a minor system quality-of-life backend update. The update had unintended consequences on an older system component which hadn’t been changed in a while. The component was to do with agent server communications, and it ended up disabled after the update. Although the same component was unaffected during testing in dev/staging and passed UAT, it turned out that there was a necessarily different configuration around rewrites in production that caused the issue.
The nature of the update meant that rolling back wasn’t straightforward, so the team resolved to fix the incompatibility in-place, and as such were unable to stop the erroneous “server down” notifications that then went out. The issue was resolved approximate 1 hour later. at 08:00 UTC - after which time servers would have started to appear as “online” again.
Subsequent to this outage, the difference in configuration has been added to our operations processes, such that this will not occur again. Our apologies for the inconvenience and concern that this may have created!
Github outage
Started September 6, 2022 at 11:53 AM UTC · 5m
OutageMajor incident
Affected components
Github
investigating
GitHub is having an outage at this time. This may adversely affect Deployments.
resolved
This incident has been resolved.
Azure reporting DNS problems (affecting Ubuntu 18.04)
Started August 30, 2022 at 9:09 AM UTC · 5d 23h
Pending
identified
From Azure (https://status.azure.com/en-gb/status)
Starting at approximately 06:00 UTC on 30 Aug 2022, a number of customers running Ubuntu 18.04 (bionic) VMs recently upgraded to systemd version 237-3ubuntu10.54 reported experiencing DNS errors when trying to access their resources. Reports of this issue are confined to this single Ubuntu version.
identified
Azure is recommending that affected customers reboot their servers to obtain an updated DHCP lease to resolve this issue
monitoring
We are continuing to monitor this incident
resolved
This incident has been resolved.
Cloud 66 outage history and incident timeline | Uptimus