We are currently investigating issues with Iron dashboard (hud-e). IronMQ, IronWorker and IronCache services should work fine. If you have any questions - write us to [email protected]
resolved
This incident has been resolved.
IronWorker degraded performance
Início 18 Genver 2026 da 22:47 UTC · 7d 22h
IssuesMinor incident
Componentes afetados
IronCacheIronWorker PublicIronWorker Dedicated
investigating
We are experiencing issues with upstream provider. Currently in touch with them.
monitoring
A fix has been implemented and we are monitoring the results.
monitoring
We are continuing to monitor for any further issues.
resolved
This incident has been resolved.
IronWorker degraded performance
Início 16 Ebrel 2024 da 22:57 UTC · 2d 11h
OutageMajor incident
Componentes afetados
IronWorker PublicIronWorker Dedicated
investigating
We are currently investigating this issue.
monitoring
A fix has been implemented and we are monitoring the results
resolved
This incident has been resolved.
IronCache issue
Início 2 Cʼhwevrer 2024 da 04:56 UTC · 0m
OutageMajor incident
resolved
Our engineering team is actively resolving the situation, and undertaking necessary steps to recover the system. We expect to post an update momentarily.
Iron has completed a review of its systems and the Log4j security issue does not affect Iron's services.
IronWorker Degraded Performance
Início 6 Here 2021 da 00:16 UTC · 19m
IssuesMinor incident
Componentes afetados
IronWorker Dedicated
identified
The issue has been identified and a fix is being implemented.
monitoring
A fix has been implemented and we are monitoring the results.
resolved
This incident has been resolved.
IronWorker Degraded Performance
Início 7 Gwengolo 2021 da 12:58 UTC · 3h 14m
IssuesMinor incident
Componentes afetados
IronWorker Public
investigating
IronWorker customers may experience delays while pushing the tasks. We are currently investigating the issue.
monitoring
A fix has been implemented and we are monitoring the results.
resolved
This incident has been resolved.
IronWorker outage
Início 27 Mezheven 2021 da 03:37 UTC · 2h 13m
OutageMajor incident
Componentes afetados
IronWorker PublicIronWorker Dedicated
investigating
We are currently investigating this issue.
resolved
This incident has been resolved.
IronWorker Degraded Performance
Início 14 Meurzh 2021 da 07:21 UTC · 0m
Pending
Componentes afetados
IronWorker Public
resolved
IronWorker customers could experience issues while pushing tasks to Public Cluster on March 14 between 12:00 am UTC and 6:30 am UTC. We've resolved the issue. Our development team is monitoring the situation.
Database Upgrade
Início 3 Meurzh 2021 da 09:30 UTC · 0m
Pending
resolved
Amazon performed an upgrade of our PostgreSQL DB version. This DB stores logins, passwords, tokens and other information about our customers. While upgrading they shutted down the database instance, performed the upgrade, and restarted the database instance. As a result, you could see HTTP 401 errors in your logs between 09:28 am UTC and 09:45 am UTC.
Here is more information: https://forums.aws.amazon.com/ann.jspa?annID=8176
IronWorker Degraded Performance
Início 30 Ebrel 2020 da 23:00 UTC · 0m
Pending
resolved
On May 1, at 11 am UTC we identified the issue with our autoscale functionality: there were errors while starting our IronWorker service on new instances but it was working fine on existing ones. After investigating into the errors thrown, we found the issue was happening while pulling our docker images from DockerHub. The root cause was problems on dockerhub side information about which you can find on their status page: https://status.docker.com/
In order to resolve the issue, we have manually launched our service on new instances and temporarily disabled the autoscale functionality.
IronCache Issue
Início 25 Here 2019 da 11:00 UTC · 53m
IssuesMinor incident
Componentes afetados
IronCache
investigating
We are currently investigating this issue.
identified
The issue has been identified and a fix is being implemented.
resolved
This incident has been resolved.
IronCache Issue
Início 9 Gouere 2019 da 21:17 UTC · 1h 34m
OutageMajor incident
Componentes afetados
IronCache
investigating
We are currently investigating this issue.
identified
The issue has been identified and a fix is being implemented.
resolved
This incident has been resolved.
IronWorker Degraded Performance
Início 14 Mae 2019 da 21:10 UTC · 2h 39m
OutageMajor incident
Componentes afetados
IronWorker PublicIronWorker Dedicated
identified
Due to a database upgrade issue, a portion of our IronWorker customers are experiencing issues with certain API commands. We've identified the issue and are in the process of resolving.
identified
We are continuing to work on a fix for this issue.
identified
Migration is still in progress. This is taking more time than expected but we're monitoring it closely.
resolved
The migration has completed and service has returned to normal.
postmortem
**Overview**
On May 13th, at 03:29 UTC, we began routine database upgrades. During the upgrade process we noticed errors in our logs indicating certain queries weren’t able to complete successfully.
**What went wrong**
After investigating into the errors thrown, we found data anomalies in our Production data set that didn’t exist in our Staging data set. This difference resulted in slow queries and errors that cascaded into service interruptions for a subset of our customers.
**What we're doing to prevent this from happening again**
Moving forward we’re taking steps to ensure our Staging data set is 100% up to date with our Production data set. If the copies of the data were exact, this would have been caught in Staging and wouldn’t have caused a disruption in service.
**Resolution time**
The incident was resolved at 11:49 UTC
Scheduler service: maintenance work
Início 6 Du 2018 da 23:52 UTC · 44m
Pending
Componentes afetados
IronWorker PublicIronWorker Dedicated
investigating
Amazon has scheduled for maintenance one of the instances where our scheduler is running. They say that the instance will be unavailable for 2 hours: on November 7, from 12:00 am to 2:00 am UTC. We're going to run another scheduler on another instance at that period of time.
resolved
This incident has been resolved.
IronWorker degraded performance
Início 17 Here 2018 da 20:52 UTC · 1h 14m
IssuesMinor incident
Componentes afetados
IronWorker Public
identified
The issue has been identified and a fix is being implemented.
**Overview**
On August 6th, at 15:07 UTC, we noticed connectivity issues across our network. These connectivity issues caused IronMQ to degrade into an unhealthy state which rendered the service un-usable.
**What went wrong**
At 12:49 AM PDT, the vendor who we rely on for DNS \(AWS Route 53\) experienced issues. In-network connectivity was broken and many components of our network were unable to communicate with each other. When the vendor issue was resolved at 1:04 AM PDT, the issue persisted within our network due to caching and TTL issues.
**What we're doing to prevent this from happening again**
* We identified the places within network that could have caused this issue and reviewed their caching strategies and TTL times. Multiple cache times were too aggressive and we’ve increased timeouts in the necessary places. We’re testing various failure scenarios within our staging network to confirm the validity of these timeout values.
* We’re currently discussing backup DNS strategies as a team and will be posting updates on our blog about our strategy moving forward, and, continued progress.
**Resolution time**
The incident was resolved at 16:04 UTC.
IronMQ (aws-us-east) service degradation
Início 20 Ebrel 2018 da 18:29 UTC · 1d 21h
IssuesMinor incident
investigating
We are currently investigating this issue.
identified
The issue has been identified and a fix is being implemented.
monitoring
A fix has been implemented and we are monitoring the results.
resolved
This incident has been resolved.
IronMQ (aws-us-east) service degradation
Início 16 Ebrel 2018 da 16:20 UTC · 1d 8h
IssuesMinor incident
investigating
We are currently investigating this issue.
identified
The issue has been identified and a fix is being implemented.
monitoring
A fix has been implemented and we are monitoring the results.