部分停用 -- -- 服务检查显示器
- investigating
我们正在调查一个在我们欧盟区域可能无法评估某些服务检查监测员的问题.
- monitoring
一项措施已经执行,我们正在监测结果.
- resolved
这一事件已经得到解决.
自动翻译自官方事件更新。
55 Datadog Eu incidents · 2023年8月 — official updates, affected components, duration and resolution details.
我们正在调查一个在我们欧盟区域可能无法评估某些服务检查监测员的问题.
一项措施已经执行,我们正在监测结果.
这一事件已经得到解决.
自动翻译自官方事件更新。
我们正在调查从Datadog容器图像中提取的错误。 由于这个问题,一些用户可能无法下载或部署Datadog Agent,Cluster Agent等Datadog容器图像. 已经运行的集装箱不受影响.
这一问题已经确定,一个解决办法正在实施之中.
一项措施已经执行,我们正在监测结果.
这一事件已经得到解决.
自动翻译自官方事件更新。
我们正在积极调查多个Datadog产品的高误差率.
我们正在调查包括舰队自动化、安全产品、云工艺、云网络监测、无服务器、Kubernetes自动化和数据库监测在内的多种产品的高误率.
我们查明了错误率高的问题并正在通过缓解行动开展工作.
我们已部署缓解行动并正在监测恢复情况.
我们正在继续监测局势,并正在采取更多的预防措施.
这一事件已经得到解决.
自动翻译自官方事件更新。
We are investigating increased latency processing Metrics. As a result of this issue, some users may see delays or gaps for metrics on graphs.
We have improvement in some Monitors.
We are still investigating. Only process metrics and process check monitors are still affected.
We have deployed a fix and we are monitoring the results. We will provide another update once the issue is fully resolved.
This incident has been resolved.
We are investigating increased latency processing data. As a result of this issue, some users may see gaps or delays in RUM and other product graphs as well as empty or partial query results. To prevent false monitor alerts due to delayed data, monitors affected by the delay will not notify and will automatically resume once current data is available. All other monitors will operate normally.
We have identified the underlying issue and are working on a fix. It is important to note that no data has been lost, and it will be backfilled and available once the services are operational again.
We have deployed a fix and we are monitoring the results. We will provide another update once the issue is fully resolved.
This incident has been resolved.
We have identified the underlying issue and are working on a fix. It is important to note that no data has been lost, and notifications will be caught up once the service is operational again.
This incident has been resolved.
We are investigating increased latency across multiple products. As a result of this issue, some users may see delays in data across the platform.
We are currently investigating an issue affecting multiple products. During this time, some customers may experience: Delayed or missing alert notifications for RUM, Synthetics, APM, and Log-based monitors Errors or delays when querying Logs, including in Log Explorer and Live Tail Delays or gaps in Event Management data Errors or failures in Error Tracking, Cloud SIEM, and Session Replay Degraded performance across Software Delivery products, including CI Visibility Broader platform degradation and intermittent request failures across services
We are continuing to investigate an issue affecting multiple products. During this time, some customers may experience: Delayed or missing alert notifications for RUM, Synthetics, APM, and Log-based monitors Errors or delays when querying Logs, including in Log Explorer and Live Tail Delays or gaps in Event Management data Errors or failures in Error Tracking, Cloud SIEM, and Session Replay Degraded performance across Software Delivery products, including CI Visibility Broader platform degradation and intermittent request failures across services
A fix has been implemented and we are monitoring the results.
We are continuing to monitor for any further issues.
This incident has been resolved.
We are investigating delays in Monitors Notifications, which began at 19:35 UTC
We have identified the underlying issue and are working on a fix. Delayed Monitor Notifications are limited to only distribution monitors.
We have deployed a fix and we are monitoring the results. We will provide another update once the issue is fully resolved.
This incident has been resolved.
We are investigating an issue submitting Azure metrics.
The issue has been identified and a fix is being implemented.
We are continuing to work on a fix for this issue.
A fix has been implemented and we are monitoring the results.
This incident has been resolved and Azure metrics are reporting as expected.
We are investigating degraded performance with the Web Application.
We have identified the underlying issue and are working on a fix.
We have deployed a fix and we are monitoring the results. We will provide another update once the issue is fully resolved.
This incident has been resolved.
We are currently investigating an issue where customers in our EU region might experience issues searching, creating, and updating their monitor and SLO configurations through the web application or API. As a result, dashboard widgets based on monitors or SLOs are also affected. Monitor alerts and SLO evaluations are not affected.
We are continuing to investigate the issue where customers in our EU region might experience issues searching, creating, and updating their monitor and SLO configurations through the web application or API. As a result, dashboard widgets based on monitors or SLOs are also affected. Monitor alerts and SLO evaluations are not affected.
We have identified the underlying issue and are working on a fix.
We have deployed a fix and we are monitoring the results. We will provide another update once the service is fully operational.
This incident has been resolved.
We are investigating an issue causing some metric monitors to intermittently report "No Data" status. This primarily affects monitors based on distribution metrics. Monitor alerting may be unreliable during this time. We are actively investigating and will provide updates as available.
We are continuing to investigate the issue and will provide updates as available.
A fix has been implemented and we are monitoring the results.
This incident has been resolved.
We are investigating loading issues on our web application. As a result, some users might be getting errors when loading the web application.
We are continuing to investigate this issue.
A fix has been implemented and we are monitoring the results.
This incident has been resolved.
We are currently investigating this issue.
We are continuing to investigate this issue.
The issue has been identified and a fix is being implemented.
We are continuing to work on a fix for this issue.
We are observing recovery for the vast majority of monitors and are continuing to work on a full fix for this issue.
We have observed full recovery of monitors and will continue to monitor
This incident has been resolved.
We are investigating loading issues on our web application. As a result, some users might be getting errors when loading the web application. Please note that data processing and alerts are not affected by this incident.
We are continuing to investigate this issue.
A fix has been implemented and we are monitoring the results.
This incident has been resolved.
We are investigating delays in RUM-based Monitors Notifications, which began at 11:30am UTC.
This incident has been resolved. Notification delays were only affecting our internal monitoring and were due to the ongoing Cloudflare incident: https://www.cloudflarestatus.com/incidents/8gmgl950y3h7/.
We are investigating user login issues with the web application. Please note that data processing and alerts are not affected by this incident.
A fix has been implemented and we are monitoring the results.
This incident has been resolved.
We are actively investigating elevated error rates for Resource Catalog views As a result of this issue, some users may see errors with Resource Catalog on the web application
A fix has been implemented and we are monitoring the results.
This incident has been resolved.
We are investigating increased latency processing Metrics. As a result of this issue, some users may see delays or gaps for metrics on graphs. To prevent spurious alerts, we have temporarily disabled monitors based on this data.
This incident has been resolved.
We are investigating user login issues with the web application via Google SSO. Please note that data processing and alerts are not affected by this incident.
We are seeing recovery in Google SSO logins. We are continuing to monitor for issues.
This incident has been resolved.