A token service issue causing elevated internal request volume impacted Dynatrace authentication services for approximately 20 hours, beginning September 1, 2026. This cascading failure affected a broad range of platform capabilities across all cloud providers (AWS, GCP, Azure), including Anomaly Detectors, BizEvents, Extracted Metrics, Problem Manager, Real User Monitoring, Smartscape, Synthetics, Workflows, and the Classic UI, leaving customers unable to access their environments or experiencing significant data processing delays. Mitigations were progressively rolled out across all cloud providers, with full deployment recovery confirmed by September 2, after which teams continued monitoring for backlog clearance and assessing potential data loss.
Trusted by 1,000+ teams
Stop finding out about outages from your users. Monitor 6,320+ cloud services and get alerted the second something breaks.
We are investigating an issue affecting authentication services. As a result, data processing and extraction services may experience errors or delayed results. Data ingestion continues to operate, but processing of ingested data may be impacted.
Our teams are actively investigating the issue and working to restore normal service. Please monitor Dynatrace Status for further updates. We apologize for the inconvenience.
We are continuing to investigate the issue affecting authentication services.
Customers may experience problems accessing Dynatrace services.
We are actively working to mitigate the issue and restore normal service. We will provide our next update within 30 minutes or sooner if significant new information becomes available.
We apologize for the inconvenience.
We have identified the issue affecting authentication services and are actively implementing mitigation measures.
Customers may experience problems accessing Dynatrace services while mitigation activities are in progress.
We are working to restore normal service as quickly as possible and will provide our next update within 30 minutes or sooner if significant new information becomes available.
We apologize for the inconvenience.
We continue to work on mitigation measures for the issue affecting authentication services.
Customers may experience problems accessing Dynatrace services.
We will provide our next update within 30 minutes or sooner if significant new information becomes available.
We apologize for the inconvenience.
We continue to work on mitigation measures for the issue affecting authentication services and are observing slight improvements.
We will provide our next update within 30 minutes or sooner if significant new information becomes available.
We apologize for the inconvenience.
We continue to work on mitigation measures for the issue affecting authentication services.
We are observing continued improvements. Access to the Dynatrace platform is improving. We have implemented additional mitigation measures to reduce the impact, while work continues to restore all affected services.
We will provide our next update within 30 minutes or sooner if significant new information becomes available.
We continue to make progress in mitigating the issue affecting authentication services. Additional mitigation measures have been implemented to reduce customer impact while our teams continue working toward full service restoration.
We will provide the next update within 30 minutes, or sooner if significant new information becomes available.
We apologize for the inconvenience and appreciate your patience as we continue working toward full mitigation.
We continue to make progress in mitigating the issue affecting authentication services. Additional mitigation measures have been implemented to reduce customer impact while our teams continue working toward full-service restoration.
We will provide the next update within 30 minutes, or sooner if significant new information becomes available.
We apologize for the inconvenience and appreciate your patience as we continue working toward full mitigation.
We continue to make progress in mitigating the issue affecting authentication services. We have identified that the token service supporting several Dynatrace capabilities remains impacted and are actively implementing additional mitigation measures while work continues toward full service restoration.
The following services may experience degraded functionality or interruptions while mitigation activities are in progress:
- Authentication Services
- Anomaly Detectors
- BizEvents
- Extracted Metrics
- Problem Manager
- Real User Monitoring
- Smartscape
- Synthetics
- Workflows
In addition, customers may begin to experience issues accessing the Dynatrace Classic UI. If you are unable to access your environment, this is related to this ongoing incident.
We will share progress and resolution updates on this Status Page as they become available. To reduce duplicate effort and keep our teams focused on remediation, please follow this incident here rather than opening a Support case for the impacted behavior.
Our teams remain actively engaged and are working to restore all affected services as quickly as possible. We will provide the next update within 30 minutes, or sooner if significant new information becomes available.
We apologize for the inconvenience and appreciate your patience as we continue working toward full mitigation.
We continue to make progress in mitigating the issue affecting Dynatrace Platform. We have identified that the token service supporting several Dynatrace capabilities remains impacted and are actively implementing additional mitigation measures while work continues toward full service restoration.
The following services may experience degraded functionality or interruptions while mitigation activities are in progress:
- Authentication Services
- Anomaly Detectors (Alerting)
- BizEvents
- Extracted Metrics
- Problem Manager
- Real User Monitoring
- Smartscape
- Synthetics
- Workflows
In addition, customers may begin to experience issues accessing the Dynatrace Classic UI. If you are unable to access your environment, this is related to this ongoing incident.
We will share progress and resolution updates on this Status Page as they become available. To reduce duplicate effort and keep our teams focused on remediation, please follow this incident here rather than opening a Support case for the impacted behavior.
Our teams remain actively engaged and are working to restore all affected services as quickly as possible. We will provide the next update within 30 minutes, or sooner if significant new information becomes available.
We apologize for the inconvenience and appreciate your patience as we continue working toward full mitigation.
We continue to make progress mitigating the issue affecting the Dynatrace Platform. The token service supporting several Dynatrace capabilities remains impacted, and we are implementing additional mitigation measures to further reduce customer impact while work continues toward full-service restoration.
The following services may experience degraded functionality or interruptions while mitigation activities are in progress:
- Authentication Services
- Anomaly Detectors (Alerting)
- BizEvents
- Extracted Metrics
- Problem Manager
- Real User Monitoring
- Smartscape
- Synthetics
- Workflows
In addition, customers may begin to experience issues accessing the Dynatrace Classic UI. If you are unable to access your environment, this is related to this ongoing incident.
We will share progress and resolution updates on this Status Page as they become available. To reduce duplicate effort and keep our teams focused on remediation, please follow this incident here rather than opening a Support case for the impacted behavior.
Our teams remain actively engaged and are working to restore all affected services as quickly as possible. We will provide the next update within 30 minutes, or sooner if significant new information becomes available.
We apologize for the inconvenience and appreciate your patience as we continue working toward full mitigation.
We are continuing to implement mitigation measures and are seeing gradual progress. Our investigation has helped narrow the issue to increase of requests affecting our authentication services coming from internal resources. The team is actively working to restore service behaviour to normal operating levels.
However, we are still processing the accumulated backlog, and customers may continue to experience degraded functionality until recovery is fully completed.
We will provide the next update within 30 minutes, or sooner if significant new information becomes available.
We apologize for the inconvenience and appreciate your patience as we continue working toward full mitigation.
We continue to implement mitigation measures, and early results are showing promising signs of improvement. Our investigation has narrowed the issue to elevated internal request volume impacting authentication services. Teams remain focused on restoring normal service behavior as quickly and safely as possible.
Recovery is still progressing as we work through the accumulated backlog, and some customers may continue to experience degraded functionality until mitigation is fully complete.
We will provide the next update within 30 minutes, or sooner if significant new information becomes available.
We apologize for the disruption and appreciate your continued patience as we work toward full mitigation.
We continue to implement mitigation measures, and some customers should begin seeing improvements in their environments. Our investigation has narrowed the issue to elevated internal request volume impacting authentication services. Teams remain fully engaged and are making steady progress toward restoring normal service behavior as quickly as possible.
Until full mitigation is reached, some customers may continue to experience degraded functionality and data ingestion lags, while recovery work continues.
We will provide the next update within 30 minutes, or sooner if significant new information becomes available.
We apologize for the disruption and appreciate your continued patience as we work toward full mitigation.
We continue to see steady improvement as mitigation efforts progress, and many customers should already be experiencing better platform performance.
While we are making significant progress toward full service restoration, some customers may still encounter degraded functionality and delays in data ingestion as recovery activities continue. We understand the impact this may have on your operations and remain fully focused on restoring all services as quickly as possible. Full recovery may take several more hours.
We will provide the next update within 60 minutes, or sooner if there are any significant developments.
We sincerely apologize for the disruption and greatly appreciate your patience and understanding as we work toward full restoration.
Mitigation efforts continue to show steady improvement, and many customers are already seeing better platform performance. While recovery work continues, some customers may still experience degraded functionality or data ingestion delays. We are closely monitoring backlog processing and overall system health to help ensure customer environments return to a healthy state following this incident.
We understand the importance of restoring normal service as quickly as possible and remain fully engaged until recovery is complete. Based on current progress, full recovery may still require several more hours, and we will continue working to minimize any remaining customer impact. The latest mitigation efforts require propagation. Full recovery may take several more hours.
We will share the next update within 30 minutes, or sooner if there is a material change.
We sincerely apologize for the disruption and appreciate your continued patience and trust as we complete recovery efforts.
Mitigation efforts continue to show progress, with GCP and Azure environments showing significant improvement. Some customers may still experience degraded functionality or data ingestion delays while recovery work continues.
Full recovery may take several more hours as mitigation propagates. Teams remain fully engaged and are monitoring system health and backlog processing.
We will provide the next update within 30 minutes, or sooner if there are any significant developments.
We apologize for the disruption and appreciate your continued patience.
Mitigation efforts continue to progress, with GCP and Azure environments demonstrating significant improvement. While recovery activities remain ongoing, our teams are focused on the restoration of AWS environments to a healthy operational state. Some customers may continue to experience degraded functionality or data ingestion delays during this period.
Full recovery may require several additional hours as mitigation measures to propagate across affected environments. Our teams remain fully engaged, with continued monitoring of system health and backlog processing.
We will provide the next update within 30 minutes, or sooner if there are any significant developments.
Mitigation continues to progress, with GCP and Azure showing significant improvement and we are also seeing a large number of AWS deployments are now in a healthy state. Some customers may still experience degraded functionality or data ingestion delays while recovery completes.
Full recovery may take several more hours. Our teams remain fully engaged and are monitoring system health and backlog processing.
We will provide the next update within 30 minutes, or sooner if there are significant developments.
Recovery efforts continue to progress across all cloud providers. Approximately 85% of AWS deployments have now returned to a healthy state, and we continue to see stability across GCP and Azure environments.
While service stability has improved for the majority of affected customers, recovery may take several hours, and some customers may still experience degraded functionality or delays in data processing and ingestion as recovery activities continue and remaining backlogs are cleared.
Our teams remain fully engaged and are closely monitoring system health and recovery progress. We will provide the next update within 30 minutes, or sooner if there are significant developments.
Mitigations have now been implemented across all cloud providers with only a small number of remaining AWS deployments to be restored.
Customers may experience degraded functionality or delays in data processing and ingestion, which may take several hours to resolve, as recovery activities continue and remaining backlogs are cleared. Some customers may experience a significant increase in alerts as services restore.
Our teams remain fully engaged and are closely monitoring system health and recovery progress. We will provide the next update within 30 minutes, or sooner if there are significant developments.
Mitigations have now been implemented across all cloud providers and all deployments are fully recovered. We are continuing to actively monitor the health of the restored deployments and currently observing stable performance.
Customers may experience degraded functionality or delays in data processing and ingestion, which may take several hours to resolve, as recovery activities continue and remaining backlogs are cleared. Some customers may experience a significant increase in alerts as services restore.
We will provide the next update within 30 minutes, or sooner if there are significant developments.
Mitigations have been successfully implemented and our team is actively monitoring the health of the restored deployments. We are continuing to observe stable performance.
Customers may experience degraded functionality or delays in data processing and ingestion, which may take several hours to resolve, as recovery activities continue and remaining backlogs are cleared. Some customers may experience a significant increase in alerts as services restore.
We will provide the next update within 60 minutes, or sooner if there are significant developments.
Our team is actively monitoring the restored deployments and we are continuing to observe stable performance.
Some customers may experience degraded functionality, delays in data processing and ingestion, or a temporary increase in alerts. These issues are expected to resolve as backlogs are cleared.
It may take several hours for some backlogs to clear and associated tenants to resume normal operations. We will provide the next update within 60 minutes, or sooner if there are significant developments.
We are continuing to observe stable performance and are maintaining active monitoring to ensure all deployments have an extended period of stable performance.
Some customers may experience degraded functionality, delays in data processing and ingestion, or a temporary increase in alerts. These issues are expected to resolve as backlogs are cleared.
It may take several hours for some backlogs to clear and associated tenants to resume normal operations. We will provide the next update within 60 minutes, or sooner if there are significant developments.
Performance across all restored deployments has remained stable as we continue active monitoring.
The temporary degradation of functionality, delays in data processing and ingestion, or increase in alerts that some customers have experienced should be subsiding. These issues are expected to resolve as backlogs are cleared.
It may take several hours for some backlogs to clear and associated tenants to resume normal operations. We will provide the next update within 60 minutes, or sooner if there are significant developments.
Performance across all restored deployments has remained stable as we continue active monitoring and providing hourly updates.
The temporary degradation of functionality, delays in data processing and ingestion, or increase in alerts that some customers have experienced should be subsiding. These issues are expected to resolve as backlogs are cleared and we are investigating any potential data loss.
Performance across all deployments remains stable, and we continue to actively monitor service health.
Recovery activities are ongoing. Some customers may still experience delays in data processing and ingestion, as well as a temporary increase in alerts, while systems return to normal operation. We are also assessing the impact of the incident on data processing completeness.
We will provide the next update within 60 minutes, or sooner if significant new information becomes available.
With IsDown, you can monitor all your critical services' official status pages from one centralized dashboard and receive instant alerts the moment an outage is detected. Say goodbye to constantly checking multiple sites for updates and stay ahead of outages with IsDown.
Start free trialNo credit card required · Cancel anytime · 6320 services available
Integrations with