Fly.io's Machines API and Dashboard experienced a 10-hour major outage caused by a failure in an internal authentication service. While existing machines and apps continued running normally, users were unable to interact with the API or Dashboard, and some Managed Postgres v1 clusters became degraded as a secondary impact. The incident was resolved after a fix was implemented and monitored, with full resolution confirmed later that day.
Trusted by 1,000+ teams
Stop finding out about outages from your users. Monitor 6,320+ cloud services and get alerted the second something breaks.
This incident has been resolved.
We are still working on fixing degraded Managed Postgres clusters.
Some Managed Postgres v1 clusters are degraded. We are working on fixing them. Managed Postgres v2 is unaffected.
A fix has been implemented and we are monitoring the results.
We've identified an internal service providing authentication to our Machines API has failed, our team is currently looking at our options for restoring this service. Existing Machines/Apps will continue to run as normal. Thank you for your patience.
We've identified an internal service providing authentication to our Machines API has failed, our team is currently looking at our options for restoring this service. Existing Machines/Apps will continue to run as normal. Thank you for your patience.
We are continuing to investigate this issue.
Existing machines are unaffected. We are investigating the issue.
With IsDown, you can monitor all your critical services' official status pages from one centralized dashboard and receive instant alerts the moment an outage is detected. Say goodbye to constantly checking multiple sites for updates and stay ahead of outages with IsDown.
Start free trialNo credit card required · Cancel anytime · 6320 services available
Integrations with