← Back to events
ResolvedTechOutage599 minResolved

High number of 5XX on the Machines API and dashboard

  • Machines: 2 events in the last 90 days

What happened

Jul 20 , 17:09 UTC Resolved - This incident has been resolved. Jul 20 , 13:34 UTC Update - We are still working on fixing degraded Managed Postgres clusters. Jul 20 , 09:14 UTC Update - Some Managed Postgres v1 clusters are degraded. We are working on fixing them. Managed Postgres v2 is unaffected. Jul 20 , 08:02 UTC Monitoring - A fix has been implemented and we are monitoring the results. Jul 20 , 07:52 UTC Update…

Summary assembled by rule from the sources below

Why it's spreading

Timeline

  1. First appeared on Fly.io StatusFly.io Status
    1. InvestigatingExisting machines are unaffected. We are investigating the issue.
    2. UpdateWe are continuing to investigate this issue.
    3. IdentifiedWe've identified an internal service providing authentication to our Machines API has failed, our team is currently looking at our options for restoring this service. Existing Machines/Apps will continue to run as normal. Thank you for your patience.
    4. UpdateWe've identified an internal service providing authentication to our Machines API has failed, our team is currently looking at our options for restoring this service. Existing Machines/Apps will continue to run as normal. Thank you for your patience.
    5. MonitoringA fix has been implemented and we are monitoring the results.
    6. UpdateSome Managed Postgres v1 clusters are degraded. We are working on fixing them. Managed Postgres v2 is unaffected.
    7. UpdateWe are still working on fixing degraded Managed Postgres clusters.
    8. ResolvedThis incident has been resolved.

Sources