Fly.io Outage History

50 incidents reported. Data sourced from the official Fly.io status page.

50
Total Incidents
16
Major/Critical
27
Minor
50
Resolved

September 2026

Depot builder failures

minor
Sep 15, 04:08 AMSep 15, 05:34 AMresolved
Sep 15, 05:34 AM
resolvedThis incident has been resolved.
Sep 15, 04:51 AM
monitoringA fix has been implemented and we are monitoring builds. Standard flyctl builds should be working again for customers in all regions.
Sep 15, 04:45 AM
identifiedWe have identified an issue with deploying via Depot for users connecting through our SYD and JNB regions. Affected customers in these regions can deploy successfully using the --depot=false or --buil...
+1 more updates

Network issues in US West Coast

minor
Sep 12, 09:22 PMSep 12, 10:12 PMresolved
Sep 12, 10:12 PM
resolvedThis incident has been resolved.
Sep 12, 09:57 PM
monitoringPrivate networking between Fly Machines is resolved, and most outbound connections are healthy. We're continuing to monitor the network, and some issues will still be expected from clients physically ...
Sep 12, 09:22 PM
investigatingWe are investigating upstream network issues from US West Coast (SJC, LAX). Apps hosted in US West regions may experience higher latency or packet loss, and requests from clients physically located in...

Sprites API Partial Outage

minor
Sep 2, 11:03 PMSep 3, 12:41 AMresolved
Sep 3, 12:41 AM
resolvedThis incident has been resolved.
Sep 2, 11:33 PM
monitoringError rates have decreased. We are continuing to monitor the API health.
Sep 2, 11:03 PM
investigatingWe're aware of a problem affecting a subset of Sprites users. We are investigating the source of the issue.

Upstream network issues

minor
Sep 2, 02:18 PMSep 2, 03:57 PMresolved
Sep 2, 03:57 PM
resolvedThis incident has been resolved.
Sep 2, 03:33 PM
monitoringA fix has been implemented and we are monitoring the results.
Sep 2, 02:56 PM
identifiedWe're updating the affected region list to also include SJC since this seems to be a wider upstream issue in US West Coast.
+2 more updates

API background job queue failure

major
Sep 2, 07:21 AMSep 2, 07:44 AMresolved
Sep 2, 07:44 AM
resolvedThis incident has been resolved.
Sep 2, 07:31 AM
monitoringA fix has been implemented and we are monitoring the results.
Sep 2, 07:21 AM
investigatingWe are investigating an issue with the background job runner for our API. Actions that require a background job, such as creating apps, assigning IP addresses, or creating/renewing certificates, may f...

August 2026

HTTP/2 traffic disruptions

none
Aug 31, 03:30 PMAug 31, 03:30 PMresolved
Aug 31, 03:42 PM
resolvedA configuration update caused temporary failures for incoming HTTP/2 traffic for Fly Machines located on a subset of hosts for a few minutes. This incident has since been resolved. Managed Postgres d...

Packet loss in ORD

minor
Aug 31, 06:26 AMAug 31, 08:10 AMresolved
Aug 31, 08:10 AM
resolvedThis incident has been resolved.
Aug 31, 07:39 AM
monitoringPacket loss in ORD is improving and impacted services are recovering; we’re continuing to monitor for intermittent issues
Aug 31, 06:26 AM
investigatingDue to an upstream provider, we are seeing ~50% packet loss on a subset of hosts in ORD. Some MPG clusters in ORD are slow to replicate as a result.

Sprite deletion jobs failing

none
Aug 30, 10:45 PMAug 30, 10:45 PMresolved
Aug 30, 10:45 PM
resolvedWe saw Sprite deletion jobs failing between 21:18 and 22:05 UTC. This issue has been resolved.

Networking Issues in GRU

minor
Aug 28, 09:56 PMAug 28, 10:36 PMresolved
Aug 28, 10:36 PM
resolvedThis incident has been resolved.
Aug 28, 10:36 PM
identifiedOur upstream provider has implemented a fix. Network performance in GRU has normalized.
Aug 28, 10:12 PM
identifiedWe are seeing a recurrance in networking issues in GRU. Some apps in the region may experience increased latency or packet loss. We are working with our upstream networking provider to resolve.
+2 more updates

Increased packet loss

minor
Aug 28, 08:09 AMAug 28, 10:22 AMresolved
Aug 28, 10:22 AM
resolvedThis incident has been resolved.
Aug 28, 08:09 AM
investigatingWe are currently investigating this issue.

WireGuard gateway issues

minor
Aug 26, 06:14 PMAug 26, 06:43 PMresolved
Aug 26, 06:43 PM
resolvedThis incident has been resolved.
Aug 26, 06:33 PM
monitoringOur testing and monitoring indicates gateways should be back to normal; if you are still having problem using `flyctl ssh console`, try restarting the `flyctl` agent by `flyctl agent restart`.
Aug 26, 06:27 PM
monitoringA fix has been implemented and we are monitoring the results.
+1 more updates

Metrics in some regions are lagging behind

minor
Aug 24, 10:23 AMAug 24, 01:24 PMresolved
Aug 24, 01:24 PM
resolvedThis is now resolved
Aug 24, 12:47 PM
monitoringAll hosts have caught up with metrics and we're monitoring the situation
Aug 24, 10:23 AM
investigatingWe are currently experiencing some metrics lag on servers in some regions. We are provisioning more metric processing instances to accommodate the backlog and catch up.

Network Issues in LAX Region

major
Aug 23, 01:28 AMAug 23, 02:10 AMresolved
Aug 23, 02:10 AM
resolvedThis incident has been resolved.
Aug 23, 02:04 AM
monitoringUpstream networking issues have resolved.
Aug 23, 01:28 AM
investigatingWe are investigating network issues in the Los Angeles region. Apps may experience higher latency or be unreachable at this time.

Temporary DNS resolution failure

minor
Aug 20, 07:30 PMAug 20, 07:30 PMresolved
Aug 20, 08:07 PM
resolvedA BGP configuration error caused our Anycast DNS to route to some nodes without the proper DNS infrastructure. The issue was temporary and was resolved as soon as we removed that node from BGP.

Oauth/Macaroon Errors from flyctl

none
Aug 20, 01:54 PMAug 20, 02:16 PMresolved
Aug 20, 02:16 PM
resolvedThis incident has been resolved.
Aug 20, 02:05 PM
monitoringA fix has been deployed and this error should no longer be occurring. We're monitoring to ensure full recovery.
Aug 20, 01:54 PM
identifiedWe have identified an issue causing authentication errors for some operations from `flyctl`. These operations are failing with an error like: `This endpoint no longer accepts legacy OAuth tokens (star...

MPG (v1) partially down in ORD

major
Aug 20, 07:25 AMAug 20, 07:57 AMresolved
Aug 20, 07:57 AM
resolvedThis incident has been resolved.
Aug 20, 07:32 AM
monitoringA fix has been implemented and we are monitoring the results.
Aug 20, 07:26 AM
investigatingWe are continuing to investigate this issue.
+1 more updates

6PN Networking issue in YYZ

none
Aug 18, 08:00 PMAug 19, 12:00 AMresolved
Aug 19, 01:15 AM
resolved6PN networking issues between some machines in YYZ during a rollout which was rolled back once we noticed errors. During this time some machines were unable to talk to internal resources like other DB...

No capacity in ARN

none
Aug 17, 01:19 PMAug 17, 03:41 PMresolved
Aug 17, 03:41 PM
resolvedThe capacity issue in the ARN region has been resolved.
Aug 17, 01:19 PM
investigatingNew machines may fail to create in ARN because we lack capacity.

Secrets service outage

major
Aug 14, 07:30 PMAug 14, 08:33 PMresolved
Aug 14, 08:33 PM
resolvedThis incident has been resolved.
Aug 14, 08:08 PM
monitoringWe have failed over the secrets database to a replica, and the Machines API appears healthy now. We are monitoring for any further issues.
Aug 14, 07:30 PM
identifiedWe are working to recover our secrets service after a failed deployment. Apps continue to run, but it is not possible to create new apps or update secrets at this time.

IPv6 Networking Issues

minor
Aug 13, 05:45 PMAug 13, 07:41 PMresolved
Aug 13, 07:41 PM
resolvedThis incident has been resolved.
Aug 13, 06:37 PM
monitoringWe are continuing to monitor for any further issues.
Aug 13, 06:36 PM
monitoringA fix has been implemented and we are monitoring the results.
+3 more updates

Increased app-not-found errors

major
Aug 9, 02:50 AMAug 9, 07:10 AMresolved
Aug 9, 07:10 AM
resolvedThis incident has been resolved.
Aug 9, 06:42 AM
monitoringA fix has been implemented and we are monitoring the results
Aug 9, 06:03 AM
identifiedWe’ve deployed an additional mitigation to further reduce Corrosion retry pressure and are seeing improvement; we’re continuing to monitor while remaining affected nodes catch up.
+3 more updates

MPG IAD data plane degraded for new clusters

minor
Aug 5, 08:26 PMAug 5, 08:41 PMresolved
Aug 5, 08:41 PM
resolvedEtcd is stable. Services are back to normal.
Aug 5, 08:26 PM
identifiedHigh CPU pressure on a shared etcd instance is causing lags on MPG creation in the IAD region

MPG creation is failing in GRU due to lack of capacity

major
Aug 4, 07:03 PMAug 5, 03:22 AMresolved
Aug 5, 03:22 AM
resolvedThis incident has been resolved.
Aug 4, 08:59 PM
monitoringWe tweaked hosts to allow for more machine allocation. We'll be monitoring the region over the next hours.
Aug 4, 07:03 PM
identifiedNew MPG clusters may fail to create in GRU because we lack capacity.

Certificate issuance delays

minor
Aug 4, 02:53 PMAug 4, 11:58 PMresolved
Aug 4, 11:58 PM
resolvedThis incident has been resolved.
Aug 4, 08:36 PM
monitoringA fix has been implemented and we are monitoring the results.
Aug 4, 06:26 PM
identifiedWe have the size of TLS certificate issuance backlog under control, but are still seeing some remaining issues and are currently working to clean up the edge cases.
+2 more updates

Inbound connection failure to Fly Apps

critical
Aug 3, 03:16 PMAug 3, 06:50 PMresolved
Aug 3, 03:16 PM
resolvedA BGP misconfiguration while provisioning new edge capacity caused most traffic from Europe endpoints to be dropped, between 14:50 UTC and 15:02 UTC. The misconfiguration has been fixed and we are imp...

Managed Postgres v2 control plane issues in iad

minor
Aug 1, 05:01 PMAug 1, 06:30 PMresolved
Aug 1, 06:30 PM
resolvedThis incident has been resolved.
Aug 1, 05:26 PM
monitoringA fix has been implemented and we are monitoring the results.
Aug 1, 05:01 PM
investigatingWe are investigating an issue with the control plane for Managed Postgres v2 in the IAD region. Creating new v2 clusters in the IAD region may fail at this time. Existing clusters continue to run, but...

July 2026

Capacity issues in CDG

none
Jul 31, 04:44 PMJul 31, 08:11 PMresolved
Jul 31, 08:11 PM
resolvedThis incident has been resolved.
Jul 31, 07:02 PM
monitoringA fix has been implemented and we are monitoring the results.
Jul 31, 04:44 PM
investigatingCreating machines in CDG region may fail at this time with an "no capacity available in cdg" message. Existing apps continue to run.

Outbound email issues

minor
Jul 31, 01:34 PMJul 31, 02:51 PMresolved
Jul 31, 02:51 PM
resolvedThis incident has been resolved.
Jul 31, 02:38 PM
monitoringA fix has been implemented and we are monitoring the results.
Jul 31, 01:34 PM
investigatingOur dashboard is failing to send outbound email. Emails such as new account verification or password reset may fail to send at this time.

Increased API latency

minor
Jul 29, 12:41 PMJul 29, 01:26 PMresolved
Jul 29, 01:26 PM
resolvedThis incident has been resolved.
Jul 29, 12:41 PM
investigatingWe are investigating some database issues causing high latency on some API endpoints and dashboard operations. You may experience intermittent "503 service unavailable" errors at this time. Currently ...

High number of 5XX on the Machines API and dashboard

critical
Jul 20, 07:10 AMJul 20, 05:09 PMresolved
Jul 20, 05:09 PM
resolvedThis incident has been resolved.
Jul 20, 01:34 PM
monitoringWe are still working on fixing degraded Managed Postgres clusters.
Jul 20, 09:14 AM
monitoringSome Managed Postgres v1 clusters are degraded. We are working on fixing them. Managed Postgres v2 is unaffected.
+5 more updates

Egress IPv6 issues in BOM

minor
Jul 19, 03:48 AMJul 19, 04:52 AMresolved
Jul 19, 04:52 AM
resolvedThis incident has been resolved.
Jul 19, 03:48 AM
identifiedWe have identified an upstream issue that is preventing egress IPv6 addresses in BOM from reaching parts of the internet, and we're currently working with an upstream provider to resolve this issue. N...

Brief Flycast / MPG interruption in YYZ

none
Jul 17, 06:30 PMJul 17, 06:30 PMresolved
Jul 17, 07:13 PM
resolvedA bad deployment momentarily caused issues with Flycast connectivity, and, by extension, MPG, in our YYZ region. The deployment was immediately rolled back and connectivity was restored shortly after.

Edge proxy issues

minor
Jul 16, 11:34 AMJul 16, 01:44 PMresolved
Jul 16, 01:44 PM
resolvedThis incident has been resolved.
Jul 16, 12:27 PM
monitoringA fix has been implemented and we are monitoring the results.
Jul 16, 11:34 AM
investigatingWe are investigating increased connection latency and "connection reset" errors from our edge proxy. Apps continue to run, but requests may experience increased connection latency or fail at this time...

App creation failing

minor
Jul 16, 11:54 AMJul 16, 12:17 PMresolved
Jul 16, 12:17 PM
resolvedThis incident has been resolved.
Jul 16, 11:54 AM
identifiedAn issue with our Machines API is causing app creations to fail in some cases. We are working on a fix.

Partial outage in SJC

major
Jul 14, 02:51 PMJul 14, 04:59 PMresolved
Jul 14, 04:59 PM
resolvedThis incident has been resolved.
Jul 14, 03:53 PM
monitoringA fix has been implemented and we are monitoring the results. Apps should be reachable at this point in time.
Jul 14, 02:51 PM
identifiedA subset of hosts in SJC are currently offline. Some apps may be unreachable at this time.

Some DFW hosts offline

minor
Jul 13, 10:25 PMJul 13, 10:58 PMresolved
Jul 13, 10:58 PM
resolvedThis incident has been resolved.
Jul 13, 10:36 PM
monitoringA fix has been implemented and we are monitoring the results.
Jul 13, 10:25 PM
investigatingA subset of hosts in DFW are currently offline, and we're investigating the issue.

Delays starting Depot Builders in IAD

minor
Jul 9, 03:21 PMJul 9, 04:18 PMresolved
Jul 9, 04:18 PM
resolvedThis incident has been resolved.
Jul 9, 03:43 PM
monitoringA fix has been implemented and we are seeing improvements in builder performance, latency, and error rates. We are continuing to monitor for a full recovery. Customers still seeing issues can trigger...
Jul 9, 03:21 PM
identifiedWe have identified an issue causing delays or failures when starting depoting builders located in the IAD region. Customers with builders in IAD may see delays or timeouts starting builds during `fly ...

Registry performance issues

minor
Jul 9, 03:33 PMJul 9, 04:04 PMresolved
Jul 9, 04:04 PM
resolvedThis incident has been resolved.
Jul 9, 03:55 PM
monitoringA fix has been implemented and we are monitoring the results.
Jul 9, 03:33 PM
identifiedThe fly.io registry is currently experiencing capacity constraints that reduced performance and may lead to temporary high latency or failed pushes. We are currently working to add capacity and restor...

Partial Outage in ORD

major
Jul 3, 05:00 PMJul 3, 06:11 PMresolved
Jul 3, 06:11 PM
resolvedThis incident has been resolved.
Jul 3, 05:49 PM
monitoringA fix has been implemented and we are monitoring the results.
Jul 3, 05:02 PM
identifiedWe've identified the issue as a networking hardware failure impacting a subset of hosts at one of our Upstream providers in ORD. We are working with our provider to restore connectivity.
+1 more updates

Partial outage in ORD

major
Jul 3, 12:11 AMJul 3, 05:59 AMresolved
Jul 3, 05:59 AM
resolvedThis incident has been resolved.
Jul 3, 04:41 AM
monitoringCustomer workloads are now starting and we're monitoring the affected hosts. Affected Managed Postgres instances will be investigated.
Jul 3, 04:21 AM
identifiedPower restoration is ongoing and we're making sure the hosts are healthy before starting customer workloads to avoid issues. Customer impact remains and updates to come.
+4 more updates

Errors issuing new SSL certificates

critical
Jul 2, 09:38 PMJul 2, 11:18 PMresolved
Jul 2, 11:18 PM
resolvedThis incident has been resolved.
Jul 2, 10:50 PM
monitoringA fix has been implemented upstream and certificates are being issued successfully. We will continue to monitor.
Jul 2, 10:26 PM
identifiedThe issue has been identified and we are awaiting a fix.
+1 more updates

Static Egress IPv6 issues in NRT

minor
Jul 1, 01:06 PMJul 1, 01:57 PMresolved
Jul 1, 01:57 PM
resolvedThis incident has been resolved.
Jul 1, 01:44 PM
monitoringA fix has been implemented and we are monitoring the results.
Jul 1, 01:06 PM
investigatingWe are investigating issues with static egress IPv6 addresses in NRT region. Apps using static egress IPs may experience connectivity failures to some destinations.

Elevated API Errors

major
Jul 1, 06:14 AMJul 1, 07:50 AMresolved
Jul 1, 07:50 AM
resolvedThis incident has been resolved.
Jul 1, 07:26 AM
monitoringBackground jobs have caught up and the API is fully operational. We are continuing to monitor service health.
Jul 1, 07:05 AM
identifiedA fix has been put in place, and we are no longer seeing elevated API errors. Some dashboard actions will be delayed while background processing catches up.
+2 more updates

June 2026

Delayed Metrics

major
Jun 29, 01:03 AMJun 30, 10:19 PMresolved
Jun 30, 10:19 PM
resolvedThis incident has been resolved.
Jun 30, 07:58 PM
identifiedAlmost all metrics have caught up aside from a small handful in `sin` and `syd`. We're continuing to monitor and expect these to complete in the next few hours.
Jun 30, 04:50 PM
identifiedBacklogged metrics are still being processed. We're bringing extra processing capacity online to speed up the process.
+9 more updates

Egress IP issues in SIN and NRT

minor
Jun 30, 01:45 PMJun 30, 02:03 PMresolved
Jun 30, 02:03 PM
resolvedThis incident has been resolved.
Jun 30, 01:49 PM
monitoringA fix has been implemented and we are monitoring the results.
Jun 30, 01:45 PM
identifiedWe are aware of egress IP issues in SIN and NRT and are working on a fix. Some machines in SIN and NRT using egress IPs may temporarily lose connectivity or otherwise see degraded performance.

Metrics currently experiencing issues

major
Jun 28, 08:10 AMJun 28, 09:20 PMresolved
Jun 28, 09:20 PM
resolvedThis incident has been resolved.
Jun 28, 08:10 AM
investigatingWe are currently investigating an issue with our metrics cluster.

IPv6 Connectivity Issues in EWR

major
Jun 26, 10:17 PMJun 26, 11:48 PMresolved
Jun 26, 11:48 PM
resolvedThis incident has been resolved.
Jun 26, 11:14 PM
monitoringA fix has been implemented and we are monitoring the results.
Jun 26, 10:17 PM
investigatingOne of our upstream providers is experiencing IPv6 network connectivity problems in EWR. Apps with machines on affected hosts may have impacted connectivity to certain IPv6 destinations while they inv...

Deploys defaulting to Fly-hosted Builders

minor
Jun 25, 01:41 PMJun 25, 03:38 PMresolved
Jun 25, 03:38 PM
resolvedThis incident has been resolved.
Jun 25, 03:08 PM
monitoringWe are seeing improvements in Depot builder provision times and are switching the default deploy strategy back to them. We will continue to monitor builder performance closely. Users with a preferenc...
Jun 25, 01:41 PM
investigatingWe are investigating delays provisioning Depot backed builders for deploys. We have switched the default `fly deploy` strategy to use fly hosted builders at this time. Users can still trigger a depo...

Elevated control plane latency

minor
Jun 25, 12:11 PMJun 25, 03:18 PMresolved
Jun 25, 03:18 PM
resolvedThis incident has been resolved.
Jun 25, 01:48 PM
monitoringA fix has been implemented and we are monitoring the results.
Jun 25, 01:20 PM
identifiedThe issue has been identified and a fix is being implemented.
+1 more updates

Degraded networking in North America

minor
Jun 24, 01:42 AMJun 24, 06:57 AMresolved
Jun 24, 06:57 AM
resolvedThis incident has been resolved.
Jun 24, 03:42 AM
identifiedSome 6PN Private Networking traffic remains impacted into and out of our LAX region, pending upstream resolution.
Jun 24, 02:04 AM
identifiedMost networking is largely healthy between primary North American regions. Some Machines may see ongoing packet loss and higher latency communicating with other Machines on certain routes. We're conti...
+1 more updates

Get Fly.io Outage Alerts

Be the first to know when Fly.io go down.