Fly.io Outage History
50 incidents reported. Data sourced from the official Fly.io status page.
50
Total Incidents
20
Major/Critical
26
Minor
50
Resolved
August 2026
Inbound connection failure to Fly Apps
criticalAug 3, 03:16 PM→Aug 3, 06:50 PMresolved
Aug 3, 03:16 PM
resolved — A BGP misconfiguration while provisioning new edge capacity caused most traffic from Europe endpoints to be dropped, between 14:50 UTC and 15:02 UTC. The misconfiguration has been fixed and we are imp...
Managed Postgres v2 control plane issues in iad
minorAug 1, 05:01 PM→Aug 1, 06:30 PMresolved
Aug 1, 06:30 PM
resolved — This incident has been resolved.
Aug 1, 05:26 PM
monitoring — A fix has been implemented and we are monitoring the results.
Aug 1, 05:01 PM
investigating — We are investigating an issue with the control plane for Managed Postgres v2 in the IAD region. Creating new v2 clusters in the IAD region may fail at this time. Existing clusters continue to run, but...
July 2026
Capacity issues in CDG
noneJul 31, 04:44 PM→Jul 31, 08:11 PMresolved
Jul 31, 08:11 PM
resolved — This incident has been resolved.
Jul 31, 07:02 PM
monitoring — A fix has been implemented and we are monitoring the results.
Jul 31, 04:44 PM
investigating — Creating machines in CDG region may fail at this time with an "no capacity available in cdg" message. Existing apps continue to run.
Outbound email issues
minorJul 31, 01:34 PM→Jul 31, 02:51 PMresolved
Jul 31, 02:51 PM
resolved — This incident has been resolved.
Jul 31, 02:38 PM
monitoring — A fix has been implemented and we are monitoring the results.
Jul 31, 01:34 PM
investigating — Our dashboard is failing to send outbound email. Emails such as new account verification or password reset may fail to send at this time.
Increased API latency
minorJul 29, 12:41 PM→Jul 29, 01:26 PMresolved
Jul 29, 01:26 PM
resolved — This incident has been resolved.
Jul 29, 12:41 PM
investigating — We are investigating some database issues causing high latency on some API endpoints and dashboard operations. You may experience intermittent "503 service unavailable" errors at this time.
Currently ...
High number of 5XX on the Machines API and dashboard
criticalJul 20, 07:10 AM→Jul 20, 05:09 PMresolved
Jul 20, 05:09 PM
resolved — This incident has been resolved.
Jul 20, 01:34 PM
monitoring — We are still working on fixing degraded Managed Postgres clusters.
Jul 20, 09:14 AM
monitoring — Some Managed Postgres v1 clusters are degraded. We are working on fixing them. Managed Postgres v2 is unaffected.
+5 more updates
Egress IPv6 issues in BOM
minorJul 19, 03:48 AM→Jul 19, 04:52 AMresolved
Jul 19, 04:52 AM
resolved — This incident has been resolved.
Jul 19, 03:48 AM
identified — We have identified an upstream issue that is preventing egress IPv6 addresses in BOM from reaching parts of the internet, and we're currently working with an upstream provider to resolve this issue. N...
Brief Flycast / MPG interruption in YYZ
noneJul 17, 06:30 PM→Jul 17, 06:30 PMresolved
Jul 17, 07:13 PM
resolved — A bad deployment momentarily caused issues with Flycast connectivity, and, by extension, MPG, in our YYZ region. The deployment was immediately rolled back and connectivity was restored shortly after.
Edge proxy issues
minorJul 16, 11:34 AM→Jul 16, 01:44 PMresolved
Jul 16, 01:44 PM
resolved — This incident has been resolved.
Jul 16, 12:27 PM
monitoring — A fix has been implemented and we are monitoring the results.
Jul 16, 11:34 AM
investigating — We are investigating increased connection latency and "connection reset" errors from our edge proxy. Apps continue to run, but requests may experience increased connection latency or fail at this time...
App creation failing
minorJul 16, 11:54 AM→Jul 16, 12:17 PMresolved
Jul 16, 12:17 PM
resolved — This incident has been resolved.
Jul 16, 11:54 AM
identified — An issue with our Machines API is causing app creations to fail in some cases. We are working on a fix.
Partial outage in SJC
majorJul 14, 02:51 PM→Jul 14, 04:59 PMresolved
Jul 14, 04:59 PM
resolved — This incident has been resolved.
Jul 14, 03:53 PM
monitoring — A fix has been implemented and we are monitoring the results. Apps should be reachable at this point in time.
Jul 14, 02:51 PM
identified — A subset of hosts in SJC are currently offline. Some apps may be unreachable at this time.
Some DFW hosts offline
minorJul 13, 10:25 PM→Jul 13, 10:58 PMresolved
Jul 13, 10:58 PM
resolved — This incident has been resolved.
Jul 13, 10:36 PM
monitoring — A fix has been implemented and we are monitoring the results.
Jul 13, 10:25 PM
investigating — A subset of hosts in DFW are currently offline, and we're investigating the issue.
Delays starting Depot Builders in IAD
minorJul 9, 03:21 PM→Jul 9, 04:18 PMresolved
Jul 9, 04:18 PM
resolved — This incident has been resolved.
Jul 9, 03:43 PM
monitoring — A fix has been implemented and we are seeing improvements in builder performance, latency, and error rates. We are continuing to monitor for a full recovery.
Customers still seeing issues can trigger...
Jul 9, 03:21 PM
identified — We have identified an issue causing delays or failures when starting depoting builders located in the IAD region. Customers with builders in IAD may see delays or timeouts starting builds during `fly ...
Registry performance issues
minorJul 9, 03:33 PM→Jul 9, 04:04 PMresolved
Jul 9, 04:04 PM
resolved — This incident has been resolved.
Jul 9, 03:55 PM
monitoring — A fix has been implemented and we are monitoring the results.
Jul 9, 03:33 PM
identified — The fly.io registry is currently experiencing capacity constraints that reduced performance and may lead to temporary high latency or failed pushes. We are currently working to add capacity and restor...
Partial Outage in ORD
majorJul 3, 05:00 PM→Jul 3, 06:11 PMresolved
Jul 3, 06:11 PM
resolved — This incident has been resolved.
Jul 3, 05:49 PM
monitoring — A fix has been implemented and we are monitoring the results.
Jul 3, 05:02 PM
identified — We've identified the issue as a networking hardware failure impacting a subset of hosts at one of our Upstream providers in ORD. We are working with our provider to restore connectivity.
+1 more updates
Partial outage in ORD
majorJul 3, 12:11 AM→Jul 3, 05:59 AMresolved
Jul 3, 05:59 AM
resolved — This incident has been resolved.
Jul 3, 04:41 AM
monitoring — Customer workloads are now starting and we're monitoring the affected hosts. Affected Managed Postgres instances will be investigated.
Jul 3, 04:21 AM
identified — Power restoration is ongoing and we're making sure the hosts are healthy before starting customer workloads to avoid issues. Customer impact remains and updates to come.
+4 more updates
Errors issuing new SSL certificates
criticalJul 2, 09:38 PM→Jul 2, 11:18 PMresolved
Jul 2, 11:18 PM
resolved — This incident has been resolved.
Jul 2, 10:50 PM
monitoring — A fix has been implemented upstream and certificates are being issued successfully. We will continue to monitor.
Jul 2, 10:26 PM
identified — The issue has been identified and we are awaiting a fix.
+1 more updates
Static Egress IPv6 issues in NRT
minorJul 1, 01:06 PM→Jul 1, 01:57 PMresolved
Jul 1, 01:57 PM
resolved — This incident has been resolved.
Jul 1, 01:44 PM
monitoring — A fix has been implemented and we are monitoring the results.
Jul 1, 01:06 PM
investigating — We are investigating issues with static egress IPv6 addresses in NRT region. Apps using static egress IPs may experience connectivity failures to some destinations.
Elevated API Errors
majorJul 1, 06:14 AM→Jul 1, 07:50 AMresolved
Jul 1, 07:50 AM
resolved — This incident has been resolved.
Jul 1, 07:26 AM
monitoring — Background jobs have caught up and the API is fully operational. We are continuing to monitor service health.
Jul 1, 07:05 AM
identified — A fix has been put in place, and we are no longer seeing elevated API errors. Some dashboard actions will be delayed while background processing catches up.
+2 more updates
June 2026
Delayed Metrics
majorJun 29, 01:03 AM→Jun 30, 10:19 PMresolved
Jun 30, 10:19 PM
resolved — This incident has been resolved.
Jun 30, 07:58 PM
identified — Almost all metrics have caught up aside from a small handful in `sin` and `syd`. We're continuing to monitor and expect these to complete in the next few hours.
Jun 30, 04:50 PM
identified — Backlogged metrics are still being processed. We're bringing extra processing capacity online to speed up the process.
+9 more updates
Egress IP issues in SIN and NRT
minorJun 30, 01:45 PM→Jun 30, 02:03 PMresolved
Jun 30, 02:03 PM
resolved — This incident has been resolved.
Jun 30, 01:49 PM
monitoring — A fix has been implemented and we are monitoring the results.
Jun 30, 01:45 PM
identified — We are aware of egress IP issues in SIN and NRT and are working on a fix. Some machines in SIN and NRT using egress IPs may temporarily lose connectivity or otherwise see degraded performance.
Metrics currently experiencing issues
majorJun 28, 08:10 AM→Jun 28, 09:20 PMresolved
Jun 28, 09:20 PM
resolved — This incident has been resolved.
Jun 28, 08:10 AM
investigating — We are currently investigating an issue with our metrics cluster.
IPv6 Connectivity Issues in EWR
majorJun 26, 10:17 PM→Jun 26, 11:48 PMresolved
Jun 26, 11:48 PM
resolved — This incident has been resolved.
Jun 26, 11:14 PM
monitoring — A fix has been implemented and we are monitoring the results.
Jun 26, 10:17 PM
investigating — One of our upstream providers is experiencing IPv6 network connectivity problems in EWR. Apps with machines on affected hosts may have impacted connectivity to certain IPv6 destinations while they inv...
Deploys defaulting to Fly-hosted Builders
minorJun 25, 01:41 PM→Jun 25, 03:38 PMresolved
Jun 25, 03:38 PM
resolved — This incident has been resolved.
Jun 25, 03:08 PM
monitoring — We are seeing improvements in Depot builder provision times and are switching the default deploy strategy back to them. We will continue to monitor builder performance closely.
Users with a preferenc...
Jun 25, 01:41 PM
investigating — We are investigating delays provisioning Depot backed builders for deploys. We have switched the default `fly deploy` strategy to use fly hosted builders at this time.
Users can still trigger a depo...
Elevated control plane latency
minorJun 25, 12:11 PM→Jun 25, 03:18 PMresolved
Jun 25, 03:18 PM
resolved — This incident has been resolved.
Jun 25, 01:48 PM
monitoring — A fix has been implemented and we are monitoring the results.
Jun 25, 01:20 PM
identified — The issue has been identified and a fix is being implemented.
+1 more updates
Degraded networking in North America
minorJun 24, 01:42 AM→Jun 24, 06:57 AMresolved
Jun 24, 06:57 AM
resolved — This incident has been resolved.
Jun 24, 03:42 AM
identified — Some 6PN Private Networking traffic remains impacted into and out of our LAX region, pending upstream resolution.
Jun 24, 02:04 AM
identified — Most networking is largely healthy between primary North American regions. Some Machines may see ongoing packet loss and higher latency communicating with other Machines on certain routes. We're conti...
+1 more updates
Network issues in SIN, NRT
majorJun 22, 08:15 PM→Jun 22, 09:56 PMresolved
Jun 22, 09:56 PM
resolved — This incident has been resolved.
Jun 22, 08:15 PM
identified — Our upstream provider is continuing to experience network issues in SIN and NRT regions. Apps running in those regions may be unreachable or experience high packet loss at this time.
SIN, NRT network issues
majorJun 22, 05:10 PM→Jun 22, 07:48 PMresolved
Jun 22, 07:48 PM
resolved — This incident has been resolved.
Jun 22, 06:33 PM
identified — We are continuing to experience network issues with an upstream provider in SIN and NRT regions.
Jun 22, 05:44 PM
monitoring — A fix has been implemented and we are monitoring the results.
+2 more updates
Log search unavailable
minorJun 19, 06:05 PM→Jun 19, 09:53 PMresolved
Jun 19, 09:53 PM
resolved — Most queued historical logs have been ingested and should now be available through log search. Log ingestion rates have returned to normal levels.
Jun 19, 06:19 PM
monitoring — We've applied a fix for this issue. Historical logs are currently backfilling. We will post an update once logs have finished backfilling and current logs are being ingested normally.
Jun 19, 06:10 PM
investigating — Log search is available; however, new app logs since ~1 hour ago are missing and new logs are not being ingested. We are continuing to investigate.
+1 more updates
Network Issues in SIN
majorJun 17, 03:21 AM→Jun 17, 04:55 AMresolved
Jun 17, 04:55 AM
resolved — This incident has been resolved.
Jun 17, 04:38 AM
monitoring — Network connectivity in SIN has been fully restored. We're continuing to monitor.
Jun 17, 04:22 AM
identified — We are seeing recovery of network connectivity between SIN and most destinations. We're continuing to work with our upstream provider to resolve the remaining issues.
+2 more updates
Macaroon Auth + Machines API Issues
criticalJun 15, 03:03 PM→Jun 15, 04:41 PMresolved
Jun 15, 04:41 PM
resolved — This incident has been resolved and we are seeing all platform functions operate normally.
Jun 15, 04:15 PM
monitoring — A fix has been implemented and we are monitoring the results.
Jun 15, 04:15 PM
identified — We have deployed another change and are seeing wider improvements in platform stability across all regions. Performance is trending to normal, though users may still see some degradation at this time....
+6 more updates
MPG cluster provisioning is broken
noneJun 15, 06:20 AM→Jun 15, 06:39 AMresolved
Jun 15, 06:39 AM
resolved — This incident has been resolved.
Jun 15, 06:20 AM
investigating — New MPG cluster provisioning is broken. Existing MPG clusters are not affected. Newly created organizations may see errors while SSHing into their machines. We are investigating the issue.
Elevated Sprites error rates in SIN
majorJun 12, 02:35 AM→Jun 12, 04:30 AMresolved
Jun 12, 04:30 AM
resolved — This incident has been resolved.
Jun 12, 03:24 AM
monitoring — A fix has been implemented and we are seeing error rates for sprites in SIN normalize. We are continuing to monitor to ensure full recovery.
Jun 12, 02:35 AM
investigating — We are investigating elevated 500 / internal server error rates with Sprites in the SIN region. Users may see increased errors when accessing sprites located in this region, or for requests to the Spr...
Increased network latency in North America
minorJun 11, 01:16 AM→Jun 11, 11:12 AMresolved
Jun 11, 02:12 PM
resolved — This incident has been resolved.
Jun 11, 10:09 AM
monitoring — Our upstream paths have been fixed. We are monitoring the results.
Jun 11, 09:35 AM
identified — We are working with our upstream network provider to address periodic loss of connectivity over transit in ord
+5 more updates
Ingress Traffic issues in GRU
majorJun 10, 07:39 PM→Jun 10, 07:52 PMresolved
Jun 10, 07:52 PM
resolved — This incident has been resolved.
Jun 10, 07:39 PM
identified — Some of our edge nodes in GRU has suffered an error that crashed some of the critical services. We're currently working to bring them back online. Some traffic entering through GRU (i.e. users connect...
Emergency maintenance of Petsem causing some control plane errors
majorJun 9, 06:20 PM→Jun 9, 06:42 PMresolved
Jun 9, 06:42 PM
resolved — This incident has been resolved.
Jun 9, 06:28 PM
monitoring — The maintenance has been completed and control plane functions should recover to normal. We're monitoring for any further complications.
Jun 9, 06:20 PM
identified — We're performing an emergency maintenance on Petsem, our secrets management service. Some control plane write operations may temporarily fail, for example, creating new apps or secrets. Existing apps ...
Managed Postgres Control Plane Issues in IAD
majorJun 9, 02:57 PM→Jun 9, 04:15 PMresolved
Jun 9, 09:31 PM
resolved — This incident has been resolved.
Jun 9, 03:19 PM
monitoring — An initial fix has been implemented and connectivity to all impacted clusters has been restored. We are continuing to monitor to ensure stable recovery.
Jun 9, 03:06 PM
identified — We are continuing to address this issue. Some clusters in IAD are unavailable at this time, some users may have seen unexpected cluter restarts. We are working on restoring normal performance for all ...
+1 more updates
egress ips are broken in ORD
minorJun 9, 09:16 AM→Jun 9, 10:01 AMresolved
Jun 9, 10:01 AM
resolved — This incident has been resolved.
Jun 9, 09:37 AM
monitoring — A fix has been implemented and we are monitoring the results.
Jun 9, 09:16 AM
investigating — Egress ips are broken in most of ORD, we are currently investigating this issue
Capacity issues in ARN region
minorJun 8, 10:31 AM→Jun 8, 12:30 PMresolved
Jun 8, 12:30 PM
resolved — This incident has been resolved.
Jun 8, 10:31 AM
investigating — The ARN region is low on available host capacity. Creating new machines, or starting currently stopped/suspended machines, may fail at this time.
We are working on provisioning new host capacity in t...
Consul cluster degradation
noneJun 4, 01:44 PM→Jun 4, 05:03 PMresolved
Jun 4, 05:03 PM
resolved — We have restored the degraded Consul cluster. All affected functionality is now working correctly: Unmanaged Postgres and LiteFS with dynamic leases.
Jun 4, 04:02 PM
monitoring — We have restored the degraded Consul cluster and are monitoring for stability. All affected functionality should now be working correctly: Unmanaged Postgres and LiteFS with dynamic leases.
Jun 4, 01:44 PM
identified — One of our Consul clusters is in degraded state due to a failed node. This can cause issues with LiteFS primary node selection, Unmanaged Postgres (14.x and older *only*), and creation of new Unmanage...
Issues with flyctl ssh console and Machines OIDC
minorJun 1, 06:33 PM→Jun 1, 06:43 PMresolved
Jun 1, 06:43 PM
resolved — This incident has been resolved.
Jun 1, 06:35 PM
monitoring — A fix has been implemented and we are monitoring the results.
Jun 1, 06:33 PM
investigating — We're currently investigating an issue affecting flyctl ssh console functionality and machines' OIDC tokens.
May 2026
IPv6 outage for some machines in ORD
majorMay 31, 06:34 AM→May 31, 10:52 AMresolved
May 31, 10:52 AM
resolved — This incident has been resolved.
May 31, 07:03 AM
monitoring — A fix has been implemented and we are monitoring the results.
May 31, 06:34 AM
investigating — We're working with our upstream providers to investigate an IPv6 networking failure in ORD.
Private networking issues in SYD
minorMay 30, 01:29 PM→May 30, 01:37 PMresolved
May 30, 01:37 PM
resolved — This incident has been resolved.
May 30, 01:32 PM
monitoring — A fix has been implemented and we are monitoring the results.
May 30, 01:29 PM
identified — Due to an upstream provider issue, Private Networking (6PN) is currently degraded in SYD region. Communication between Machines in SYD region and Machines in other regions may fail at this time. Newly...
Elevated deployment errors
minorMay 30, 02:42 AM→May 30, 04:21 AMresolved
May 30, 04:21 AM
resolved — This incident has been resolved.
May 30, 03:37 AM
monitoring — A fix has been implemented and we are monitoring the results.
May 30, 03:13 AM
identified — We identified the issue and are working on a fix.
+1 more updates
Networking issues in ORD
minorMay 29, 07:09 PM→May 29, 07:45 PMresolved
May 29, 07:45 PM
resolved — This incident has been resolved.
May 29, 07:21 PM
monitoring — A fix has been implemented and we are monitoring the results.
May 29, 07:09 PM
identified — We are aware of increased latency and connection drops for clients located near Chicago (ORD) and are currently working on a fix.
Networking issues in ORD
minorMay 29, 06:59 AM→May 29, 04:50 PMresolved
May 29, 04:50 PM
resolved — This incident has been resolved.
May 29, 10:22 AM
monitoring — A fix has been implemented and we are monitoring the results.
May 29, 06:59 AM
investigating — We are currently investigating increased latency and dropped connections in ORD (Chicago).
Networking issues in ORD
minorMay 29, 12:42 AM→May 29, 01:27 AMresolved
May 29, 01:27 AM
resolved — This incident has been resolved.
May 29, 01:00 AM
monitoring — A fix has been implemented and we are monitoring the results.
May 29, 12:42 AM
investigating — We are currently investigating increased latency and dropped connections in ORD (Chicago).
Increased latency in SJC
minorMay 28, 09:08 PM→May 28, 11:07 PMresolved
May 28, 11:07 PM
resolved — This incident has been resolved.
May 28, 10:34 PM
monitoring — We are continuing to monitor for any further issues.
May 28, 10:33 PM
monitoring — A fix has been implemented and we are monitoring the results.
+3 more updates
Private networking issues in SYD
majorMay 27, 11:54 AM→May 27, 12:52 PMresolved
May 27, 12:52 PM
resolved — This incident has been resolved.
May 27, 12:13 PM
monitoring — A fix has been implemented and we are monitoring the results.
May 27, 11:54 AM
identified — Due to an upstream provider issue, Private Networking (6PN) is currently degraded in SYD region. Communication between Machines in SYD region and Machines in other regions may fail at this time. Newly...
Elevated GraphQL API Latency
minorMay 27, 03:50 AM→May 27, 04:36 AMresolved
May 27, 04:36 AM
resolved — This incident has been resolved.
May 27, 04:07 AM
monitoring — A fix has been implemented and we are monitoring the results.
May 27, 03:50 AM
investigating — We are investigating elevated API Latency. Users may see delays or errors creating apps, as well as on some dashboard pages.
Related Incident Histories
Get Fly.io Outage Alerts
Be the first to know when Fly.io go down.