- Detected by Pingoru
- Aug 01, 2026, 05:01 PM UTC
- Resolved
- Aug 01, 2026, 06:30 PM UTC
- Duration
- 1h 28m
Affected: Management Plane - IAD
Timeline · 3 updates
-
investigating Aug 01, 2026, 05:01 PM UTC
We are investigating an issue with the control plane for Managed Postgres v2 in the IAD region. Creating new v2 clusters in the IAD region may fail at this time. Existing clusters continue to run, but may experience lagging backups.
-
monitoring Aug 01, 2026, 05:26 PM UTC
A fix has been implemented and we are monitoring the results.
-
resolved Aug 01, 2026, 06:30 PM UTC
This incident has been resolved.
Read the full incident report →
- Detected by Pingoru
- Jul 31, 2026, 04:44 PM UTC
- Resolved
- Jul 31, 2026, 08:11 PM UTC
- Duration
- 3h 27m
Affected: CDG - Paris, France
Timeline · 3 updates
-
investigating Jul 31, 2026, 04:44 PM UTC
Creating machines in CDG region may fail at this time with an "no capacity available in cdg" message. Existing apps continue to run.
-
monitoring Jul 31, 2026, 07:02 PM UTC
A fix has been implemented and we are monitoring the results.
-
resolved Jul 31, 2026, 08:11 PM UTC
This incident has been resolved.
Read the full incident report →
- Detected by Pingoru
- Jul 31, 2026, 01:34 PM UTC
- Resolved
- Jul 31, 2026, 02:51 PM UTC
- Duration
- 1h 16m
Affected: Dashboard
Timeline · 3 updates
-
investigating Jul 31, 2026, 01:34 PM UTC
Our dashboard is failing to send outbound email. Emails such as new account verification or password reset may fail to send at this time.
-
monitoring Jul 31, 2026, 02:38 PM UTC
A fix has been implemented and we are monitoring the results.
-
resolved Jul 31, 2026, 02:51 PM UTC
This incident has been resolved.
Read the full incident report →
- Detected by Pingoru
- Jul 29, 2026, 12:41 PM UTC
- Resolved
- Jul 29, 2026, 01:26 PM UTC
- Duration
- 44m
Affected: DashboardMachines APIDeployments
Timeline · 2 updates
-
investigating Jul 29, 2026, 12:41 PM UTC
We are investigating some database issues causing high latency on some API endpoints and dashboard operations. You may experience intermittent "503 service unavailable" errors at this time. Currently deployed apps continue to run.
-
resolved Jul 29, 2026, 01:26 PM UTC
This incident has been resolved.
Read the full incident report →
- Detected by Pingoru
- Jul 20, 2026, 07:10 AM UTC
- Resolved
- Jul 20, 2026, 05:09 PM UTC
- Duration
- 9h 59m
Affected: DashboardMachines API
Timeline · 8 updates
-
investigating Jul 20, 2026, 07:10 AM UTC
Existing machines are unaffected. We are investigating the issue.
-
investigating Jul 20, 2026, 07:48 AM UTC
We are continuing to investigate this issue.
-
identified Jul 20, 2026, 07:51 AM UTC
We've identified an internal service providing authentication to our Machines API has failed, our team is currently looking at our options for restoring this service. Existing Machines/Apps will continue to run as normal. Thank you for your patience.
-
identified Jul 20, 2026, 07:52 AM UTC
We've identified an internal service providing authentication to our Machines API has failed, our team is currently looking at our options for restoring this service. Existing Machines/Apps will continue to run as normal. Thank you for your patience.
-
monitoring Jul 20, 2026, 08:02 AM UTC
A fix has been implemented and we are monitoring the results.
-
monitoring Jul 20, 2026, 09:14 AM UTC
Some Managed Postgres v1 clusters are degraded. We are working on fixing them. Managed Postgres v2 is unaffected.
-
monitoring Jul 20, 2026, 01:34 PM UTC
We are still working on fixing degraded Managed Postgres clusters.
-
resolved Jul 20, 2026, 05:09 PM UTC
This incident has been resolved.
Read the full incident report →
- Detected by Pingoru
- Jul 19, 2026, 03:48 AM UTC
- Resolved
- Jul 19, 2026, 04:52 AM UTC
- Duration
- 1h 4m
Affected: BOM - Mumbai, India
Timeline · 2 updates
-
identified Jul 19, 2026, 03:48 AM UTC
We have identified an upstream issue that is preventing egress IPv6 addresses in BOM from reaching parts of the internet, and we're currently working with an upstream provider to resolve this issue. Normal IPv6 addresses remain unaffected.
-
resolved Jul 19, 2026, 04:52 AM UTC
This incident has been resolved.
Read the full incident report →
- Detected by Pingoru
- Jul 17, 2026, 07:13 PM UTC
- Resolved
- Jul 17, 2026, 06:30 PM UTC
- Duration
- —
Timeline · 1 update
-
resolved Jul 17, 2026, 07:13 PM UTC
A bad deployment momentarily caused issues with Flycast connectivity, and, by extension, MPG, in our YYZ region. The deployment was immediately rolled back and connectivity was restored shortly after.
Read the full incident report →
- Detected by Pingoru
- Jul 16, 2026, 11:54 AM UTC
- Resolved
- Jul 16, 2026, 12:17 PM UTC
- Duration
- 22m
Affected: Machines API
Timeline · 2 updates
-
identified Jul 16, 2026, 11:54 AM UTC
An issue with our Machines API is causing app creations to fail in some cases. We are working on a fix.
-
resolved Jul 16, 2026, 12:17 PM UTC
This incident has been resolved.
Read the full incident report →
- Detected by Pingoru
- Jul 16, 2026, 11:34 AM UTC
- Resolved
- Jul 16, 2026, 01:44 PM UTC
- Duration
- 2h 10m
Affected: Customer Applications
Timeline · 3 updates
-
investigating Jul 16, 2026, 11:34 AM UTC
We are investigating increased connection latency and "connection reset" errors from our edge proxy. Apps continue to run, but requests may experience increased connection latency or fail at this time.
-
monitoring Jul 16, 2026, 12:27 PM UTC
A fix has been implemented and we are monitoring the results.
-
resolved Jul 16, 2026, 01:44 PM UTC
This incident has been resolved.
Read the full incident report →
- Detected by Pingoru
- Jul 14, 2026, 02:51 PM UTC
- Resolved
- Jul 14, 2026, 04:59 PM UTC
- Duration
- 2h 7m
Affected: SJC - San Jose, California (US)
Timeline · 3 updates
-
identified Jul 14, 2026, 02:51 PM UTC
A subset of hosts in SJC are currently offline. Some apps may be unreachable at this time.
-
monitoring Jul 14, 2026, 03:53 PM UTC
A fix has been implemented and we are monitoring the results. Apps should be reachable at this point in time.
-
resolved Jul 14, 2026, 04:59 PM UTC
This incident has been resolved.
Read the full incident report →
- Detected by Pingoru
- Jul 13, 2026, 10:25 PM UTC
- Resolved
- Jul 13, 2026, 10:58 PM UTC
- Duration
- 33m
Affected: DFW - Dallas, Texas (US)
Timeline · 3 updates
-
investigating Jul 13, 2026, 10:25 PM UTC
A subset of hosts in DFW are currently offline, and we're investigating the issue.
-
monitoring Jul 13, 2026, 10:36 PM UTC
A fix has been implemented and we are monitoring the results.
-
resolved Jul 13, 2026, 10:58 PM UTC
This incident has been resolved.
Read the full incident report →
- Detected by Pingoru
- Jul 09, 2026, 03:33 PM UTC
- Resolved
- Jul 09, 2026, 04:04 PM UTC
- Duration
- 31m
Affected: Deployments
Timeline · 3 updates
-
identified Jul 09, 2026, 03:33 PM UTC
The fly.io registry is currently experiencing capacity constraints that reduced performance and may lead to temporary high latency or failed pushes. We are currently working to add capacity and restore service.
-
monitoring Jul 09, 2026, 03:55 PM UTC
A fix has been implemented and we are monitoring the results.
-
resolved Jul 09, 2026, 04:04 PM UTC
This incident has been resolved.
Read the full incident report →
- Detected by Pingoru
- Jul 09, 2026, 03:21 PM UTC
- Resolved
- Jul 09, 2026, 04:18 PM UTC
- Duration
- 56m
Affected: DeploymentsRemote Builds
Timeline · 3 updates
-
identified Jul 09, 2026, 03:21 PM UTC
We have identified an issue causing delays or failures when starting depoting builders located in the IAD region. Customers with builders in IAD may see delays or timeouts starting builds during `fly deploy`. We are working on a fix. In the meantime customers can trigger a fly-hosted builder based deploy with `fly deploy --depot=false`. You can also change your builder region away from IAD via the `settings` tab of your Fly dashboard. We recommend DFW or ORD as alternate builder regions at this time. Please note, your builder region can differ from the region your machines are in.
-
monitoring Jul 09, 2026, 03:43 PM UTC
A fix has been implemented and we are seeing improvements in builder performance, latency, and error rates. We are continuing to monitor for a full recovery. Customers still seeing issues can trigger a fly-hosted builder based deploy with `fly deploy --depot=false`.
-
resolved Jul 09, 2026, 04:18 PM UTC
This incident has been resolved.
Read the full incident report →
- Detected by Pingoru
- Jul 03, 2026, 05:00 PM UTC
- Resolved
- Jul 03, 2026, 06:11 PM UTC
- Duration
- 1h 11m
Affected: Management Plane - ORDORD - Chicago, Illinois (US)
Timeline · 4 updates
-
investigating Jul 03, 2026, 05:00 PM UTC
We are investigating an issue with one of our upstream providers in ORD. Machines across a subset of hosts may be unreachable or not running correctly. Deploys with machines on these hosts may fail at this time. Some Managed Postgres clusters in ORD region may be unavailable or see connectivity issues at this time.
-
identified Jul 03, 2026, 05:02 PM UTC
We've identified the issue as a networking hardware failure impacting a subset of hosts at one of our Upstream providers in ORD. We are working with our provider to restore connectivity.
-
monitoring Jul 03, 2026, 05:49 PM UTC
A fix has been implemented and we are monitoring the results.
-
resolved Jul 03, 2026, 06:11 PM UTC
This incident has been resolved.
Read the full incident report →
- Detected by Pingoru
- Jul 03, 2026, 12:11 AM UTC
- Resolved
- Jul 03, 2026, 05:59 AM UTC
- Duration
- 5h 48m
Affected: Management Plane - ORDORD - Chicago, Illinois (US)
Timeline · 7 updates
-
investigating Jul 03, 2026, 12:11 AM UTC
We are investigating an issue with one of our upstream providers in ORD. Machines across a subset of hosts may be unreachable or not running correctly. Deploys with machines on these hosts may fail at this time. Some Managed Postgres clusters in ORD region may be unavailable at this time.
-
identified Jul 03, 2026, 12:29 AM UTC
We've identified and reported power issues with one of our upstream providers in ORD. We're waiting for an update from our upstream for a resolution. Some Managed Postgres clusters in ORD will be unavailable due to placement.
-
identified Jul 03, 2026, 12:54 AM UTC
Our provider has advised us their facilities team is working on restoring the power, we'll provide another update as soon as we learn more.
-
identified Jul 03, 2026, 02:59 AM UTC
Power restoration work in a subset of ORD is still in progress and impact remains ongoing for a subset of hosts and some Managed Postgres clusters.
-
identified Jul 03, 2026, 04:21 AM UTC
Power restoration is ongoing and we're making sure the hosts are healthy before starting customer workloads to avoid issues. Customer impact remains and updates to come.
-
monitoring Jul 03, 2026, 04:41 AM UTC
Customer workloads are now starting and we're monitoring the affected hosts. Affected Managed Postgres instances will be investigated.
-
resolved Jul 03, 2026, 05:59 AM UTC
This incident has been resolved.
Read the full incident report →
- Detected by Pingoru
- Jul 02, 2026, 09:38 PM UTC
- Resolved
- Jul 02, 2026, 11:18 PM UTC
- Duration
- 1h 39m
Affected: SSL/TLS Certificate Provisioning
Timeline · 4 updates
-
investigating Jul 02, 2026, 09:38 PM UTC
We are currently investigating errors when issuing new SSL certificates for hostnames.
-
identified Jul 02, 2026, 10:26 PM UTC
The issue has been identified and we are awaiting a fix.
-
monitoring Jul 02, 2026, 10:50 PM UTC
A fix has been implemented upstream and certificates are being issued successfully. We will continue to monitor.
-
resolved Jul 02, 2026, 11:18 PM UTC
This incident has been resolved.
Read the full incident report →
- Detected by Pingoru
- Jul 02, 2026, 08:05 AM UTC
- Resolved
- Jul 02, 2026, 08:17 AM UTC
- Duration
- 11m
Affected: CorrosionNRT - Tokyo, Japan
Timeline · 1 update
-
investigating Jul 02, 2026, 08:05 AM UTC
A single host in the NTR region is experiencing some performance issues. The issue has been identified and our engineers are working on a resolution.
Read the full incident report →
- Detected by Pingoru
- Jul 01, 2026, 01:06 PM UTC
- Resolved
- Jul 01, 2026, 01:57 PM UTC
- Duration
- 50m
Affected: NRT - Tokyo, Japan
Timeline · 3 updates
-
investigating Jul 01, 2026, 01:06 PM UTC
We are investigating issues with static egress IPv6 addresses in NRT region. Apps using static egress IPs may experience connectivity failures to some destinations.
-
monitoring Jul 01, 2026, 01:44 PM UTC
A fix has been implemented and we are monitoring the results.
-
resolved Jul 01, 2026, 01:57 PM UTC
This incident has been resolved.
Read the full incident report →
- Detected by Pingoru
- Jul 01, 2026, 06:14 AM UTC
- Resolved
- Jul 01, 2026, 07:50 AM UTC
- Duration
- 1h 35m
Affected: Dashboard
Timeline · 5 updates
-
investigating Jul 01, 2026, 06:14 AM UTC
We are investigating elevated errors with our GraphQL API and background job processing
-
identified Jul 01, 2026, 06:27 AM UTC
The cause of the errors has been identified and we are working on a fix.
-
identified Jul 01, 2026, 07:05 AM UTC
A fix has been put in place, and we are no longer seeing elevated API errors. Some dashboard actions will be delayed while background processing catches up.
-
monitoring Jul 01, 2026, 07:26 AM UTC
Background jobs have caught up and the API is fully operational. We are continuing to monitor service health.
-
resolved Jul 01, 2026, 07:50 AM UTC
This incident has been resolved.
Read the full incident report →
- Detected by Pingoru
- Jun 30, 2026, 01:45 PM UTC
- Resolved
- Jun 30, 2026, 02:03 PM UTC
- Duration
- 18m
Affected: NRT - Tokyo, JapanSIN - Singapore
Timeline · 3 updates
-
identified Jun 30, 2026, 01:45 PM UTC
We are aware of egress IP issues in SIN and NRT and are working on a fix. Some machines in SIN and NRT using egress IPs may temporarily lose connectivity or otherwise see degraded performance.
-
monitoring Jun 30, 2026, 01:49 PM UTC
A fix has been implemented and we are monitoring the results.
-
resolved Jun 30, 2026, 02:03 PM UTC
This incident has been resolved.
Read the full incident report →
- Detected by Pingoru
- Jun 29, 2026, 01:03 AM UTC
- Resolved
- Jun 30, 2026, 10:19 PM UTC
- Duration
- 1d 21h
Affected: Metrics
Timeline · 12 updates
Read the full incident report →
- Detected by Pingoru
- Jun 28, 2026, 08:10 AM UTC
- Resolved
- Jun 28, 2026, 09:20 PM UTC
- Duration
- 13h 10m
Affected: Metrics
Timeline · 2 updates
-
investigating Jun 28, 2026, 08:10 AM UTC
We are currently investigating an issue with our metrics cluster.
-
resolved Jun 28, 2026, 09:20 PM UTC
This incident has been resolved.
Read the full incident report →
- Detected by Pingoru
- Jun 26, 2026, 10:17 PM UTC
- Resolved
- Jun 26, 2026, 11:48 PM UTC
- Duration
- 1h 30m
Affected: EWR - Secaucus, NJ (US)
Timeline · 3 updates
-
investigating Jun 26, 2026, 10:17 PM UTC
One of our upstream providers is experiencing IPv6 network connectivity problems in EWR. Apps with machines on affected hosts may have impacted connectivity to certain IPv6 destinations while they investigate and resolve this issue.
-
monitoring Jun 26, 2026, 11:14 PM UTC
A fix has been implemented and we are monitoring the results.
-
resolved Jun 26, 2026, 11:48 PM UTC
This incident has been resolved.
Read the full incident report →
- Detected by Pingoru
- Jun 25, 2026, 01:41 PM UTC
- Resolved
- Jun 25, 2026, 03:38 PM UTC
- Duration
- 1h 56m
Affected: Deployments
Timeline · 3 updates
-
investigating Jun 25, 2026, 01:41 PM UTC
We are investigating delays provisioning Depot backed builders for deploys. We have switched the default `fly deploy` strategy to use fly hosted builders at this time. Users can still trigger a depot based deploy with `fly deploy --depot=true`
-
monitoring Jun 25, 2026, 03:08 PM UTC
We are seeing improvements in Depot builder provision times and are switching the default deploy strategy back to them. We will continue to monitor builder performance closely. Users with a preference can trigger a Fly builder based deploy with `fly deploy --depot=false` or use `fly deploy --depot=true` to force a depot-based deployment.
-
resolved Jun 25, 2026, 03:38 PM UTC
This incident has been resolved.
Read the full incident report →
- Detected by Pingoru
- Jun 25, 2026, 12:11 PM UTC
- Resolved
- Jun 25, 2026, 03:18 PM UTC
- Duration
- 3h 7m
Affected: BOM - Mumbai, IndiaNRT - Tokyo, Japan
Timeline · 4 updates
-
investigating Jun 25, 2026, 12:11 PM UTC
We're addressing elevated control plane latency and saturation affecting the BOM and NRT regions. Apps with machines in this region might experience longer response times and possible timeouts (502 errors).
-
identified Jun 25, 2026, 01:20 PM UTC
The issue has been identified and a fix is being implemented.
-
monitoring Jun 25, 2026, 01:48 PM UTC
A fix has been implemented and we are monitoring the results.
-
resolved Jun 25, 2026, 03:18 PM UTC
This incident has been resolved.
Read the full incident report →