Nanonets Outage History
Nanonets is up right nowNanonets had 49 outages in the last 2 years totaling 23h 8m of downtime — averaging 2 incidents per month.
There were 49 Nanonets outages since February 1, 2025 totaling 23h 8m of downtime. Each is summarised below — incident details, duration, and resolution information.
Processing delay on Agent Tasks
Timeline · 1 update
- resolved Jun 30, 2026, 12:07 PM UTC
Our database was running at capacity due to disk usage surge, this lasted for an hour before we recognized the problem and scaled the database.
Processing Delay on Agents
Timeline · 1 update
- resolved Jun 30, 2026, 12:03 PM UTC
We observed processing delay on tasks in Agents Platform for a couple of hours caused by architectural changes introduced by our team to increase observability of our platform.
Brief Service Unavailability on June 29 for app.nanonets.com
Timeline · 1 update
- resolved Jun 29, 2026, 04:41 AM UTC
Between 04:21 AM UTC and 04:25 AM UTC on June 29th, app.nanonets.com experienced a brief service disruption, during which the application was unavailable. The issue was caused by an unexpected error encountered during backend infrastructure operations, resulting in a temporary interruption of service. The issue was identified and resolved promptly, and all services have been fully restored. We apologize for the inconvenience and appreciate your patience.
Agents UI freezing issue
Timeline · 1 update
- resolved Jun 26, 2026, 07:06 AM UTC
Frontend resources taking time to load causing UI to freeze in loading state. Appropriate CDN fixes were implemented to fix the issue.
Brief Service Unavailability on June 1 for app.nanonets.com
Timeline · 1 update
- resolved May 31, 2026, 07:35 PM UTC
Between 12:53 AM IST and 01:00 AM IST on June 1, app.nanonets.com experienced a brief service disruption, during which the application was unavailable. The issue was caused by an unexpected error encountered during backend infrastructure operations, resulting in a temporary interruption of service. The issue was identified and resolved promptly, and all services have been fully restored. We are reviewing the incident and strengthening our deployment and validation processes to further reduce the likelihood of similar occurrences. We apologize for the inconvenience and appreciate your patience.
RCA: Performance Degradation for LLM Nano Models on May 14
Timeline · 1 update
- resolved May 15, 2026, 07:14 AM UTC
From 22:00 UTC to 23:15 UTC on May 14th, we observed performance degradation in the oneshot service affecting LLM Nano models. The issue was primarily caused by a network issue with our GPU service provider. Our team worked with the provider to stabilize the service and performance has since recovered. As an additional preventive measure, we are introducing stronger fallback mechanisms to route requests to more accurate backup models during such events. We apologize for the inconvenience caused and appreciate your patience.
Elevated 5xx Errors & Timeouts for Instant Learning (LLM Nano)
Timeline · 3 updates
- investigating Apr 28, 2026, 05:16 AM UTC
We are currently investigating an issue affecting our GPU service provider infrastructure, which is causing elevated 5xx errors and timeouts for Instant Learning models using LLM Nano. Our team is actively working with the provider to identify and resolve the underlying network issues. We will share further updates as soon as we have more information.
- monitoring Apr 28, 2026, 05:37 AM UTC
We’ve temporarily routed LLM Nano traffic to LLM Mini, a more stable, accurate and higher-capacity variant, to mitigate errors. File processing should now be faster and more reliable while we continue working on resolving the underlying issue.
- resolved Apr 28, 2026, 06:02 AM UTC
This incident has been resolved.
Intermittent 503 Errors for Subset of Users on app.nanonets.com
Timeline · 1 update
- resolved Apr 10, 2026, 07:13 AM UTC
Between 7:19 UTC and 7:29 UTC, a subset of users experienced intermittent 503 errors when accessing app.nanonets.com. A backend node went down and was replaced by a new node. Due to a DNS caching issue, traffic for some users could not be routed to the new node, resulting in failed requests. Other users were unaffected. The issue was identified and resolved by updating the load balancer configuration to ensure proper traffic routing. Preventive measures have been put in place to avoid recurrence.
Delayed file processing for all ILM models on app.nanonets.com
Timeline · 5 updates
RCA for Processing Delays for Instant Learning Models on app.nanonets.com – 04 Mar
Timeline · 1 update
Post-Processing Failures Impacting app.nanonets.com Models
Timeline · 3 updates
- investigating Feb 11, 2026, 09:10 AM UTC
We are currently investigating this issue.
- monitoring Feb 11, 2026, 09:15 AM UTC
A fix has been implemented and we are monitoring the results.
- resolved Feb 11, 2026, 09:26 AM UTC
This incident has been resolved.
Delayed File Processing and Intermittent Failures Across Regions
Timeline · 3 updates
- identified Feb 03, 2026, 03:13 PM UTC
One of our GPU service provider is facing issues. We are working on resolving the issue with them.
- monitoring Feb 03, 2026, 03:22 PM UTC
A fix has been implemented and we are monitoring the results.
- resolved Feb 03, 2026, 03:35 PM UTC
This incident has been resolved.
Delayed file processing for all ILM models in app.nanonets.com
Timeline · 6 updates
Delayed file processing for models using async api in EU region
Timeline · 5 updates
Delayed File Processing in EU & EU-Open Region
Timeline · 4 updates
- investigating Jan 23, 2026, 05:55 PM UTC
We are currently investigating this issue.
- identified Jan 23, 2026, 06:25 PM UTC
The issue has been identified and a fix is being implemented.
- monitoring Jan 23, 2026, 06:48 PM UTC
A fix has been implemented and we are monitoring the results.
- resolved Jan 23, 2026, 07:04 PM UTC
This incident has been resolved.
Async file processing & Extract data section file visibility issue for users using app.nanonets.com
Timeline · 3 updates
- investigating Jan 21, 2026, 08:58 AM UTC
We are currently investigating this issue.
- monitoring Jan 21, 2026, 09:07 AM UTC
A fix has been implemented and we are monitoring the results.
- resolved Jan 21, 2026, 09:34 AM UTC
This incident has been resolved.
Delay in file processing for Instant learning models in US region
Timeline · 4 updates
- investigating Dec 11, 2025, 07:56 AM UTC
We are currently investigating this issue.
- monitoring Dec 11, 2025, 09:30 AM UTC
Our sync API is now operating normally. For async uploads, most results are already available and any remaining pending results will be processed within the next few minutes as we clear the backlog. Our team is addressing the issue on priority. We sincerely apologize for the inconvenience caused.
- resolved Dec 11, 2025, 09:52 AM UTC
This incident has been resolved.
- postmortem Dec 12, 2025, 07:45 AM UTC
One of our core processing services experienced an unexpected surge in load, which slowed down parts of our system and led to a backlog in processing. Our engineering team identified the underlying bottleneck and implemented fixes to stabilize performance. We are also rolling out improvements to make our platform more resilient to sudden spikes in usage. We sincerely apologize for the inconvenience caused.
IN region Outage
Timeline · 3 updates
- investigating Dec 08, 2025, 07:52 AM UTC
Service Disruption on https://in.nanonets.com US and EU regions are not affected
- identified Dec 08, 2025, 07:58 AM UTC
The issue has been identified and a fix is being implemented.
- resolved Dec 08, 2025, 07:59 AM UTC
This incident has been resolved.
High Response Times for Instant learning models in US region
Timeline · 6 updates
Failures in some instant learning models
Timeline · 4 updates
- identified Nov 08, 2025, 06:33 PM UTC
The issue has been identified and a fix is being implemented.
- monitoring Nov 08, 2025, 08:04 PM UTC
A fix has been implemented and we are monitoring the results.
- monitoring Nov 08, 2025, 08:04 PM UTC
We are continuing to monitor for any further issues.
- resolved Nov 08, 2025, 08:46 PM UTC
This incident has been resolved.
Service Disruption Due to Global AWS Outage
Timeline · 4 updates
- identified Oct 20, 2025, 09:11 PM UTC
On 20th Oct, a global AWS outage impacted one of our feature flag service providers, resulting in intermittent failures for some of our models. The issue originated from the provider's infrastructure dependency on affected AWS regions, which caused disruptions in feature flag evaluations within our system. Our engineering team promptly identified the root cause and implemented mitigation measures to restore functionality. We are continuously monitoring the situation and working with our partners to ensure full stability. We apologize for the inconvenience caused and appreciate your patience and understanding.
- monitoring Oct 20, 2025, 09:28 PM UTC
A fix has been implemented and we are monitoring the results.
- monitoring Oct 20, 2025, 09:29 PM UTC
We are continuing to monitor for any further issues.
- resolved Oct 20, 2025, 09:43 PM UTC
This incident has been resolved.
app.nanonets.com down
Timeline · 5 updates
- investigating Sep 22, 2025, 12:01 PM UTC
We are currently investigating this issue.
- identified Sep 22, 2025, 12:07 PM UTC
The issue has been identified and a fix is being implemented.
- monitoring Sep 22, 2025, 12:09 PM UTC
A fix has been implemented and we are monitoring the results.
- resolved Sep 22, 2025, 12:17 PM UTC
This incident has been resolved.
- postmortem Sep 22, 2025, 02:09 PM UTC
**Temporary Service Disruption on Sept 22nd \(US Region\)** We experienced a temporary outage affecting our US region application \([_app.nanonets.com_](http://app.nanonets.com)\) between **11:58 and 12:09 UTC on Sept 22nd**. Other regions remained unaffected. The disruption was caused by an unexpected resource issue on one of our database nodes, which impacted application stability. Our team quickly identified the issue, applied a fix, and restored normal operations. We sincerely apologize for the inconvenience and we are implementing additional safeguards to prevent such issues in the future. Thank you for your understanding.
Clarification: No Platform-wide Downtime on September 4th
Timeline · 1 update
- resolved Sep 04, 2025, 05:51 AM UTC
At 1:00 AM UTC on September 4th, we posted an incident titled “Zero learning model file processing failures.” Upon further investigation, we identified that the issue was isolated to a single model and was not platform-wide. No other models or regions were impacted. The previously posted incident has been removed. We would like to clarify that there has been no platform-wide downtime in any region. Thank you for your understanding.
app.nanonets.com – User Logout Issue
Timeline · 1 update
- resolved Aug 20, 2025, 12:13 PM UTC
During a planned migration activity, users were unexpectedly logged out of the platform between 11:50 UTC and 12:05 UTC. The disruption occurred because database reads failed for old pods during the migration, while new pods came up late, leading to temporary session validation failures. The issue was resolved once the new pods were fully operational. We sincerely apologize for the inconvenience caused and will ensure future migrations are executed with better coordination to avoid service disruption.