Monte Carlo Data Outage History

Monte Carlo Data is up right now

Monte Carlo Data had 10 outages in the last 2 years totaling 559h 16m of downtime — averaging 0.4 incidents per month.

There were 10 Monte Carlo Data outages since March 30, 2026 totaling 559h 16m of downtime. Each is summarised below — incident details, duration, and resolution information.

Source: https://status.getmontecarlo.com

Minor October 3, 2026

False monitor run timeout errors

Detected by Pingoru
Oct 03, 2026, 06:15 PM UTC
Resolved
Oct 04, 2026, 10:25 AM UTC
Duration
16h 10m
Timeline · 1 update
  1. resolved Oct 03, 2026, 06:15 PM UTC

    Monte Carlo has resolved an issue that affected a subset of customers between 18:15 UTC on October 3 and 10:25 UTC on October 4. For affected customers, some monitor runs were incorrectly marked as "Timeout" and triggered failure notifications, even though in most cases the monitor queries had completed successfully. Those runs will continue to show as "Timeout" in run history, and the notifications can be dismissed; no action is needed from customers who did not see these errors.

Read the full incident report →

Minor September 4, 2026

Snowflake data collection issues (Snowflake EU region)

Detected by Pingoru
Sep 04, 2026, 07:00 AM UTC
Resolved
Sep 04, 2026, 08:30 AM UTC
Duration
1h 30m
Timeline · 2 updates
  1. investigating Sep 04, 2026, 07:00 AM UTC

    We’re investigating connectivity issues affecting Snowflake instances in the EU region. Anomaly detection is affected for accounts hosted in Monte Carlo's EU region, as well as customers with Snowflake integrations hosted in Snowflake EU. Metadata collection may be delayed or fail for accounts connecting to a Snowflake instance hosted in the EU, regardless of their Monte Carlo region. For more information, please refer to Snowflake’s status page: https://status.snowflake.com/incidents/fl9tkpj8d7hk

  2. resolved Sep 04, 2026, 11:08 AM UTC

    The Snowflake incident has been resolved, and data collection and anomaly detection are back to normal on all affected accounts.

Read the full incident report →

Minor August 29, 2026

Data health updates to Alation catalogs not being sent

Detected by Pingoru
Aug 29, 2026, 12:45 AM UTC
Resolved
Sep 07, 2026, 08:00 PM UTC
Duration
9d 19h
Timeline · 2 updates
  1. monitoring Jul 29, 2026, 12:45 AM UTC

    Since August 29, some Alation integrations have stopped receiving data health updates from Monte Carlo after their stored credentials stopped being accepted. Monitoring and alerting inside Monte Carlo were not affected. A fix is now live. Integrations configured with a username and password have reconnected automatically. Integrations using a refresh token need a new token generated in Alation and saved in Monte Carlo will need to update their refresh token.

  2. resolved Sep 08, 2026, 01:00 PM UTC

    A fix was deployed on September 7 and integrations configured with a username and password reconnected automatically — data health updates to their Alation catalogs have resumed. The remaining affected integrations are configured with a refresh token only, which requires the customer to generate a new token in Alation and save it in Monte Carlo. Those customers have been notified directly with instructions, and their updates will resume as soon as the new token is saved. No Monte Carlo monitoring or alerting was affected at any point.

Read the full incident report →

Minor August 28, 2026

Application is down when accessed from getmontecarlo.com

Detected by Pingoru
Aug 28, 2026, 05:45 AM UTC
Resolved
Aug 28, 2026, 06:29 AM UTC
Duration
44m
Timeline · 2 updates
  1. investigating Aug 28, 2026, 06:25 AM UTC

    Browser fails to load getmontecarlo.com. To access the app while we resolve the issue, please use the link getmontecarlo.com/signin

  2. resolved Aug 28, 2026, 06:30 AM UTC

    The app is now available at getmontecarlo.com as usual.

Read the full incident report →

Minor June 28, 2026

MS Teams notifications failing for some accounts

Detected by Pingoru
Jun 28, 2026, 11:00 AM UTC
Resolved
Jun 28, 2026, 12:30 PM UTC
Duration
1h 30m
Timeline · 2 updates
  1. investigating Jun 28, 2026, 11:00 AM UTC

    MS Teams notifications are not working for accounts using the Monte Carlo app from the Microsoft Teams Store. Accounts on the legacy integration (non-app) are unaffected.

  2. resolved Jun 28, 2026, 01:50 PM UTC

    MS Teams notifications are working again.

Read the full incident report →

Minor June 25, 2026

Snowflake query logs delayed

Detected by Pingoru
Jun 25, 2026, 03:53 PM UTC
Resolved
Jun 27, 2026, 03:30 AM UTC
Duration
1d 11h
Timeline · 2 updates
  1. monitoring Jun 26, 2026, 03:53 PM UTC

    Query log collection is delayed by lag in Snowflake's account_usage metadata. We're tracking recovery.

  2. resolved Jun 27, 2026, 12:25 PM UTC

    Query log collection is normalized.

Read the full incident report →

Minor May 8, 2026

Databricks SQL outage impacting Databricks integration (AWS us-east-1)

Detected by Pingoru
May 08, 2026, 12:00 AM UTC
Resolved
May 09, 2026, 02:30 AM UTC
Duration
1d 2h
Timeline · 2 updates
  1. monitoring May 08, 2026, 12:00 AM UTC

    A Databricks incident (~7 May 2026, 23:00 UTC) affected Databricks SQL warehouses and Unity Catalog, causing metadata collection delays and monitor evaluation failures or timeouts for a subset of customers using the Databricks integration. The upstream incident is now resolved on Databricks' side, but some customer warehouses are still working through query backlog accumulated during the incident, so individual customers may continue to see elevated runtimes or sporadic timeouts until their warehouse capacity returns to baseline.

  2. resolved May 09, 2026, 04:22 PM UTC

    As of May 9, 2026 02:32 UTC, Databricks has marked the incident as Resolved. Databricks SQL warehouses and Unity Catalog access have stabilized, and metadata collection and monitor evaluations are operating normally for all customers using the Databricks integration.

Read the full incident report →

Minor April 14, 2026

Email notifications may be quarantined by third-party security gateways

Detected by Pingoru
Apr 14, 2026, 12:00 AM UTC
Resolved
Apr 24, 2026, 12:00 AM UTC
Duration
10d
Timeline · 1 update
  1. identified Apr 14, 2026, 12:00 AM UTC

    Some Monte Carlo email notifications may be quarantined by third-party email security gateways (e.g., Proofpoint) before reaching recipients, despite successful delivery from our platform. We recommend allowlisting mail.getmontecarlo.com on your email security tool to ensure reliable delivery.

Read the full incident report →

Major March 30, 2026

Anomaly detection for legacy metric monitors (v1) was temporarily delayed due to a Databricks outage on AWS.

Detected by Pingoru
Mar 30, 2026, 09:15 PM UTC
Resolved
Mar 30, 2026, 11:15 PM UTC
Duration
2h
Timeline · 1 update
  1. resolved Mar 30, 2026, 09:15 PM UTC

    Due to a Databricks outage on AWS, anomaly detection for legacy metric monitors (v1) was temporarily delayed. Data collection and all other monitoring capabilities were not affected. The underlying issue has been resolved by Databricks, and all affected processes have been re-run and completed successfully. Normal operation has fully resumed.

Read the full incident report →