Replicated Outage History

Replicated is up right now

Replicated had 35 outages in the last 2 years totaling 296h 41m of downtime — averaging 1.4 incidents per month.

There were 35 Replicated outages since October 21, 2024 totaling 296h 41m of downtime. Each is summarised below — incident details, duration, and resolution information.

Source: https://status.replicated.com

Minor June 29, 2025

Some AKS instance types are unavailable

Detected by Pingoru
Jun 29, 2025, 09:47 PM UTC
Resolved
Jun 30, 2025, 02:29 AM UTC
Duration
4h 41m
Affected: AKS
Timeline · 2 updates
  1. identified Jun 29, 2025, 09:47 PM UTC

    AKS instance types Standard_D2S_v5, Standard_D4S_v5, Standard_D8S_v5, Standard_D16S_v5, Standard_D32S_v5, and Standard_D48S_v5 are currently unavailable. We are working with Azure to correct this issue.

  2. resolved Jun 30, 2025, 02:29 AM UTC

    The incident has been resolved

Read the full incident report →

Minor June 12, 2025

High error rate from upstream cloud provider

Detected by Pingoru
Jun 12, 2025, 06:23 PM UTC
Resolved
Jun 12, 2025, 09:31 PM UTC
Duration
3h 7m
Affected: GKEReplicated Registry
Timeline · 5 updates
  1. monitoring Jun 12, 2025, 06:23 PM UTC

    We are aware of upstream issues with Google Cloud Platform. This is presently impacting GKE components of our Compatibility Matrix product, and upstream GCR registries with our Replicated Registry. We are monitoring our cloud providers for updates.

  2. monitoring Jun 12, 2025, 07:24 PM UTC

    We are continuing to monitor upstream issues with Google Cloud Platform and Cloudflare

  3. monitoring Jun 12, 2025, 09:29 PM UTC

    We are continuing to monitor for any further issues.

  4. monitoring Jun 12, 2025, 09:30 PM UTC

    We are continuing to monitor for any further issues.

  5. resolved Jun 12, 2025, 09:31 PM UTC

    This incident has been resolved.

Read the full incident report →

Minor May 29, 2025

CMX: Network Issues when creating RKE2 clusters

Detected by Pingoru
May 29, 2025, 05:55 PM UTC
Resolved
Jun 09, 2025, 07:18 PM UTC
Duration
11d 1h
Affected: RKE2
Timeline · 2 updates
  1. investigating May 29, 2025, 05:55 PM UTC

    We're aware of network issues when creating an RKE2 cluster and are investigating

  2. resolved Jun 09, 2025, 07:18 PM UTC

    This incident has been resolved.

Read the full incident report →

Minor March 31, 2025

Compatibility Matrix: GKE 1.32 clusters failing to create

Detected by Pingoru
Mar 31, 2025, 10:31 PM UTC
Resolved
Apr 01, 2025, 03:09 AM UTC
Duration
4h 37m
Affected: GKE
Timeline · 5 updates
  1. investigating Mar 31, 2025, 10:31 PM UTC

    We have noticed that GKE 1.32 clusters are not reaching a 'running' state. We are currently investigating this issue.

  2. identified Mar 31, 2025, 11:39 PM UTC

    We have identified an issue with pulling images from Google's registry, and are working with Google to find a resolution.

  3. monitoring Apr 01, 2025, 02:43 AM UTC

    We have rolled out a fix, pinning GKE 1.32 to version 1.32.2, and are monitoring the results now.

  4. monitoring Apr 01, 2025, 02:44 AM UTC

    We are continuing to monitor for any further issues.

  5. resolved Apr 01, 2025, 03:09 AM UTC

    This incident has been resolved.

Read the full incident report →

Notice March 24, 2025

Email deliverability spam issue

Detected by Pingoru
Mar 24, 2025, 04:30 PM UTC
Resolved
Mar 25, 2025, 02:26 AM UTC
Duration
9h 56m
Affected: Vendor Portal
Timeline · 3 updates
  1. identified Mar 24, 2025, 04:30 PM UTC

    Check your spam folder. Our transactional email provider is seeing deliverability problems with gmail. https://status.postmarkapp.com/notices/bt3ky3r8zlaapqlo-increased-gmail-spam-reports

  2. monitoring Mar 24, 2025, 11:27 PM UTC

    A fix has been implemented and we are monitoring the results.

  3. resolved Mar 25, 2025, 02:26 AM UTC

    This incident has been resolved.

Read the full incident report →

Minor March 17, 2025

Issues provisioning CMX multi-node vm clusters

Detected by Pingoru
Mar 17, 2025, 11:49 PM UTC
Resolved
Mar 18, 2025, 01:57 AM UTC
Duration
2h 8m
Affected: OpenShiftkURLk3skindOKE - Alpha
Timeline · 4 updates
  1. investigating Mar 17, 2025, 11:49 PM UTC

    We are currently investigating this issue.

  2. investigating Mar 18, 2025, 12:08 AM UTC

    We are continuing to investigate this issue.

  3. monitoring Mar 18, 2025, 01:49 AM UTC

    A fix has been implemented and we are monitoring the results.

  4. resolved Mar 18, 2025, 01:57 AM UTC

    This incident has been resolved.

Read the full incident report →

Major March 12, 2025

Issues provisioning k3s and kind clusters in CMX

Detected by Pingoru
Mar 12, 2025, 09:33 PM UTC
Resolved
Mar 12, 2025, 11:40 PM UTC
Duration
2h 6m
Affected: k3skind
Timeline · 4 updates
  1. investigating Mar 12, 2025, 09:33 PM UTC

    We are currently investigating issues with provisioning k3s CMX clusters leading to timeouts

  2. investigating Mar 12, 2025, 09:35 PM UTC

    This issue is likely impacting kind cluster creation as well

  3. monitoring Mar 12, 2025, 09:43 PM UTC

    We are seeing k3s and kind CMX clusters spinning up successfully at this time without timing out. We are continuing to monitor this situation.

  4. resolved Mar 12, 2025, 11:40 PM UTC

    This incident has been resolved.

Read the full incident report →

Minor February 24, 2025

CMX: Openshift connectivity issues

Detected by Pingoru
Feb 24, 2025, 07:47 PM UTC
Resolved
Feb 24, 2025, 08:16 PM UTC
Duration
28m
Affected: OpenShift
Timeline · 2 updates
  1. investigating Feb 24, 2025, 07:47 PM UTC

    We are currently investigating issues with name resolution in some containers in OpenShift clusters, in Compatibility Matrix

  2. resolved Feb 24, 2025, 08:16 PM UTC

    We have run diagnostics across the OpenShift cluster and found the source of the report, however it appears that the limitation is built into OpenShift by design and working as intended.

Read the full incident report →

Minor February 3, 2025

Degraded Performance for Embedded Cluster and kURL in CMX

Detected by Pingoru
Feb 03, 2025, 03:04 PM UTC
Resolved
Feb 03, 2025, 06:32 PM UTC
Duration
3h 27m
Affected: kURLk3skindEmbedded Clusters
Timeline · 6 updates
  1. investigating Feb 03, 2025, 03:04 PM UTC

    We are currently investigating this issue

  2. investigating Feb 03, 2025, 03:14 PM UTC

    We are continuing to investigate this issue.

  3. investigating Feb 03, 2025, 04:13 PM UTC

    We are continuing to investigate this issue.

  4. identified Feb 03, 2025, 04:38 PM UTC

    We believe we have identified the cause of these issues and are working to remediate them.

  5. monitoring Feb 03, 2025, 06:00 PM UTC

    We have deployed changes to help mitigate issues creating k3s, kind, kURL, and Embedded Clusters on CMX. We are currently monitoring these changes.

  6. resolved Feb 03, 2025, 06:32 PM UTC

    Our monitoring confirms that remediation efforts are now showing successful k3s, kind, kURL, and Embedded Cluster in CMX.

Read the full incident report →

Minor October 21, 2024

Increased 5xx rate for Vendor API

Detected by Pingoru
Oct 21, 2024, 04:51 PM UTC
Resolved
Oct 21, 2024, 05:35 PM UTC
Duration
43m
Affected: Vendor Portal
Timeline · 3 updates
  1. identified Oct 21, 2024, 04:51 PM UTC

    We have identified the issues causing a higher than expected 5xx rate for the Vendor API. We have mitigated this by rolling back some recent changes and are working on a long-term solution.

  2. monitoring Oct 21, 2024, 05:29 PM UTC

    A long-term fix for this issue has been released and we are currently monitoring.

  3. resolved Oct 21, 2024, 05:35 PM UTC

    This issue should now be resolved

Read the full incident report →