Scaleway incident

[COCKPIT] - [fr-par/pl-waw] - Scaleway Metrics/Logs ingestion down

Notice Resolved View vendor source →

Scaleway experienced a notice incident on June 30, 2026 affecting Cockpit, lasting 8d 3h. The incident has been resolved; the full update timeline is below.

Started
Jun 30, 2026, 08:23 AM UTC
Resolved
Jul 08, 2026, 12:20 PM UTC
Duration
8d 3h
Detected by Pingoru
Jun 30, 2026, 08:23 AM UTC

Affected components

Cockpit

Update timeline

  1. monitoring Jun 30, 2026, 08:23 AM UTC

    Cockpit impacted this night betwen 1am30 and 8am30 with data loss for all Products metrics and logs on FR-PAR & PL-WAW Regions Root cause : based on last week Cockpit ingestion solution, a high RAM consumption saturated nodes and corrupted replay solution. Status : Cockpit fixed at 8am30 and fix in development to avoid replay corruption on OOM events

  2. monitoring Jul 03, 2026, 01:21 PM UTC

    Root cause was finally identified yesterday evening, we make an evolution on staging environments and monitored its behavior today. We plan to deploy this fix next monday morning. Metrics & Logs acquisition are fine since yesterday with a scaling workaround, we keep this solution for the week-end to not impact production on a friday afternoon.

  3. monitoring Jul 06, 2026, 04:17 PM UTC

    We applied and monitored today the last fix for high RAM consumption. All seems fine now on Cockpit performances for all regions.

  4. resolved Jul 08, 2026, 12:20 PM UTC

    This incident has been resolved.