Siit incident

Investigate high latency on bot response

Minor Resolved View vendor source →

Siit experienced a minor incident on October 8, 2026 affecting Customer Portal and Public API & MCP and 1 more component, lasting 2h 12m. The incident has been resolved; the full update timeline is below.

Started
Oct 08, 2026, 07:31 AM UTC
Resolved
Oct 08, 2026, 09:43 AM UTC
Duration
2h 12m
Detected by Pingoru
Oct 08, 2026, 07:31 AM UTC

Affected components

Customer PortalPublic API & MCPAdmin DashboardAI Agents

Update timeline

  1. investigating Oct 08, 2026, 07:31 AM UTC

    Status: Investigating We are receiving multiple reports and alerts of issues affecting our app Affected components Admin Dashboard (Degraded performance) Public API & MCP (Degraded performance) AI Agents (Partial outage) Customer Portal (Degraded performance)

  2. investigating Oct 08, 2026, 07:42 AM UTC

    Status: Investigating We have pushed a fix and expect recovery to happen during the next 10-40 minutes Affected components Admin Dashboard (Degraded performance) Public API & MCP (Degraded performance) AI Agents (Partial outage) Customer Portal (Degraded performance)

  3. monitoring Oct 08, 2026, 07:54 AM UTC

    Status: Monitoring Our fix appears to be working - our process are still catching their queue so you could still expect some latency in the delivery of pilled up notifications. We’ll come up with a detailed reports once we confirm that things settled down for good Affected components Public API & MCP (Operational) AI Agents (Degraded performance) Customer Portal (Degraded performance) Admin Dashboard (Operational)

  4. monitoring Oct 08, 2026, 08:50 AM UTC

    Status: Monitoring We are still having impactful latency until we reach full recovery Affected components Customer Portal (Degraded performance) Admin Dashboard (Degraded performance) Public API & MCP (Degraded performance) AI Agents (Degraded performance)

  5. monitoring Oct 08, 2026, 09:01 AM UTC

    Status: Monitoring All our monitoring data now appears back to normal, we will continue monitoring and provide a post mortem shortly. Affected components AI Agents (Operational) Customer Portal (Operational) Admin Dashboard (Operational) Public API & MCP (Operational)

  6. resolved Oct 08, 2026, 09:43 AM UTC

    Status: Resolved The issue is now fixed, the postmortem of the incident is available below. -- Summary On the morning of October 8, Siit stopped processing most of our background tasks between about 07:45 and 09:40 CEST. During that time, notifications, emails, Slack and Teams messages, workflows and integration syncs were delayed. Some interactive actions, such as clicking a button or talking to the AI Agent in a Slack or Teams message, couldn't be completed and may have needed a retry. While we caught up on the delayed work that had piled up, the app was slower than usual until about 11:00 CEST. All delayed work was queued safely and processed once service was restored. No data was lost. What happened A routine configuration cleanup after a recent release left our background processing unable to restart during scaling phases. The effect built up gradually overnight and became visible in the morning when activity picked up. How we resolved it We found the cause a few minutes after the investigation started and restored processing, handling the most time-sensitive tasks first. Catching up on the backlog put heavy load on our database, so we added database capacity and spread out the remaining work. Everything was back to normal by 11:00 CEST. Where we failed We had no alert for this specific failure, and the alerts we did have relied on the affected system itself As a result, the problem went unnoticed until it impacted you. What we're changing • We are revamping our Alerting, so such cases are caught right away. • Stricter safeguards and automated checks on production configuration changes. • We’ve upgraded our database capacity • Improving out-of-hours alerting and escalation We're sorry for the disruption to your teams. If an action from that morning still looks incomplete, retrying it should work. If not, our support team is happy to help. Affected components Public API & MCP (Operational) AI Agents (Operational) Customer Portal (Operational) Admin Dashboard (Operational)