DigitalOcean incident

Serverless Inference - Gemma4 Latency Issues Causing Timeouts & Slow Responses

Minor Resolved View vendor source →

DigitalOcean experienced a minor incident on July 13, 2026 affecting Inference, lasting 2d 2h. The incident has been resolved; the full update timeline is below.

Started
Jul 13, 2026, 07:04 PM UTC
Resolved
Jul 15, 2026, 09:46 PM UTC
Duration
2d 2h
Detected by Pingoru
Jul 13, 2026, 07:04 PM UTC

Affected components

Inference

Update timeline

  1. identified Jul 13, 2026, 07:04 PM UTC

    We are currently experiencing an issue where customers using the Gemma 4 model on our Serverless Inference and Dedicated Inference platforms may experience severe latency or request timeouts. Our engineering team has identified a backend configuration issue as the root cause, which is temporarily degrading performance. We are actively working on a fix to restore normal response times and will provide another update as soon as the mitigation is in place

  2. monitoring Jul 13, 2026, 10:24 PM UTC

    A fix has been deployed to resolve the backend configuration issue. We are closely monitoring system performance to ensure full recovery and normal response times for all Gemma 4 inference workflows.

  3. identified Jul 14, 2026, 11:07 AM UTC

    We are currently experiencing an issue affecting customers using the Gemma 4 model on our Serverless Inference platform. Customers may experience significantly increased latency or request timeouts. Our Engineering team has identified a backend configuration issue as the root cause, which is temporarily impacting model performance. Please be assured that our Engineering team is actively working on a fix and is treating this issue with high priority. We sincerely apologise for any inconvenience this may have caused and appreciate your patience and understanding. If you have any further questions, please create a support ticket so that we can investigate your specific case further.

  4. monitoring Jul 15, 2026, 10:06 AM UTC

    A fix has been deployed to resolve the issue. We are closely monitoring system performance to ensure full recovery and normal response times for all Gemma 4 inference workflows.

  5. resolved Jul 15, 2026, 09:46 PM UTC

    The deployed fix has successfully restored full functionality, and our monitoring shows that system performance has completely stabilized. Response times for all Gemma 4 inference workflows have returned to normal baseline levels. We will continue to track platform stability moving forward to ensure long-term reliability. We apologize for any disruption this may have caused to your workflows and appreciate your patience throughout the recovery process.