DigitalOcean incident
Serverless Inference - Gemma4 Latency Issues Causing Timeouts & Slow Responses
DigitalOcean experienced a minor incident on July 13, 2026 affecting Inference, lasting 2d 2h. The incident has been resolved; the full update timeline is below.
Affected components
Update timeline
- identified Jul 13, 2026, 07:04 PM UTC
We are currently experiencing an issue where customers using the Gemma 4 model on our Serverless Inference and Dedicated Inference platforms may experience severe latency or request timeouts. Our engineering team has identified a backend configuration issue as the root cause, which is temporarily degrading performance. We are actively working on a fix to restore normal response times and will provide another update as soon as the mitigation is in place
- monitoring Jul 13, 2026, 10:24 PM UTC
A fix has been deployed to resolve the backend configuration issue. We are closely monitoring system performance to ensure full recovery and normal response times for all Gemma 4 inference workflows.
- identified Jul 14, 2026, 11:07 AM UTC
We are currently experiencing an issue affecting customers using the Gemma 4 model on our Serverless Inference platform. Customers may experience significantly increased latency or request timeouts. Our Engineering team has identified a backend configuration issue as the root cause, which is temporarily impacting model performance. Please be assured that our Engineering team is actively working on a fix and is treating this issue with high priority. We sincerely apologise for any inconvenience this may have caused and appreciate your patience and understanding. If you have any further questions, please create a support ticket so that we can investigate your specific case further.
- monitoring Jul 15, 2026, 10:06 AM UTC
A fix has been deployed to resolve the issue. We are closely monitoring system performance to ensure full recovery and normal response times for all Gemma 4 inference workflows.
- resolved Jul 15, 2026, 09:46 PM UTC
The deployed fix has successfully restored full functionality, and our monitoring shows that system performance has completely stabilized. Response times for all Gemma 4 inference workflows have returned to normal baseline levels. We will continue to track platform stability moving forward to ensure long-term reliability. We apologize for any disruption this may have caused to your workflows and appreciate your patience throughout the recovery process.