DigitalOcean Status · History · Incident #5485

RESOLVED

Serverless Inference - Gemma4 Latency Issues Causing Timeouts & Slow Responses

Minor · Started Jul 13, 2026 · 7:04 PM

  • Duration

    2d 2h 42m

  • Severity

    Minor

  • Detection lead

  • User reports

Summary

Serverless Inference - Gemma4 Latency Issues Causing Timeouts & Slow Responses

The deployed fix has successfully restored full functionality, and our monitoring shows that system performance has completely stabilized. Response times for all Gemma 4 inference workflows have returned to normal baseline levels. We will continue to track platform stability moving forward to ensure long-term reliability. We apologize for any disruption this may have caused to your workflows and appreciate your patience throughout the recovery process.


  • Started

    Jul 13, 2026 · 7:04 PM

  • Resolved

    Jul 15, 2026 · 9:46 PM

  • Duration

    2d 2h 42m

  • Severity

    Minor

Event timeline

How this incident unfolded

  • Identified

    Jul 13 · 7:04 PM DigitalOcean

    We are currently experiencing an issue where customers using the Gemma 4 model on our Serverless Inference and Dedicated Inference platforms may experience severe latency or request timeouts. Our engineering team has identified a backend configuration issue as the root cause, which is temporarily degrading performance. We are actively working on a fix to restore normal response times and will provide another update as soon as the mitigation is in place

  • Monitoring

    Jul 13 · 10:24 PM DigitalOcean

    A fix has been deployed to resolve the backend configuration issue. We are closely monitoring system performance to ensure full recovery and normal response times for all Gemma 4 inference workflows.

  • Identified

    Jul 14 · 11:07 AM DigitalOcean

    We are currently experiencing an issue affecting customers using the Gemma 4 model on our Serverless Inference platform. Customers may experience significantly increased latency or request timeouts. Our Engineering team has identified a backend configuration issue as the root cause, which is temporarily impacting model performance. Please be assured that our Engineering team is actively working on a fix and is treating this issue with high priority. We sincerely apologise for any inconvenience this may have caused and appreciate your patience and understanding. If you have any further questions, please create a support ticket so that we can investigate your specific case further.

  • Monitoring

    Jul 15 · 10:06 AM DigitalOcean

    A fix has been deployed to resolve the issue. We are closely monitoring system performance to ensure full recovery and normal response times for all Gemma 4 inference workflows.

  • Resolved

    Jul 15 · 9:46 PM DigitalOcean

    The deployed fix has successfully restored full functionality, and our monitoring shows that system performance has completely stabilized. Response times for all Gemma 4 inference workflows have returned to normal baseline levels. We will continue to track platform stability moving forward to ensure long-term reliability. We apologize for any disruption this may have caused to your workflows and appreciate your patience throughout the recovery process.

Get alerted before the next DigitalOcean outage.

Pulsetic catches degradations minutes before vendors acknowledge them.

Start monitoring free