Groq Status · History · Incident #5076

RESOLVED

Data Center Failure Impacting Capacity

Minor · Started Jul 1, 2026 · 11:01 PM

  • Duration

    1h 33m

  • Severity

    Minor

  • Detection lead

  • User reports

Summary

Data Center Failure Impacting Capacity

The **capacity issue has been resolved**. All services are operating at normal capacity. The incident was caused by a **power loss issue that led to a cooling system failure at one of our US Central data centers**. We apologize for any disruption and **appreciate your patience**.


  • Started

    Jul 1, 2026 · 11:01 PM

  • Resolved

    Jul 2, 2026 · 12:35 AM

  • Duration

    1h 33m

  • Severity

    Minor

Event timeline

How this incident unfolded

  • Investigating

    Jul 1 · 11:01 PM Groq

    We are currently investigating a potential issue at one of our US Central data centers that is impacting system capacity. Users **may experience higher latencies** due to reduced resources. Our engineering team is working to identify the root cause and remediate the problem.

  • Identified

    Jul 1 · 11:11 PM Groq

    We have **identified a cooling system failure** at one of our **US Central** data centers that is causing **reduced capacity**. The team is **working on restoring capacity**. Users **continue to see elevated latencies for specific models**.

  • Identified

    Jul 1 · 11:27 PM Groq

    A power loss issue at approximately 22:20pm UTC led to a subsequent **cooling system failure** at one of our **US Central** data centers that is causing **reduced capacity**, primarily for the `openai/gpt-oss-20b` model. The team is **working on restoring capacity**.

  • Monitoring

    Jul 1 · 11:48 PM Groq

    We have **redistributed and allocated more capacity to production models** in **the US Central region.** Users should **see performance return to normal**. We’re now **monitoring** the infrastructure to ensure stability.

  • Resolved

    Jul 2 · 12:35 AM Groq

    The **capacity issue has been resolved**. All services are operating at normal capacity. The incident was caused by a **power loss issue that led to a cooling system failure at one of our US Central data centers**. We apologize for any disruption and **appreciate your patience**.

Get alerted before the next Groq outage.

Pulsetic catches degradations minutes before vendors acknowledge them.

Start monitoring free