Skip to main content

groq · 74d ago

Data Center Failure Impacting Capacity

Resolved

Timeline

Started: Jul 1, 2026, 11:01 PM UTC

Last update: Jul 2, 2026, 12:35 AM UTC

Resolved: Jul 2, 2026, 12:35 AM UTC

Updates

  • resolved

    The **capacity issue has been resolved**. All services are operating at normal capacity. The incident was caused by a **power loss issue that led to a cooling system failure at one of our US Central data centers**. We apologize for any disruption and **appreciate your patience**.

    Jul 2, 2026, 12:35 AM UTC

  • monitoring

    We have **redistributed and allocated more capacity to production models** in **the US Central region.** Users should **see performance return to normal**. We’re now **monitoring** the infrastructure to ensure stability.

    Jul 1, 2026, 11:48 PM UTC

  • identified

    A power loss issue at approximately 22:20pm UTC led to a subsequent **cooling system failure** at one of our **US Central** data centers that is causing **reduced capacity**, primarily for the `openai/gpt-oss-20b` model. The team is **working on restoring capacity**.

    Jul 1, 2026, 11:27 PM UTC

  • identified

    We have **identified a cooling system failure** at one of our **US Central** data centers that is causing **reduced capacity**. The team is **working on restoring capacity**. Users **continue to see elevated latencies for specific models**.

    Jul 1, 2026, 11:11 PM UTC

  • investigating

    We are currently investigating a potential issue at one of our US Central data centers that is impacting system capacity. Users **may experience higher latencies** due to reduced resources. Our engineering team is working to identify the root cause and remediate the problem.

    Jul 1, 2026, 11:01 PM UTC

← All outage history