groq · 74d ago
Data Center Failure Impacting Capacity
ResolvedTimeline
Started: Jul 1, 2026, 11:01 PM UTC
Last update: Jul 2, 2026, 12:35 AM UTC
Resolved: Jul 2, 2026, 12:35 AM UTC
Updates
resolved
The **capacity issue has been resolved**. All services are operating at normal capacity. The incident was caused by a **power loss issue that led to a cooling system failure at one of our US Central data centers**. We apologize for any disruption and **appreciate your patience**.
Jul 2, 2026, 12:35 AM UTC
monitoring
We have **redistributed and allocated more capacity to production models** in **the US Central region.** Users should **see performance return to normal**. We’re now **monitoring** the infrastructure to ensure stability.
Jul 1, 2026, 11:48 PM UTC
identified
A power loss issue at approximately 22:20pm UTC led to a subsequent **cooling system failure** at one of our **US Central** data centers that is causing **reduced capacity**, primarily for the `openai/gpt-oss-20b` model. The team is **working on restoring capacity**.
Jul 1, 2026, 11:27 PM UTC
identified
We have **identified a cooling system failure** at one of our **US Central** data centers that is causing **reduced capacity**. The team is **working on restoring capacity**. Users **continue to see elevated latencies for specific models**.
Jul 1, 2026, 11:11 PM UTC
investigating
We are currently investigating a potential issue at one of our US Central data centers that is impacting system capacity. Users **may experience higher latencies** due to reduced resources. Our engineering team is working to identify the root cause and remediate the problem.
Jul 1, 2026, 11:01 PM UTC