Skip to content

Incident detail

Data Center Failure Impacting Capacity - DMM1

Resolved incidentMinor

Timeline window

to

Outage alerts

Get alerted the next time Groq breaks

Free email alerts for up to 5 providers. No card, live in about a minute. Paid plans add Slack, Teams, Discord, and webhook delivery across your whole stack, plus higher API quotas.

Timeline

Incident updates

Every update Groq posted, oldest to newest, exactly as it appeared on their official status page.

  1. Resolved

    The data center capacity issue has been resolved. All services are operating at normal capacity. The incident was caused by a failed ARP entry on the network gateway device caused the Salam network link to go down in the DMM1 data center, which has been addressed. We apologize for the disruption and appreciate your patience.

  2. Monitoring

    We have restored capacity in DMM 1 and services are beginning to recover. Users should gradually see network performance return to normal. We’re now monitoring the infrastructure to ensure stability.

  3. Identified

    We have identified a failure in our DMM1 infrastructure that is causing reduced capacity and network latency. The team has isolated the cause to the Salam link confirmed down and the Mobily link operating at full capacity, resulting in significant packet loss and is working on restoring capacity. Users continue to see Network latency.

  4. Investigating

    We are currently investigating a potential network issue at our DMM1 that is impacting system capacity. Users may experience slower response times or intermittent failures due to reduced resources. Our engineering team is working to identify the root cause and remediate the problem.

Keep exploring

More from Groq

Neighboring incidents on Groq's timeline and the rest of their record on OutageDeck.