Skip to content

Incident detail

Incident with CodeQL, Webhooks, Notifications, and Slack Integration

Resolved incidentMinor2 affected services

Timeline window

to

Outage alerts

Get alerted the next time GitHub breaks

Free email alerts for up to 5 providers. No card, live in about a minute. Paid plans add Slack, Teams, Discord, and webhook delivery across your whole stack, plus higher API quotas.

Timeline

Incident updates

Every update GitHub posted, oldest to newest, exactly as it appeared on their official status page.

  1. Investigating

    We are investigating reports of degraded performance for CodeQL

  2. Investigating

    CodeQL actions are currently experiencing delays, which may result in those actions being stuck in a pending state or having failed due to a timeout.

  3. Investigating

    We're continuing to investigate issues with CodeQL actions workflows. We're additionally seeing delays for notifications, webhooks, and the Slack integration.

  4. Investigating

    Webhooks is experiencing degraded performance. We are continuing to investigate.

  5. Investigating

    We've established that most delays are related to a queuing service and are working to scale out. Early signals from the scale-out are showing signs of recovery for some services. We'll provide an update when services are fully recovered.

  6. Investigating

    Webhooks is operating normally.

  7. Investigating

    Webhooks have fully recovered. Continuing to work on recovery for the other services.

  8. Investigating

    CodeQL has fully recovered. We're continuing to work on recovery for the remaining impacted services.

  9. Investigating

    All services have fully recovered.

  10. Resolved

    On May 12, 2026, between 13:41 and 17:43 UTC, some services experienced delays in processing. For the Code Scanning service, 53% of check runs took over 15 minutes to complete. Additionally, notifications took an average of 22 minutes to be delivered and Slack integration webhooks took an average of 20 minutes to be delivered. The delays were caused by replication lag due to an internal database migration, resulting in insufficient worker capacity for our high rate of job enqueues. <br /><br />We mitigated the impact by scaling our processing workers to handle the increased load. All services returned to normal processing times after the mitigation was applied. <br /><br />We are working to create dedicated worker pools for some of our high usage shared queues to help prevent this in the future.

Keep exploring

More from GitHub

Neighboring incidents on GitHub's timeline and the rest of their record on OutageDeck.