Skip to main content

GitHub Copilot · 48d ago

Latency issues across a number of services

Resolved

Timeline

Started: Jul 23, 2026, 7:53 AM UTC

Last update: Jul 27, 2026, 9:51 PM UTC

Resolved: Jul 23, 2026, 9:39 AM UTC

Affected: 4230lsnqdsld, kr09ddfgbfsf, hhtssxt0f5v2, br0l2tvcx85d

GitHub Copilot's official report →

Updates

  • resolved

    On July 23, 2026, between 07:08 and 09:39 UTC, several services experienced delays: 8% of actions workflow runs experienced an average run start delay of 10 minutes, 5% of webhook deliveries exceeded SLO, and code scanning, repos, notifications, issues and pull requests experienced increased latency over the life of the incident. <br /><br />The root cause of the incident was a node of our background job processing system which did not recover after entering scheduled host maintenance. The incident was mitigated by identifying the problematic shard and restoring its correct state, after which queue backlogs drained and services recovered. <br /><br />To speed mitigation, we have added monitors for nodes in this unhealthy state after maintenance operations. To prevent future recurrence, we are adapting our lifecycle automation to verify host rejoin after a scheduled reboot.

    Jul 23, 2026, 9:39 AM UTC

  • monitoring

    The degradation has been mitigated. We are monitoring to ensure stability.

    Jul 23, 2026, 9:39 AM UTC

  • investigating

    Webhooks is operating normally.

    Jul 23, 2026, 9:35 AM UTC

  • investigating

    The degradation affecting Pull Requests has been mitigated. We are monitoring to ensure stability.

    Jul 23, 2026, 9:27 AM UTC

  • investigating

    We identified the source of latency affecting multiple services and applied a fix. Issues and Actions are recovering, and remaining affected services are seeing improvement as processing backlogs clear. We are actively monitoring recovery across all services.

    Jul 23, 2026, 9:22 AM UTC

  • investigating

    The degradation affecting Actions has been mitigated. We are monitoring to ensure stability.

    Jul 23, 2026, 9:19 AM UTC

  • investigating

    The degradation affecting Issues has been mitigated. We are monitoring to ensure stability.

    Jul 23, 2026, 9:18 AM UTC

  • investigating

    We're currently investigating latency across multiple services. This can show as Actions jobs taking longer to start, Issues search serving stale results, and other listed services being similarly impacted.

    Jul 23, 2026, 8:34 AM UTC

  • investigating

    Pull Requests is experiencing degraded performance. We are continuing to investigate.

    Jul 23, 2026, 8:25 AM UTC

  • investigating

    We are investigating reports of degraded availability for Actions, Issues and Webhooks

    Jul 23, 2026, 7:53 AM UTC

← All outage history