Latency issues across a number of services


Incident resolved in 1h45m20s

Resolved

On July 23, 2026, between 07:08 and 09:39 UTC, several services experienced delays: 8% of actions workflow runs experienced an average run start delay of 10 minutes, 5% of webhook deliveries exceeded SLO, and code scanning, repos, notifications, issues and pull requests experienced increased latency over the life of the incident. The root cause of the incident was a node of our background job processing system which did not recover after entering scheduled host maintenance. The incident was mitigated by identifying the problematic shard and restoring its correct state, after which queue backlogs drained and services recovered. To speed mitigation, we have added monitors for nodes in this unhealthy state after maintenance operations. To prevent future recurrence, we are adapting our lifecycle automation to verify host rejoin after a scheduled reboot.

1784799559

Update

The degradation has been mitigated. We are monitoring to ensure stability.

1784799551

Investigating

Webhooks is operating normally.

1784799317

Investigating

The degradation affecting Pull Requests has been mitigated. We are monitoring to ensure stability.

1784798855

Investigating

We identified the source of latency affecting multiple services and applied a fix. Issues and Actions are recovering, and remaining affected services are seeing improvement as processing backlogs clear. We are actively monitoring recovery across all services.

1784798541

Investigating

The degradation affecting Actions has been mitigated. We are monitoring to ensure stability.

1784798389

Investigating

The degradation affecting Issues has been mitigated. We are monitoring to ensure stability.

1784798317

Investigating

We're currently investigating latency across multiple services. This can show as Actions jobs taking longer to start, Issues search serving stale results, and other listed services being similarly impacted.

1784795663

Investigating

Pull Requests is experiencing degraded performance. We are continuing to investigate.

1784795116

Investigating

We are investigating reports of degraded availability for Actions, Issues and Webhooks

1784793239