Jul 24, 17:36 UTC
Resolved - On July 24th at 16:04 UTC, a loss of connectivity occurred in network paths in one of our three physical data center availability zones (AZs). This resulted in packet loss due to the remaining active paths becoming saturated. Our data centers use a leaf-spine switch fabric in each compute cage, and an aggregation layer interconnecting the spines from each cage within each AZ. The loss of connectivity affected links between one cage’s spine switches and the aggregation layer within that specific AZ. Workloads depending on compute resources in this cage became degraded due to packet loss, and exhibited intermittent errors: - Actions saw 10% of jobs fail during the impact window, and 5% of jobs succeeded but with delayed starts. - 27% of GitHub issues interactions saw slow requests or timeouts. - 4% of GitHub Copilot requests experienced errors, though most automatically retry. - 4% of git push operations saw impacts during the affected window. - Authentication requests saw increased latency during the affected window, but error rates, while elevated, were < 1% in all cases. We were able to mitigate the outage by re-routing affected connections to available fiber paths that were allocated for future capacity upgrades. Sufficient network capacity to eliminate packet loss was restored at 17:07, with most services showing full recovery by 17:16. All paths were restored and services healthy at 17:36. This incident affected 25% of available network interconnect capacity. Older cages utilize a 100Gbps network interface standard. To remove risk of reoccurrence, a planned upgrade to 400Gbps interfaces is being accelerated as much as possible, ensuring increased bandwidth available at all layers of the switch fabric for resiliency to path or device loss. Jul 24, 17:24 UTC
Update - We are seeing recovery across all services Jul 24, 17:16 UTC
Update - The degradation affecting API Requests, Actions, Copilot, Issues, Pages and Pull Requests has been mitigated. We are monitoring to ensure stability. Jul 24, 16:41 UTC
Update - Actions is experiencing degraded performance. We are continuing to investigate. Jul 24, 16:40 UTC
Update - We have applied a mitigation and are monitoring for recovery Jul 24, 16:28 UTC
Update - Actions is experiencing degraded availability. We are continuing to investigate. Jul 24, 16:27 UTC
Update - Pages is experiencing degraded performance. We are continuing to investigate. Jul 24, 16:26 UTC
Update - Copilot is experiencing degraded performance. We are continuing to investigate. Jul 24, 16:22 UTC
Update - We are investigating timeouts to some GitHub services Jul 24, 16:20 UTC
Update - Pull Requests is experiencing degraded performance. We are continuing to investigate. Jul 24, 16:19 UTC
Update - Actions is experiencing degraded performance. We are continuing to investigate. Jul 24, 16:17 UTC
Investigating - We are investigating reports of degraded performance for API Requests and Issues
On July 24th at 16:04 UTC, a loss of connectivity occurred in network paths in one of our three physical data center availability zones (AZs). This resulted in packet loss due to the remaining active paths becoming saturated. Our data centers use a leaf-spine switch fabric in each compute cage, and an aggregation layer interconnecting the spines from each cage within each AZ. The loss of connectivity affected links between one cage’s spine switches and the aggregation layer within that specific AZ.
Workloads depending on compute resources in this cage became degraded due to packet loss, and exhibited intermittent errors:
- Actions saw 10% of jobs fail during the impact window, and 5% of jobs succeeded but with delayed starts.
- 27% of GitHub issues interactions saw slow requests or timeouts.
- 4% of GitHub Copilot requests experienced errors, though most automatically retry.
- 4% of git push operations saw impacts during the affected window.
- Authentication requests saw increased latency during the affected window, but error rates, while elevated, were < 1% in all cases.
We were able to mitigate the outage by re-routing affected connections to available fiber paths that were allocated for future capacity upgrades. Sufficient network capacity to eliminate packet loss was restored at 17:07, with most services showing full recovery by 17:16. All paths were restored and services healthy at 17:36.
This incident affected 25% of available network interconnect capacity. Older cages utilize a 100Gbps network interface standard. To remove risk of reoccurrence, a planned upgrade to 400Gbps interfaces is being accelerated as much as possible, ensuring increased bandwidth available at all layers of the switch fabric for resiliency to path or device loss.
Posted Jul 24, 2026 - 17:36 UTC
We are seeing recovery across all services
Posted Jul 24, 2026 - 17:24 UTC
The degradation affecting API Requests, Actions, Copilot, Issues, Pages and Pull Requests has been mitigated. We are monitoring to ensure stability.
Posted Jul 24, 2026 - 17:16 UTC
Actions is experiencing degraded performance. We are continuing to investigate.
Posted Jul 24, 2026 - 16:41 UTC
We have applied a mitigation and are monitoring for recovery
Posted Jul 24, 2026 - 16:40 UTC
Actions is experiencing degraded availability. We are continuing to investigate.
Posted Jul 24, 2026 - 16:28 UTC
Pages is experiencing degraded performance. We are continuing to investigate.
Posted Jul 24, 2026 - 16:27 UTC
Copilot is experiencing degraded performance. We are continuing to investigate.
Posted Jul 24, 2026 - 16:26 UTC
We are investigating timeouts to some GitHub services
Posted Jul 24, 2026 - 16:22 UTC
Pull Requests is experiencing degraded performance. We are continuing to investigate.
Posted Jul 24, 2026 - 16:20 UTC
Actions is experiencing degraded performance. We are continuing to investigate.
Posted Jul 24, 2026 - 16:19 UTC
We are investigating reports of degraded performance for API Requests and Issues
Posted Jul 24, 2026 - 16:17 UTC
This incident affected: API Requests, Issues, Pull Requests, Actions, Pages, and Copilot.
| # | Наименование новости | Тональность | Информативность | Дата публикации |
|---|---|---|---|---|
| 1 | Latency issues across a number of services | 0 | 12.62 | 23-07-2026 |
| 2 | Incident with Actions | 0 | 8.82 | 06-08-2026 |
| 3 | Incident with Actions | 0 | 14.96 | 25-07-2026 |
| 4 | Actions run failures and delays | 0 | 10.08 | 25-07-2026 |
| 5 | Incident with Actions | 0 | 12.56 | 29-07-2026 |
| 6 | Disruption with actions hosted runners | 0 | 13.62 | 22-07-2026 |
| 7 | Incident With Blocked GitHub.com Traffic | 0 | 11.66 | 24-07-2026 |
| 8 | Some Copilot Cloud Agent jobs not starting | 0 | 10.28 | 05-08-2026 |
| 9 | Incident with Pull Requests | 0 | 10.96 | 24-07-2026 |
| 10 | Incident with Copilot AI Model Providers | 0 | 7.81 | 29-07-2026 |