starslingdev - Delays starting Actions runs – Incident details

All systems operational

Delays starting Actions runs

Resolved
Major outage
Started 18 days agoLasted about 9 hours

Affected

Third Party: GitHub → Actions

Operational from 4:34 AM to 4:51 AM, Degraded performance from 4:51 AM to 10:07 AM, Partial outage from 10:07 AM to 12:54 PM, Major outage from 12:54 PM to 1:16 PM, Partial outage from 1:16 PM to 1:39 PM, Operational from 1:39 PM to 1:52 PM

Updates
  • Resolved
    UTC
    Resolved

    On July 9, 2026, between 03:29 UTC and 13:39 UTC, GitHub Actions experienced delayed and failed job starts on GitHub-hosted runners. The incident was caused by an unhealthy state in a backend data service responsible for provisioning hosted runners, preventing runner acquisition for a subset of workloads. During most of the incident, approximately 8% of workflow runs on hosted runners were delayed by more than 5 minutes, while roughly 2% failed to start.

    At 13:39 UTC, we restored the health of the backend data replication system, allowing provisioning to recover and the accumulated workflow backlog to drain. Service performance then returned to expected levels. We are improving provisioning-service resiliency, workload distribution, and capacity balancing to reduce the likelihood and impact of similar incidents.

  • Update
    UTC
    Update

    Actions, Pages builds, Copilot Cloud Agent, and Copilot Code review have all recovered and are mitigated.

    We are continuing to monitor to ensure full recovery, and investigating the health of the affected infrastructure.

  • Monitoring
    UTC
    Monitoring

    The degradation affecting Actions and Pages has been mitigated. We are monitoring to ensure stability.

  • Update
    UTC
    Update

    We are continuing to monitor slow recovery in Actions and Pages builds as the system works through the high volume of backlog.

    Customers may see a small rate of  API and job failures as the system is recovering.

    Copilot Cloud Agent and Copilot Code Review also failed to start for approximately 30 minutes during this incident, and we are monitoring recovery.

    Pages were accessible throughout the incident.

  • Update
    UTC
    Update

    Actions is experiencing degraded availability. We are continuing to investigate.

  • Update
    UTC
    Update

    We're seeing Actions and Pages recovery.

    For a period of approximate 20 minutes ~96% of GitHub Actions runs on GitHub-hosted runners were failing to start, but has now recovered and we are seeing jobs processing.

    GitHub pages builds were also failing during that period, but Pages are still accessible.

    We are continuing to monitor for full recovery.

  • Update
    UTC
    Update

    Pages is experiencing degraded performance. We are continuing to investigate.

  • Update
    UTC
    Update

    Approximately 30% of GitHub Actions runs on GitHub-hosted runners are experiencing run start delays exceeding 5 minutes. A smaller percentage of those are exhausting retries and failing to start.
    This has caused some customers to exceed their hosted compute concurrency and experience increased impact.

    We are continuing to working on infrastructure mitigations.

    Next update in one hour.

  • Update
    UTC
    Update

    Approximately 30% of GitHub Actions runs on GitHub-hosted runners are experiencing run start delays exceeding 5 minutes. A smaller percentage of those are exhausting retries and failing to start.

  • Update
    UTC
    Update

    Actions is experiencing degraded availability. We are continuing to investigate.

  • Update
    UTC
    Update

    We are continuing to work on a mitigation.

  • Investigating
    UTC
    Investigating

    Approximately 5% of GitHub Actions runs on GitHub-hosted runners are experiencing run start delays exceeding 5 minutes. A small portion of these runs may fail after extended delays. We have identified the cause and are working on a mitigation.