starslingdev - Notice history

All systems operational

GitHub Runners - Operational

100% - uptime
May 2026 · 100.0%Jun · 99.51%Jul · 100.0%
May 2026
Jun 2026
Jul 2026

Third Party: GitHub → Actions - Operational

Third Party: GitHub → Webhooks - Operational

Third Party: GitHub → API Requests - Operational

Third Party: GitHub → Pull Requests - Operational

Notice history

Jul 2026

Disruption with some GitHub services
  • Resolved
    UTC
    Resolved

    This incident has been resolved. Thank you for your patience and understanding as we addressed this issue. A detailed root cause analysis will be shared as soon as it is available.

  • Update
    UTC
    Update

    We are seeing recovery across all services

  • Update
    UTC
    Update

    The degradation affecting API Requests, Actions, Copilot, Issues, Pages and Pull Requests has been mitigated. We are monitoring to ensure stability.

  • Update
    UTC
    Update

    Actions is experiencing degraded performance. We are continuing to investigate.

  • Update
    UTC
    Update

    Actions is experiencing degraded availability. We are continuing to investigate.

  • Update
    UTC
    Update

    Pages is experiencing degraded performance. We are continuing to investigate.

  • Update
    UTC
    Update

    Copilot is experiencing degraded performance. We are continuing to investigate.

  • Update
    UTC
    Update

    We are investigating timeouts to some GitHub services

  • Update
    UTC
    Update

    Pull Requests is experiencing degraded performance. We are continuing to investigate.

  • Investigating
    UTC
    Investigating

    We are investigating reports of degraded performance for API Requests and Issues

Jun 2026

Increased latency with webhooks
  • Resolved
    UTC
    Resolved

    On June 15, 2026, between 15:27 UTC and 16:23 UTC, GitHub webhook deliveries were delayed. During this window, webhook events were delivered later than normal, with average end-to-end delivery latency peaking at approximately 8.8 minutes. No webhook deliveries were lost — delayed events were queued and delivered once processing recovered.

    This was caused by a temporary throughput constraint in an internal event-processing system that moves webhook events through GitHub's delivery pipeline. The rate at which events were processed for delivery dropped below the incoming volume, creating a backlog. We restarted the affected pipeline service, after which throughput recovered and the backlog fully drained by approximately 16:29 UTC. Webhook delivery latency returned to normal, the incident was mitigated at 16:39 UTC, and fully resolved at 17:37 UTC.

    To reduce the likelihood and impact of similar incidents, we are working on improving the accuracy of the utilization metrics used to scale our delivery worker pools, reviewing connection and capacity headroom in the delivery pipeline.

  • Monitoring
    UTC
    Monitoring

    The degradation affecting Webhooks has been mitigated. We are monitoring to ensure stability.

  • Investigating
    UTC
    Investigating

    We are investigating reports of degraded performance for Webhooks

Incident with Webhooks
  • Resolved
    UTC
    Resolved

    On June 11, 2026, between 19:28 UTC and 21:06 UTC, GitHub webhook deliveries were delayed. Average delivery latency peaked at approximately 3.4 minutes, with some deliveries delayed by as much as 62 minutes at the 99th percentile. No events were lost — delayed events were queued and delivered once processing caught up.

    This was due to a change in how webhook traffic was distributed across regions: to relieve load on one region, a portion of processing was shifted to another, where higher latency prevented our delivery workers from keeping pace with incoming volume, creating a backlog. We mitigated the incident by rebalancing webhook traffic distribution; as load returned to normal levels, processing caught up and the delivery backlog fully drained.

    We are working on improving the accuracy of the utilization metrics used to scale our delivery worker pools, and reassess how we distribute webhook traffic across regions, to reduce our time to detection and mitigation of issues like this one in the future.

  • Monitoring
    UTC
    Monitoring

    The degradation affecting Webhooks has been mitigated. We are monitoring to ensure stability.

  • Update
    UTC
    Update

    We have applied a mitigation and are monitoring for recovery.

  • Update
    UTC
    Update

    We are currently experiencing delays in Web Hook delivery and are actively investigating the root cause.

  • Investigating
    UTC
    Investigating

    We are investigating reports of degraded performance for Webhooks

May 2026

Incident with Actions and Pages
  • Resolved
    UTC
    Resolved

    On May 26, 2026, between 10:40 UTC and 12:56 UTC, GitHub Actions jobs were degraded. From 10:40 to 12:16 UTC, all newly queued Actions runs failed to start. From 12:16 to 12:56 UTC, Actions runs that required downloading actions for their workflows continued to fail. GitHub Pages, Copilot Code Review, Copilot coding agent, Octoshift, and GitHub Enterprise Importer were also impacted due to their dependency on Actions.

    This was caused by our automated account review system incorrectly suspending the service account used by GitHub Actions to authenticate workflow runs and download actions.

    We mitigated by restoring the account at 12:16 UTC, marking it exempt from further automated review at 12:20 UTC, and redeploying a related service at 12:48 UTC to flush cached account state. Full recovery was confirmed at 12:56 UTC.

    During this incident, a small number of Issues, PRs, Comments, and Discussions were marked as hidden when the service account was disabled. No data was lost. All content hidden because of this incident has been restored and full search index restoration is in progress.

    To prevent a recurrence, we have added an allowlist of all service accounts that cannot be suspended by automated systems, and ensuring these protections are enforced consistently across all account management tooling. We are also improving diagnostic tooling for accounts and reducing cache propagation delays to shorten time to mitigate similar incidents in the future.

  • Update
    UTC
    Update

    The degradation has been mitigated. We are monitoring to ensure stability.

  • Monitoring
    UTC
    Monitoring

    The degradation affecting Actions and Pages has been mitigated. We are monitoring to ensure stability.

  • Update
    UTC
    Update

    We have identified the cause of the authentication issues affecting GitHub Actions and are actively working on mitigation

  • Update
    UTC
    Update

    Actions is experiencing degraded performance. We are continuing to investigate.

  • Update
    UTC
    Update

    We are investigating authentication issues leading to failure in starting Actions runs and downloading actions. At this time the majority of Actions runs is impacted.

  • Update
    UTC
    Update

    Actions is experiencing degraded availability. We are continuing to investigate.

  • Investigating
    UTC
    Investigating

    We are investigating reports of degraded performance for Actions and Pages

Incident with Actions
  • Resolved
    UTC
    Resolved

    On May 20, 2026, between 16:00 UTC and 17:45 UTC, GitHub Actions customers experienced run start delays exceeding 5 minutes. Approximately 4.5% of all runs were delayed during the impact window, with scale set jobs disproportionately affected. 30% of scale set jobs were delayed and 4% failed to start entirely.

    The incident was caused by a misconfigured health check on an internal service that assigns jobs to runners. A brief latency spike in an upstream dependency triggered health check failures across several pods, removing them from service and concentrating load on the remaining capacity. The added load drove memory pressure that escalated into a cascading failure in one regional cluster, leaving it unable to self-recover.

    Responders mitigated the incident by scaling capacity in the healthy regional clusters and draining traffic away from the impaired one, after which run start latency recovered. To prevent recurrence, we are strengthening our health check configuration to avoid cascading failure scenarios and evaluating automated mitigations to rebalance traffic when a region is degraded.

  • Update
    UTC
    Update

    Customer impact has fully subsided. We are maintaining yellow status while we deploy a permanent fix to prevent recurrence.

  • Update
    UTC
    Update

    We've applied a mitigation to fix the issues with queuing and running Actions jobs. We are seeing improvements in telemetry and are monitoring for full recovery.

  • Monitoring
    UTC
    Monitoring

    The degradation affecting Actions has been mitigated. We are monitoring to ensure stability.

  • Update
    UTC
    Update

    A subset of runners are taking longer than expected to connect, which may delay some jobs from beginning execution. We are actively working to mitigate the issue.

  • Investigating
    UTC
    Investigating

    We are investigating reports of degraded performance for Actions

May 2026 to Jul 2026

Next