보관된 공식 INCIDENT

Actions delays in starting runs

GitHub · 마지막 저장 제공업체 상태: resolved

현재 서비스 보고 →

공식 출처 문구는 원어로 표시됩니다.

과거 증거입니다.

마지막 저장 사고 상태는 현재 서비스 상태가 아닙니다. 출처 목록은 불완전할 수 있고 소멸은 복구 확인이 아닙니다. 최신 증거는 현재 보고 또는 공식 출처를 확인하세요.

저장 사고 상세

제공업체 상태
resolved
제공업체 영향
minor
제공업체 기록 생성
24 Aug 2026, 13:56:54 UTC
제공업체 기록 갱신
25 Aug 2026, 01:36:32 UTC
제공업체가 명시한 시작
24 Aug 2026, 13:56:54 UTC
제공업체가 명시한 종료
24 Aug 2026, 14:34:42 UTC

제공업체가 보고한 영향받은 구성 요소

  • Actions br0l2tvcx85d

연결은 사고 보고 범위이며 현재 구성 요소 가용성이나 검증 의존 관계를 입증하지 않습니다.

기록 생성 시각이 반드시 장애 시작 시각인 것은 아닙니다. 보고되지 않은 시각은 알 수 없음으로 유지하며 수집 시각으로 중단 시간을 계산하지 않습니다.

저장 수정본의 제공업체 업데이트

최신순입니다. 최신 저장 내용 수정본 20개에서 최대 100개의 서로 다른 업데이트를 표시합니다. 같은 제공업체 시각의 문구 변경도 별도 보존됩니다.

  1. resolved

    On August 24, 2026, between 13:33 UTC and 14:04 UTC, 3.8% of Actions runs experienced start delays over 5 minutes with 1.25% of Actions runs failing outright. <br /> <br />The incident was caused by a disk failure on a node hosting one of many service instances responsible for processing runner assignment events. Typically, pods on unhealthy nodes are removed and replaced automatically without impact. In this case, although the node was severely degraded and unable to perform disk operations, it continued sending healthy signals, preventing the system from immediately moving its work elsewhere. During this period, events assigned to the affected component accumulated until an automatic rebalance redirected processing to healthy components at 13:54 UTC. The queue backlog was cleared at 14:00 UTC, and processing returned to normal by 14:04 UTC. <br /><br />To prevent a recurrence, we are improving detection and automated remediation for unhealthy nodes that aren’t fully offline. We are also strengthening application-level resiliency, so stalled consumers are automatically removed quickly and their work reassigned without waiting for the affected node to recover.

    저장 수정본에서 확인 06 Oct 2026, 13:58:39 UTC
  2. monitoring

    The degradation affecting Actions has been mitigated. We are monitoring to ensure stability.

    저장 수정본에서 확인 06 Oct 2026, 13:58:39 UTC
  3. investigating

    Failures while queuing and running Actions jobs for a subset of customers are now resolving. We are monitoring for full recovery.

    저장 수정본에서 확인 06 Oct 2026, 13:58:39 UTC
  4. investigating

    We are investigating reports of degraded performance for Actions

    저장 수정본에서 확인 06 Oct 2026, 13:58:39 UTC