‹ 返回事件历史

运行启动操作延迟

原文Actions delays in starting runs
已恢复轻微故障GitHub
2026年8月24日星期一 21:56 ~ 2026年8月24日星期一 22:34(37 分钟)
受影响组件
Actions
更新记录
已恢复2026年8月24日星期一 14:34

2026年8月24日,在13:33 UTC至14:04 UTC期间,3.8%的Actions运行遭遇超过5分钟的启动延迟,其中1.25%的Actions运行直接失败。<br /> <br />此次事件由托管多个负责处理runner分配事件的服务实例之一的节点上的磁盘故障引起。通常,不健康节点上的pod会被自动移除并替换,不会产生影响。但在本例中,尽管节点严重降级且无法执行磁盘操作,它仍持续发送健康信号,导致系统未能立即将其工作转移至他处。在此期间,分配给受影响组件的事件不断累积,直到13:54 UTC自动重新平衡将处理工作重定向至健康组件。队列积压于14:00 UTC清除,处理工作于14:04 UTC恢复正常。<br /><br />为防止再次发生,我们正在改进对未完全离线的非健康节点的检测和自动修复机制。同时,我们也在加强应用层面的韧性,以便停滞的消费者能被快速自动移除,其工作无需等待受影响节点恢复即可重新分配。

原文On August 24, 2026, between 13:33 UTC and 14:04 UTC, 3.8% of Actions runs experienced start delays over 5 minutes with 1.25% of Actions runs failing outright. <br /> <br />The incident was caused by a disk failure on a node hosting one of many service instances responsible for processing runner assignment events. Typically, pods on unhealthy nodes are removed and replaced automatically without impact. In this case, although the node was severely degraded and unable to perform disk operations, it continued sending healthy signals, preventing the system from immediately moving its work elsewhere. During this period, events assigned to the affected component accumulated until an automatic rebalance redirected processing to healthy components at 13:54 UTC. The queue backlog was cleared at 14:00 UTC, and processing returned to normal by 14:04 UTC. <br /><br />To prevent a recurrence, we are improving detection and automated remediation for unhealthy nodes that aren’t fully offline. We are also strengthening application-level resiliency, so stalled consumers are automatically removed quickly and their work reassigned without waiting for the affected node to recover.

监控中2026年8月24日星期一 14:26

影响Actions的服务降级问题已得到缓解。我们正在监控以确保稳定性。

原文The degradation affecting Actions has been mitigated. We are monitoring to ensure stability.

排查中2026年8月24日星期一 14:22

部分客户在排队和运行Actions作业时遇到的故障正在逐步解决。我们正在监控以确认完全恢复。

原文Failures while queuing and running Actions jobs for a subset of customers are now resolving. We are monitoring for full recovery.

排查中2026年8月24日星期一 13:56

我们正在调查有关Actions性能下降的报告。

原文We are investigating reports of degraded performance for Actions

事件内容来自 GitHub 官方状态页:查看官方原文。中文由 AI 翻译,仅供快速理解,以官方英文原文为准。

关于这些数据

事件从哪来?

全部来自各厂商官方状态页的公开接口,由本站每 5 分钟同步一次,保留最近 90 天。事件标题、时间、影响级别与更新记录均为官方原文,本站不做改写;点详情页底部的链接可回到厂商原始记录核对。

为什么有的服务查不到历史?

本页只收录对外提供官方状态页的服务。没有公开状态页的厂商(多数国产大模型属于此类)无法取得可信数据,本站不做自建拨测去猜,因此也不会出现在状态总览里。

持续时长怎么算?

按官方标注的开始时间到恢复时间计算;尚未恢复的事件按「至今」计算并标为进行中。跨天的事件在状态条上会覆盖它经过的每一天。