‹ 返回事件历史

跨多个服务的延迟问题

原文Latency issues across a number of services
已恢复较大故障GitHub
2026年7月23日星期四 15:53 ~ 2026年7月23日星期四 17:39(1 小时 45 分钟)
受影响组件
Webhooks
Issues
Pull Requests
Actions
更新记录
已恢复2026年7月23日星期四 09:39

2026年7月23日,协调世界时07:08至09:39期间,多项服务出现延迟:8%的Actions工作流运行平均启动延迟10分钟,5%的Webhook投递超出服务等级目标,代码扫描、仓库、通知、议题及拉取请求在事件期间均经历了延迟增加。<br /><br />事件的根本原因是我们后台作业处理系统的一个节点在进入计划主机维护后未能恢复。通过识别问题分片并恢复其正确状态,事件得到缓解,随后队列积压得以清空,服务恢复正常。<br /><br />为加快缓解速度,我们已为维护操作后处于此不健康状态的节点添加了监控。为防止未来再次发生,我们正在调整生命周期自动化,以在计划重启后验证主机重新加入。

原文On July 23, 2026, between 07:08 and 09:39 UTC, several services experienced delays: 8% of actions workflow runs experienced an average run start delay of 10 minutes, 5% of webhook deliveries exceeded SLO, and code scanning, repos, notifications, issues and pull requests experienced increased latency over the life of the incident. <br /><br />The root cause of the incident was a node of our background job processing system which did not recover after entering scheduled host maintenance. The incident was mitigated by identifying the problematic shard and restoring its correct state, after which queue backlogs drained and services recovered. <br /><br />To speed mitigation, we have added monitors for nodes in this unhealthy state after maintenance operations. To prevent future recurrence, we are adapting our lifecycle automation to verify host rejoin after a scheduled reboot.

监控中2026年7月23日星期四 09:39

故障已缓解。我们正在监控以确保稳定性。

原文The degradation has been mitigated. We are monitoring to ensure stability.

排查中2026年7月23日星期四 09:35

Webhooks 运行正常。

原文Webhooks is operating normally.

排查中2026年7月23日星期四 09:27

影响Pull Requests的服务降级问题已得到缓解。我们正在持续监控以确保稳定性。

原文The degradation affecting Pull Requests has been mitigated. We are monitoring to ensure stability.

排查中2026年7月23日星期四 09:22

我们已定位影响多项服务的延迟源头并应用了修复。相关问题和操作正在恢复中,其余受影响服务随着处理积压的清除也在逐步改善。我们正积极监控所有服务的恢复情况。

原文We identified the source of latency affecting multiple services and applied a fix. Issues and Actions are recovering, and remaining affected services are seeing improvement as processing backlogs clear. We are actively monitoring recovery across all services.

排查中2026年7月23日星期四 09:19

影响Actions的服务降级问题已得到缓解。我们正在监控以确保稳定性。

原文The degradation affecting Actions has been mitigated. We are monitoring to ensure stability.

排查中2026年7月23日星期四 09:18

影响Issues的服务降级问题已得到缓解。我们正在监控以确保稳定性。

原文The degradation affecting Issues has been mitigated. We are monitoring to ensure stability.

排查中2026年7月23日星期四 08:34

我们目前正在调查多个服务的延迟问题。这可能导致Actions任务启动时间变长、Issues搜索返回过期结果,以及其他列出的服务受到类似影响。

原文We're currently investigating latency across multiple services. This can show as Actions jobs taking longer to start, Issues search serving stale results, and other listed services being similarly impacted.

排查中2026年7月23日星期四 08:25

Pull Requests 正经历性能下降。我们正在继续调查。

原文Pull Requests is experiencing degraded performance. We are continuing to investigate.

排查中2026年7月23日星期四 07:53

我们正在调查有关Actions、Issues和Webhooks可用性下降的报告。

原文We are investigating reports of degraded availability for Actions, Issues and Webhooks

事件内容来自 GitHub 官方状态页:查看官方原文。中文由 AI 翻译,仅供快速理解,以官方英文原文为准。

关于这些数据

事件从哪来?

全部来自各厂商官方状态页的公开接口,由本站每 5 分钟同步一次,保留最近 90 天。事件标题、时间、影响级别与更新记录均为官方原文,本站不做改写;点详情页底部的链接可回到厂商原始记录核对。

为什么有的服务查不到历史?

本页只收录对外提供官方状态页的服务。没有公开状态页的厂商(多数国产大模型属于此类)无法取得可信数据,本站不做自建拨测去猜,因此也不会出现在状态总览里。

持续时长怎么算?

按官方标注的开始时间到恢复时间计算;尚未恢复的事件按「至今」计算并标为进行中。跨天的事件在状态条上会覆盖它经过的每一天。