Nobody Is Saying Why OpenAI and Anthropic Had Outages Today

Nobody Is Saying Why OpenAI and Anthropic Had Outages Today

没人说明 OpenAI 和 Anthropic 今天为何发生宕机

Frontier models from Anthropic, OpenAI, and xAI all experienced rare outages on Thursday morning, creating downtime for their corresponding AI chatbots. SpaceX, xAI’s parent company, said on Thursday afternoon that the issues with Grok resulted from “an outage at our Memphis compute center this morning.”

周四上午,来自 Anthropic、OpenAI 和 xAI 的前沿模型均出现了罕见的宕机,导致其对应的 AI 聊天机器人无法使用。xAI 的母公司 SpaceX 在周四下午表示,Grok 出现的问题是由于“我们孟菲斯计算中心今天上午发生了一次宕机”。

The issues initially appeared to be linked because they coincided—perhaps the result of a shared third-party service provider—but neither OpenAI nor Anthropic cited an external source in comments to WIRED on Thursday. SpaceX, xAI’s parent company, did not respond to WIRED’s request for comment. But the company said as part of its public comments on Thursday: “We’d also like to apologize to our impacted compute partners.” Anthropic and xAI announced a “compute partnership” with SpaceX in May.

这些问题起初看起来似乎有关联,因为它们发生的时间重合——这可能是由于共享了第三方服务提供商所致——但 OpenAI 和 Anthropic 在周四给《连线》(WIRED)的评论中均未提及外部原因。xAI 的母公司 SpaceX 没有回应《连线》的置评请求。但该公司在周四的公开评论中表示:“我们也想向受影响的计算合作伙伴道歉。”Anthropic 和 xAI 曾在五月份宣布与 SpaceX 建立“计算合作伙伴关系”。

OpenAI spokesperson Kathleen Chaykowski tells WIRED: “A routing error starting around 7:43 am PT on Thursday, September 3, made ChatGPT and Codex unavailable for some users across platforms. As of about 8:17 am PT on Thursday, a solution was successfully implemented and is continuing to be monitored.”

OpenAI 发言人 Kathleen Chaykowski 告诉《连线》:“太平洋时间 9 月 3 日周四上午 7:43 左右开始的一个路由错误,导致部分用户无法跨平台使用 ChatGPT 和 Codex。截至太平洋时间周四上午 8:17 左右,解决方案已成功实施,目前正在持续监控中。”

Anthropic declined to comment on the episode. The company began alerting about a “partial outage” at 6:23 am PT on Thursday that involved “elevated errors on requests to Claude Mythos 5.1, Claude Fable 5.1, and Claude Opus 5.” Shortly after, the company said it had “identified the cause” and that “a fix has been deployed.” The company marked the issue as resolved by 9:16 am PT. Claude Sonnet 5 seemed to briefly have similar issues shortly after 9 am PT.

Anthropic 拒绝就此事发表评论。该公司在太平洋时间周四上午 6:23 开始发出“部分宕机”警报,涉及“对 Claude Mythos 5.1、Claude Fable 5.1 和 Claude Opus 5 的请求错误率升高”。不久之后,该公司表示已“查明原因”并“已部署修复程序”。该公司在太平洋时间上午 9:16 将该问题标记为已解决。Claude Sonnet 5 在太平洋时间上午 9 点后不久似乎也短暂出现了类似问题。

xAI reported Grok outages across all of its platforms and services beginning at 6:30 am PT when the company posted “investigating outage” on its service status page. “Grok is experiencing issues. We are working on restoring service as quickly as possible,” the page said. At 10:05 am PT the episode was marked complete. “We have resolved the situation, and traffic is healthy again,” the company wrote.

xAI 报告称,其所有平台和服务从太平洋时间上午 6:30 开始出现 Grok 宕机,当时该公司在其服务状态页面上发布了“正在调查宕机”的通知。页面显示:“Grok 遇到问题。我们正在努力尽快恢复服务。”太平洋时间上午 10:05,该事件被标记为结束。该公司写道:“我们已经解决了问题,流量已恢复正常。”

There were scattered reports of a possible Google Gemini outage on Thursday morning as well, but the company did not confirm this or record any incidents on its service status dashboard. Google did not respond to WIRED’s request for comment ahead of publication.

周四上午也有零星报道称 Google Gemini 可能出现了宕机,但谷歌公司并未证实这一点,其服务状态仪表板上也未记录任何事故。在本文发布前,谷歌未回应《连线》的置评请求。

Typically, multiple outages in the same sector at the same time would point to a cloud provider, content delivery network, or other third-party vendor having issues affecting multiple customers. But OpenAI and Anthropic did not point to a potential shared cause, and major players in the internet infrastructure space—including Cloudflare, Amazon Web Services, and Microsoft Azure—did not report outages on Thursday.

通常情况下,同一行业在同一时间发生多次宕机,往往指向云服务提供商、内容分发网络或其他第三方供应商出现问题,从而影响了多个客户。但 OpenAI 和 Anthropic 并未指出潜在的共同原因,且互联网基础设施领域的主要参与者——包括 Cloudflare、亚马逊云科技(AWS)和微软 Azure——在周四均未报告宕机。