今天没有人说出OpenAI和Anthropic为何发生宕机。

内容来源:https://www.wired.com/story/nobody-is-saying-why-openai-and-anthropic-had-outages-today/
内容总结:
周四上午,Anthropic、OpenAI 和 xAI 三家前沿人工智能公司均遭遇罕见的大规模服务中断,导致各自的AI聊天机器人出现临时无法访问的情况。xAI母公司SpaceX于周四下午表示,Grok出现的问题源于“今晨位于孟菲斯计算中心的一次中断”。
由于三家公司服务中断时间高度重合,外界最初猜测可能存在关联,或源于某个共同的第三方服务提供商。然而,OpenAI和Anthropic在周四均未向《连线》杂志提及外部原因。SpaceX也未回应采访请求,但在公开声明中表示:“我们还要向受影响的计算合作伙伴致歉。”今年5月,Anthropic和xAI已与SpaceX宣布建立“计算合作伙伴关系”。
OpenAI发言人凯瑟琳·柴科夫斯基对《连线》表示:“太平洋时间周四上午约7点43分开始的一次路由错误,导致部分用户在不同平台上无法使用ChatGPT和Codex。截至太平洋时间上午约8点17分,修复方案已成功实施并将持续监控。”Anthropic则拒绝就此事置评。该公司于太平洋时间凌晨6点23分开始发出“部分中断”警报,涉及“Claude Mythos 5.1、Claude Fable 5.1和Claude Opus 5的请求错误率升高”。不久后,公司称已“查明原因”并“部署修复”。太平洋时间上午9点16分,问题宣告解决。约9点后,Claude Sonnet 5也短暂出现类似问题。
xAI自太平洋时间凌晨6点30分起在其服务状态页面上报告Grok全部平台和服务出现中断,页面显示:“Grok正遭遇问题,我们正尽快恢复服务。”太平洋时间上午10点5分,事件标记为已完成。公司写道:“情况已解决,流量已恢复正常。”
周四上午还有零散报告称Google Gemini可能也出现了中断,但谷歌未予确认,也未在其服务状态仪表板上记录任何事件。截至发稿前,谷歌未回应《连线》的置评请求。
通常情况下,同行业多家公司同时发生中断,往往指向云服务商、内容分发网络或其他第三方供应商的问题,因其共同受影响。但OpenAI和Anthropic均未提及潜在共同原因,而包括Cloudflare、亚马逊云服务(AWS)和微软Azure在内的主要互联网基础设施服务商,周四均未报告出现中断。
中文翻译:
来自Anthropic、OpenAI和xAI的前沿模型周四上午均遭遇了罕见的宕机,导致其相应的人工智能聊天机器人出现服务中断。xAI的母公司SpaceX在周四下午表示,Grok的问题源于“今晨我们孟菲斯计算中心发生的一次宕机”。
这些问题最初看似相互关联,因为发生时间恰好重合——或许是由于共享某家第三方服务提供商所致——但OpenAI和Anthropic周四在给《连线》杂志的评论中均未提及外部原因。xAI的母公司SpaceX未回应《连线》杂志的置评请求。但该公司在周四的公开声明中表示:“我们也想向受影响的算力合作伙伴致歉。”Anthropic和xAI于今年5月宣布与SpaceX达成“算力合作”。
OpenAI发言人凯瑟琳·查伊科夫斯基告诉《连线》杂志:“太平洋时间9月3日周四上午7点43分左右开始的一次路由错误,导致ChatGPT和Codex在多个平台上对部分用户不可用。截至周四上午8点17分左右,解决方案已成功实施,并正在持续监控中。”
Anthropic拒绝对此事发表评论。该公司于太平洋时间周四上午6点23分开始发出“部分宕机”警报,涉及“Claude Mythos 5.1、Claude Fable 5.1和Claude Opus 5的请求出现较高错误率”。不久后,该公司表示已“确定原因”且“修复方案已部署”。该公司于太平洋时间上午9点16分将问题标记为已解决。Claude Sonnet 5在太平洋时间上午9点刚过时似乎也短暂出现了类似问题。
xAI报告称,Grok从太平洋时间上午6点30分起在其所有平台和服务上出现宕机,当时该公司在其服务状态页面上发布了“正在调查宕机”的公告。“Grok目前遇到问题。我们正在尽快恢复服务,”该页面写道。太平洋时间上午10点05分,该事件被标记为已结束。“我们已解决问题,流量已恢复正常,”该公司写道。
周四上午也有零散的关于Google Gemini可能宕机的报告,但该公司未予确认,也未在其服务状态面板上记录任何事件。谷歌在发布前未回应《连线》杂志的置评请求。
通常情况下,同一行业在同一时间发生多起宕机事件,会指向某个云服务提供商、内容分发网络或其他第三方供应商出现问题而影响了多家客户。但OpenAI和Anthropic均未指向可能存在的共同原因,而互联网基础设施领域的主要参与者——包括Cloudflare、亚马逊网络服务和微软Azure——周四均未报告宕机。
评论
返回顶部
英文来源:
Frontier models from Anthropic, OpenAI, and xAI all experienced rare outages on Thursday morning, creating downtime for their corresponding AI chatbots. SpaceX, xAI’s parent company, said on Thursday afternoon that the issues with Grok resulted from “an outage at our Memphis compute center this morning.”
The issues initially appeared to be linked because they coincided—perhaps the result of a shared third-party service provider—but neither OpenAI nor Anthropic cited an external source in comments to WIRED on Thursday. SpaceX, xAI’s parent company, did not respond to WIRED’s request for comment. But the company said as part of its public comments on Thursday: “We’d also like to apologize to our impacted compute partners.” Anthropic and xAI announced a “compute partnership” with SpaceX in May.
OpenAI spokesperson Kathleen Chaykowski tells WIRED: “A routing error starting around 7:43 am PT on Thursday, September 3, made ChatGPT and Codex unavailable for some users across platforms. As of about 8:17 am PT on Thursday, a solution was successfully implemented and is continuing to be monitored.”
Anthropic declined to comment on the episode. The company began alerting about a “partial outage” at 6:23 am PT on Thursday that involved “elevated errors on requests to Claude Mythos 5.1, Claude Fable 5.1, and Claude Opus 5.” Shortly after, the company said it had “identified the cause” and that “a fix has been deployed.” The company marked the issue as resolved by 9:16 am PT. Claude Sonnet 5 seemed to briefly have similar issues shortly after 9 am PT.
xAI reported Grok outages across all of its platforms and services beginning at 6:30 am PT when the company posted “investigating outage” on its service status page. “Grok is experiencing issues. We are working on restoring service as quickly as possible,” the page said. At 10:05 am PT the episode was marked complete. “We have resolved the situation, and traffic is healthy again,” the company wrote.
There were scattered reports of a possible Google Gemini outage on Thursday morning as well, but the company did not confirm this or record any incidents on its service status dashboard. Google did not respond to WIRED’s request for comment ahead of publication.
Typically, multiple outages in the same sector at the same time would point to a cloud provider, content delivery network, or other third-party vendor having issues affecting multiple customers. But OpenAI and Anthropic did not point to a potential shared cause, and major players in the internet infrastructure space—including Cloudflare, Amazon Web Services, and Microsoft Azure—did not report outages on Thursday.
Comments
Back to top
文章标题:今天没有人说出OpenAI和Anthropic为何发生宕机。
文章链接:https://news.qimuai.cn/?post=4971
本站文章均为原创,未经授权请勿用于任何商业用途