Home / News

ChatGPT, Claude and Grok Suffer Overlapping Service Outages

OpenAI, Anthropic and xAI experienced overlapping service disruptions on September 3, temporarily degrading several of the most widely used hosted AI platforms. The providers have not identified a common cause, and available status information does not establish that the incidents resulted from a shared cloud or network failure.

Grok was the earliest confirmed failure. xAI’s status pages reported model outages beginning at 13:30 UTC across Grok’s web service, API, mobile applications and integrations, saying it was investigating and working to restore service.

Anthropic reported elevated errors beginning at about 13:23 UTC for several Claude models. The company later said it had identified the cause and deployed a fix. An Anthropic spokesperson described the disruption as an “infrastructure issue” affecting Claude.ai, Claude Code, Claude Cowork and the Claude API, without identifying an external infrastructure provider. Most services subsequently recovered, although some models remained affected for longer.

OpenAI’s problems followed later in the overlapping window. Its status service reported elevated errors affecting ChatGPT and Codex shortly before 15:00 UTC and subsequently said a mitigation had been applied while it monitored recovery. The incident affected multiple ChatGPT and Codex components.

There were also user reports of problems with Google’s Gemini during roughly the same period, prompting some reports to describe a four-provider outage. Google had not publicly acknowledged a Gemini incident, however, and its Cloud Service Health history showed no corresponding September 3 Gemini event. The Gemini disruption therefore remains less firmly established than the incidents acknowledged by OpenAI, Anthropic and xAI.

No shared upstream failure has been demonstrated. Microsoft Azure, Amazon Web Services and Cloudflare had not reported major concurrent incidents when Ars Technica checked their public status systems, and Microsoft’s public Azure dashboard showed no broad active event. That leaves open the possibility that the timing was coincidental or that a dependency not reflected on public status pages was involved.

The overlap is nevertheless notable as enterprises increasingly depend on remotely operated AI models through web services, APIs and coding tools. Separate failures occurring within the same window can produce much the same operational problem as a shared infrastructure outage for organizations whose fallback strategy consists largely of switching between a small number of external AI providers.

NORDVPN DISCOUNT - CircleID x NordVPN
Get NordVPN  [74% +3 extra months, from $2.99/month]
By CircleID Reporter

CircleID’s internal staff reporting on news tips and developing stories. Do you have information the professional Internet community should be aware of? Contact us.

Visit Page

Filed Under

Comments

Comment Title:

  Notify me of follow-up comments

We encourage you to post comments and engage in discussions that advance this post through relevant opinion, anecdotes, links and data. If you see a comment that you believe is irrelevant or inappropriate, you can report it using the link at the end of each comment. Views expressed in the comments do not represent those of CircleID. For more information on our comment policy, see Codes of Conduct.

CircleID Newsletter The Weekly Wrap

More and more professionals are choosing to publish critical posts on CircleID from all corners of the Internet industry. If you find it hard to keep up daily, consider subscribing to our weekly digest. We will provide you a convenient summary report once a week sent directly to your inbox. It's a quick and easy read.

Related

Topics

Cybersecurity

Sponsored byVerisign

DNS Security

Sponsored byWhoisXML API

Brand Protection

Sponsored byCSC

Domain Names

Sponsored byVerisign

IPv4 Markets

Sponsored byIPv4.Global

New TLDs

Sponsored byRadix

DNS

Sponsored byDNIB.com