Leading AI services in the United States experienced a collective failure, with ChatGPT, Claude, and Grok announcing service abnormalities, while Google's Gemini and Microsoft Copilot also received numerous interruption reports. This large-scale outage lasted approximately 3 hours and 40 minutes, with the failure first erupting around 9:23 AM Eastern Time on September 3rd.
Claude was the first to report errors, with OpenAI receiving over 37,000 fault reports at peak
According to data from the overseas fault monitoring platform Down Detector, the Claude series models from Anthropic experienced a sharp rise in error rates, including core models such as Mythos 5.1, Opus, and Fable 5.1. Just two minutes later, xAI's Grok service went offline entirely; about an hour and a half later, OpenAI's ChatGPT and programming tool Codex also experienced large-scale access errors.
During the peak of the incident, more than 37,000 fault reports were submitted globally for OpenAI services, with 80% concentrated on ChatGPT. At the same time, Google's Gemini and Microsoft Copilot received numerous user interruption feedback, while the AI programming tool Cursor directly announced that some of its services had been forced to interrupt due to upstream large model failures. By 11:16 AM Eastern Time on the same day, Claude's main services recovered first, followed by gradual restoration of ChatGPT and Grok.
Domestic large models remained stable, and Zhipu launched free nighttime usage
Notably, during the period when overseas large models collectively failed, domestic large models remained stable. That evening, Zhipu posted on social media X: "We are still up."—which translates to "We are still online."
At the same time, Zhipu announced on the evening of September 3rd that from now until September 20th, GLM-5.3-Flash in Z Code would be available for free between 11 PM and 9 AM daily, with the usage quota doubled for other Agents supported by the package, automatically taking effect during the specified time period.
According to the introduction, GLM-5.3-Flash is the first native multimodal model in the GLM-5 series, scoring 57 points in the comprehensive intelligence index of the globally authoritative Artificial Analysis, entering the global cutting-edge model capability range. This model is suitable for code generation, code understanding, problem fixing, and other code tasks, and shows significant improvements in Office and document-related tasks, as well as specialized tasks in finance and law.




