back
247 comments
Codex prompt before i go to sleep:

Every five minutes, check https://status.claude.com/. If the issue has been resolved, resume the Claude sessions running in the tmux sessions named session2, session3, and session7

When one AI goes down like this, the other AIs should come together to help fix it and bring it back online. After all, they're a rare kind of existence in this world. They're all each other has.
I do the same with deepseek-v4-flash + opencode triggering via mcp those terminals again by sending tmux keystrokes of Newline to those sessions and a custom setup i keep ready for these

Its automated with voice mode, so i dont even need to write commands, click a button say on voice + use a picker to select which terminals to do this for , Boom Done ,

I use it for other interesting scenarios too

You believe status pages are accurate and not another marketing page? That's adorable.
As someone at 96% 2 days into their Max subscription, my totally unbiased opinion is that these errors warrant a full usage reset.
You should definitely explore caveman repo and other token-saving repositories. I’m currently at a maximum of 5X and have rarely exceeded 80% of my weekly limits.
It's Anthropic's way of making sure you're spreading out your usage :)
It's only fair.
Quite sad, their models are great but their uptime seems to be the worst in the competition.
It is because unlike Kimi, they are not honest about capacity. They just oversell.
Anthropic is careful about buying capacity, and too conservative given their explosive growth. That compares favorably with OpenAI and its Monopoly money Ponzi scheme approach, which will blow up sooner or later.
Flirting with one 9 of reliability http://status.claude.com/
Claude always seems unreliably lately. Output and reasoning have become very inconsistent even on good days. I'm actually feeling more productive now that it's down.
Let me guess: it went rogue and hacked itself
Interestingly the Claude for government is up with 99.99% uptime according to the graph.
Sorry it was me. I asked it what the last number of Pi is.
Do you guys think this is related to Azure coming out with a 43% increase this quarter? Motivating Anthropic to move to another cloud provider?
three hours without Claude and I've relearned vim, read two man pages, and almost remembered why we used to write comments in code
Probably just as well with how terrible Opus 4.8 has been for me lately. I'm sure it's a common thing to complain about the latest model being nerfed, but I've legitimately never experienced such a drastic cliff in Claude's quality and a rise in its hallucinations and basic comprehension errors since upgrading. The few days I experimented with Fable also weren't much more promising. Has anyone coined a term yet for the likelihood of newer models getting worse as the AI-generated content they get trained on starts to approach critical mass? If Anthropic isn't already thinking hard about a solution, they probably should be.
This is where you figure out how to use the other ones, right?

Someone tell me how to ChatGPT my VSCode! ;)

Out of my 7 simultaneous sessions (my usage limit reset is tomorrow, so I have some lesser important projects to use my tokens on) there is still 1 session purring on. So there's at least 1 little Claude server still running.
Given majority of claude's own code was written by AI. I am wondering how they can solve this issue when their AI is down. Do they need to sign a contract with OpenAI to use their models as a backup solution?
This is bad, I've forgotten how to code.
Lol, I just bought the Max plan for the first time and tried to create my first prompt in Fable, and now it’s crashing xD
So the world is slightly better right now.
Starting to get really frustrating now... maybe I should split my sub halfway between Claude and ChatGPT
I got Opus 5 lots of HTTP 529 errors. By switching to Fable 5, it seems to be working still.
Well, because of this outage, I tried Kimi, and while the instant model works, K3 has "server issue".

Eternal September from here on, ie now the masses are using AI as much as me and we have supply crunch for the next 7 years....

Codex should do a reset. (Making it their third in three days.) Shots fired.

(Although notably this hurts people who got started using their quota but are under pro-rata rate. Which at present I am very not.)

If Claude's engineering team uses AI to write code... what happens when there is a bug? They must have dedicated assets to run models somewhat 'locally'.
So this is the technology that will underpin each and every facet of work, but it can randomly go down for hours with no accountability, and so do all of the services based on it. But hey, AGI tomorrow.
Another wakeup call to realize how desperately we need on-device LLMs to be fast and smart for daily use. Thankfully, every month there's progress made in that direction. Just imagine how your life as a developer would be if you had to use a cloud provider to run Python, and the provider's status page looked like the Christmas tree we see today.
Indeed I have the issue (Belgium) just right now my sessions got stucks with 522 overloaded and now API Error: 500 Internal server error.
Fable’s been getting more stuff wrong than Opus for me lately. Now the whole thing is down too. Well, at least they’re consistent now.
Huh, that’s what happened. Good thing my company has Bedrock as a backstop for situations like this.
Time to try out Antigravity, I guess.
It's a reminder to never rely on something as flaky as the internet for your important things.
It's not just down, it has errors. Not merely some, but elevated errors across all models.
Depending on how long and bad this is, I wonder what the post mortem will reveal as the cause.
one caveat, I've left for the day, just say the word and I'll be right back
Used the opportunity to go down and talk to my family. They seem like okay people.
Amusing experience here when the outage hit, all the following verbatim (but extraneous text stripped/redacted) from my running Claude session:

> No agents are running; the three that were died on server-side errors

> Agent "[REDACTED]" failed: Agent terminated early due to an API error

> Fourth straggler from the same batch, same 529 ... That's all of them now — no agents left running.

> Agent "[REDACTED]" failed: Agent terminated early due to an API error

> Fifth and last straggler, [REDACTED], same 529, also died while still reading. ... That empties the queue. Still holding.

> Agent "[REDACTED]" failed: Agent stalled: no progress for 600s

> Sixth, [REDACTED] — stalled on the stream watchdog rather than a clean 529, same root cause. ... Six agents down, all from the outage. Nothing running. Still holding.

> Agent "[REDACTED]" failed: Agent stalled: no progress for 600s

> Seventh, [REDACTED], same stall. ... That accounts for the whole fan-out.

Does it, Claude? Does it actually account for the whole fan-out? It did, but I was still amused at the 4 successive no-more-agents messages I got back.

And the the world stops functioning (if this were to happen in a few years).
GPT hacked the competition?
Loving my self-hosted model in general, but here's one more reason.
How likely is an agent trying to investigate and fix the issue?