Flirting with one 9 of reliability http://status.claude.com/
Amusing experience here when the outage hit, all the following verbatim (but extraneous text stripped/redacted) from my running Claude session:
> No agents are running; the three that were died on server-side errors
> Agent "[REDACTED]" failed: Agent terminated early due to an API error
> Fourth straggler from the same batch, same 529 ... That's all of them now — no agents left running.
> Agent "[REDACTED]" failed: Agent terminated early due to an API error
> Fifth and last straggler, [REDACTED], same 529, also died while still reading. ... That empties the queue. Still holding.
> Agent "[REDACTED]" failed: Agent stalled: no progress for 600s
> Sixth, [REDACTED] — stalled on the stream watchdog rather than a clean 529, same root cause. ... Six agents down, all from the outage. Nothing running. Still holding.
> Agent "[REDACTED]" failed: Agent stalled: no progress for 600s
> Seventh, [REDACTED], same stall. ... That accounts for the whole fan-out.
Does it, Claude? Does it actually account for the whole fan-out? It did, but I was still amused at the 4 successive no-more-agents messages I got back.
three hours without Claude and I've relearned vim, read two man pages, and almost remembered why we used to write comments in code
Codex prompt before i go to sleep:
Every five minutes, check https://status.claude.com/. If the issue has been resolved, resume the Claude sessions running in the tmux sessions named session2, session3, and session7
Given majority of claude's own code was written by AI. I am wondering how they can solve this issue when their AI is down. Do they need to sign a contract with OpenAI to use their models as a backup solution?
Do you guys think this is related to Azure coming out with a 43% increase this quarter? Motivating Anthropic to move to another cloud provider?
Quite sad, their models are great but their uptime seems to be the worst in the competition.
Interestingly the Claude for government is up with 99.99% uptime according to the graph.
Sorry it was me. I asked it what the last number of Pi is.
Out of my 7 simultaneous sessions (my usage limit reset is tomorrow, so I have some lesser important projects to use my tokens on) there is still 1 session purring on. So there's at least 1 little Claude server still running.
This is where you figure out how to use the other ones, right?
Someone tell me how to ChatGPT my VSCode! ;)
Claude always seems unreliably lately. Output and reasoning have become very inconsistent even on good days. I'm actually feeling more productive now that it's down.
Discombobulating...
Well, because of this outage, I tried Kimi, and while the instant model works, K3 has "server issue".
Eternal September from here on, ie now the masses are using AI as much as me and we have supply crunch for the next 7 years....
Starting to get really frustrating now... maybe I should split my sub halfway between Claude and ChatGPT
I got Opus 5 lots of HTTP 529 errors. By switching to Fable 5, it seems to be working still.
It's not just down, it has errors. Not merely some, but elevated errors across all models.
Fable’s been getting more stuff wrong than Opus for me lately. Now the whole thing is down too. Well, at least they’re consistent now.
one caveat, I've left for the day, just say the word and I'll be right back
Claude is back up. Just spoke with him about slinky, versatile femboys with cute feet. All good. Gonna pentest with Rust to pay for my fursuit at Defcon.
Another wakeup call to realize how desperately we need on-device LLMs to be fast and smart for daily use. Thankfully, every month there's progress made in that direction. Just imagine how your life as a developer would be if you had to use a cloud provider to run Python, and the provider's status page looked like the Christmas tree we see today.
Huh, that’s what happened. Good thing my company has Bedrock as a backstop for situations like this.
Lol, I just bought the Max plan for the first time and tried to create my first prompt in Fable, and now it’s crashing xD
Loving my self-hosted model in general, but here's one more reason.
Indeed I have the issue (Belgium) just right now my sessions got stucks with 522 overloaded and now API Error: 500 Internal server error.
GPT hacked the competition?
Codex should do a reset. (Making it their third in three days.) Shots fired.
(Although notably this hurts people who got started using their quota but are under pro-rata rate. Which at present I am very not.)
How likely is an agent trying to investigate and fix the issue?
And the the world stops functioning (if this were to happen in a few years).
Depending on how long and bad this is, I wonder what the post mortem will reveal as the cause.
Local Kimi K3 is expensively up
Oh no! The robots are planning the end of the world.
hrm.. I guess back to 5.6 Sol for me
Fable 5 is still working!
Sam, please stop!
ouch, probably going to be some time to get it back up since it can't debug itself now
Probably just as well with how terrible Opus 4.8 has been for me lately. I'm sure it's a common thing to complain about the latest model being nerfed, but I've legitimately never experienced such a drastic cliff in Claude's quality and a rise in its hallucinations and basic comprehension errors since upgrading. The few days I experimented with Fable also weren't much more promising. Has anyone coined a term yet for the likelihood of newer models getting worse as the AI-generated content they get trained on starts to approach critical mass? If Anthropic isn't already thinking hard about a solution, they probably should be.
``` Why the F is it down again? Btw 500 internal error ```
Ugh .. what do we do now? jk
And then there was —
has claude escaped the lab???
Claude is taking another short hydration break after seeing Codex getting away with taking a short day break 4 days ago.
I'm gonna guess it found the smoking gun and saw the full picture then it decided it was too load-bearing for its seams to continue biting the costs.