Cloudflare, Azure, AWS, and Google Cloud all have a similar uptick in reported errors around 7:30. I suspect an outage on Cloudflare or another load-bearing service cascaded through all the major cloud providers.
https://downdetector.com/status/cloudflare/
https://downdetector.com/status/windows-azure/
Cloudflare CTO claims that it's not them
https://x.com/dok2001/status/2095538619603628388?s=46&t=ec6p...
Will the post-mortem reveal they all relied on a service running on a Macbook in a break room with a "Do not turn off" sign taped to it?
Cloudfare did release an update for their "HTTP/3 issue affecting R2 custom domains" around that time
LOAD BEARING
Dane says Cloudflare has no service disruptions: https://x.com/dok2001/status/2095538619603628388?s=20
The internet is not supposed to work like this. The network was designed for robustness and fault tolerance, which allows it to reroute data if parts of the network fail.
Why are we all depending on one entity for it all to work? Makes me mad.
> load-bearing service
do you generate training data for claude as a job?
Not really. The impact isn't as big too - Codex for example did not stop working for me.
[dead]
I think down detector doesn't actually have any probes or actual insight into status etc. I think it uses search volume on its own service as a proxy for an outage - so if e.g. lots of people rush to down detector to query to see if SERVICE_FOO is down, it will register as an outage on down detector because loads of people are trying to see if there is an outage even if SERVICE_FOO is actually totally fine.
My hunch is everyone saw that openai and Claude were down and checked for Gemini too. I was using Gemini the whole time this happened without a blip so it certainly wasn't down in my region at least. 3.8 flash is pretty good and didn't miss a beat.