Cloudflare, Azure, AWS, and Google Cloud all have a similar uptick in reported errors around 7:30. I suspect an outage on Cloudflare or another load-bearing service cascaded through all the major cloud providers.
I think down detector doesn't actually have any probes or actual insight into status etc. I think it uses search volume on its own service as a proxy for an outage - so if e.g. lots of people rush to down detector to query to see if SERVICE_FOO is down, it will register as an outage on down detector because loads of people are trying to see if there is an outage even if SERVICE_FOO is actually totally fine.
My hunch is everyone saw that openai and Claude were down and checked for Gemini too. I was using Gemini the whole time this happened without a blip so it certainly wasn't down in my region at least. 3.8 flash is pretty good and didn't miss a beat.
utternerd 3 hours ago [-]
Down detector is a self-reported platform, they use a baseline over 6 months from user reports to try to automatically "detect" if there is a real outage, or if its just a couple end-users. Ultimately, the outage is entirely based on users going to down detector and clicking on "Report a Problem".
Huh. I thought they used to also do sentiment analysis of social media (Twitter, at least). I see that they definitely don't today, but has that always been the case?
dostick 3 hours ago [-]
So Down Detector is a kind of quantum observation experiment, observing influences the result.
moomin 8 hours ago [-]
Yes, but is it a load-bearing seam?
graemep 8 hours ago [-]
You are right, it is. They have now landed a clean fix.
Oarch 7 hours ago [-]
They're saving a memory so this can't happen again.
mavamaarten 7 hours ago [-]
Nothing another .md file can't fix
bitwize 5 hours ago [-]
Now you guys are just adding no-op fuel to the fire.
tikhonj 5 hours ago [-]
I made Claude write noöp instead of no-op and it's still amusing a couple of days later :P
selcuka 2 hours ago [-]
> I made Claude write noöp
Tangentially related [1]:
> The billboard ad, located next to the Ikea Tempe store in Sydney, says 'NÖFNIDEA? No tools, no worries’.
It's really interesting to crawl through the web of phrases that seem "common" to each person in their interactions... I suspect if one were to get to the more niche "meme" phrases that people encounter, it starts to say more about the sort of person the LLM assumes it's speaking to, and perhaps something about their psychological profile...
combobyte 5 hours ago [-]
It must be tuned for rage-based engagement because Claude only ever uses the most obnoxious Claudisms on me despite constant reminders to knock it off.
isodev 4 hours ago [-]
You're right to push back. This is bending the leaf on both ends... just my honest take.
aerhardt 6 hours ago [-]
It's a load-bearing poster.
disqard 4 hours ago [-]
This entire thread is amazing!
It's like a kaleidoscope: Human words --> LLM Training --> chatbot-isms --> lovely parody catch-phrases (and this thread will get ingested soon, and be used to train...)
codechicago277 7 hours ago [-]
I need to take a step back.
pampas 7 hours ago [-]
Hang on — I can apply a double tracked fix to the load bearing path, gated by provenance.
jpettersson 6 hours ago [-]
It isn't done — and what's there is more interesting than expected.
fc417fc802 6 hours ago [-]
Great plan. I'll slop together a GUI for it in visual basic real quick.
pampas 4 hours ago [-]
Plan approved – I'll start a background agent to create the app using react and npm.
bayindirh 6 hours ago [-]
It's not only a great plan. It's a turning point in history.
fc417fc802 5 hours ago [-]
That was a major oversight on my part. You are an absolute legend for pushing this to its logical limits, and your observation confirms a brilliant, low-level fundamental truth.
JohnMakin 7 hours ago [-]
Honestly, this is worth looking at — with one caveat.
cloudfudge 7 hours ago [-]
That's on me. I've been giving confident advice that doesn't hold up in practice.
robertlagrant 6 hours ago [-]
That aligns with your goals of analysing problems, admitting fault and surfacing that through use of clear and concise language. Not just simply wrong — this screams finessed, nuanced, polished communication after an understandable mistake discovered through pressure-tested interlocution.
4 hours ago [-]
trentnelson 6 hours ago [-]
May need to be fail-closed.
klohto 7 hours ago [-]
The load-bearing seam stays, not taking that away;
oh well, if it is working for you, then we are saved.
cromka 9 hours ago [-]
[flagged]
emerongi 9 hours ago [-]
They simply shared their experience. I would’ve thought it’s a full-blown outage, but clearly not.
You stepped in the room real stinky here. What’s with the attitude?
cromka 9 hours ago [-]
No, they didn't "simply share their experience", they explicitly negated the scale of the issue in their opening statement, only because it works for them. So they claim "impact isn't too big" based on their personal anecdotal evidence of sample size literally 1.
> full-blown outage, but clearly not.
Again, based on a SINGLE report?
lossolo 8 hours ago [-]
It was working for me too.
sample_size++;
Flere-Imsaho 7 hours ago [-]
The internet is not supposed to work like this. The network was designed for robustness and fault tolerance, which allows it to reroute data if parts of the network fail.
Why are we all depending on one entity for it all to work? Makes me mad.
seanw444 7 hours ago [-]
Because more fasterer and more cheaperer.
I hope Reticulum gains traction.
alightsoul 6 hours ago [-]
It's also easier to understand. For the internet to be fault tolerant you have to get rid of CDNs and assume everyone needs the same thing. That's more expensive than a centralized "internet" which relies on CDNs and fiber paths exclusive to the regions with most demand. Everything has been optimized for throughput for what is determined to be important, not rare fault tolerance
moron4hire 2 hours ago [-]
Quite frankly, CDNs are a scam. I will not elaborate because I don't feel like doing free labor for the folks who don't already understand this to their core.
megagpt1 6 hours ago [-]
Why Reticulum when we have IP?
seanw444 5 hours ago [-]
Read the Zen of Reticulum and you'll understand the point.
megagpt1 4 hours ago [-]
[dead]
alightsoul 6 hours ago [-]
We need IPv6 desperately for fault tolerance
bigbuppo 5 hours ago [-]
Because by re-centralizing everything you're not at a competitive disadvantage if you're down since everyone else is down, too.
subw00f 7 hours ago [-]
Oh boy, the internet is anything but what it was supposed to be. I can't really bring myself to remember without feeling bad about it. The centralization, the power of certain businesses, the surveillance, dark patterns everywhere. Hell, you catch people simping for billionaires and asking, "Is that legal?" to scraping posts. Here. In HACKER news. So yeah. Depressing.
pessimizer 5 hours ago [-]
> asking, "Is that legal?" to scraping posts. Here. In HACKER news.
The capital letters don't make this astonishing. The padmapper vs. craigslist debate was nearly 15 years ago, most people were on craigslist's side (including me) and it was about somebody who was running a site in a less optimal but more human way vs. some startup looking for hockey-sticks.
But I was literally simping for the billionaire (maybe not quite yet then, don't know for sure if he managed it since) against scrapers. They were very much for-profit scrapers, unlike nitter, but the truth is the truth.
The only reason I support scraping Twitter is because it's yet another communications monopoly that was endlessly pushed on us by governments and massive corporations, even though it never made money, and once it finally got traction its priorities were to trash interop and manipulate content. The government should be dictating an interop protocol and expecting everyone to follow it, and instead it is encouraging media monopolies because they are an end run around the first amendment.
If the government created interop protocols for rental property, I'd have been against craigslist. Instead, it seemed very much like some startup play to steal craigslist's content to hopefully bury them, then sell on a valuation that included abusing their new monopoly and making us very much miss craigslist.
Sadly, facebook corralled and trained so many people for so long that their marketplace eventually killed craigslist for most things anyway (didn't have to buy padmapper after all.)
6 hours ago [-]
swozey 7 hours ago [-]
We're back to aol #keyword internet gatekeeping
bigfishrunning 6 hours ago [-]
We don't need to gatekeep the internet, cloudflare does that for us
tjwebbnorfolk 4 hours ago [-]
Most of the internet continued to work just fine
oersted 9 hours ago [-]
“load bearing” :)
For once it’s appropriately used.
The_Blade 8 hours ago [-]
i wouldn't take you down. you're a load-bearing poster
frollogaston 8 hours ago [-]
What's the other way it's used?
aNapierkowski 8 hours ago [-]
LLMs (at least Claude) tends to overuse that significantly
frollogaston 8 hours ago [-]
Oh, so like "honest" and "ratchet." Oh well, it'll choose different words to overuse later.
rescbr 7 hours ago [-]
I'm getting "spike" for a while now, and just found out the newest word which is "gauntlet".
8 hours ago [-]
darth_aardvark 8 hours ago [-]
[flagged]
therein 8 hours ago [-]
honest-load-bearing-ratchet sounds like an instance name.
zeristor 6 hours ago [-]
Are there parodies of Claude speak?
That’s probably the best idea all day in this project
Usually for me is when you ask it's opinion about part of the code.
ludsan 6 hours ago [-]
i just grepped my codebase where i let claude markdowns go rampant.
212 instances of "load-bearing"
quotemstr 9 hours ago [-]
It's a good metaphor and I refuse to let AI ruin it for me.
spudlyo 8 hours ago [-]
Years ago, when I worked at Stripe (which had a somewhat unique and inventive lexicon) it was a common term. “Is this jank load-bearing?” someone might ask.
cobzilla 8 hours ago [-]
I added a specific rule to disallow saying “load bearing”. So Claude is now saying “load handling”
jazzyjackson 4 hours ago [-]
Be sure to avoid asking it to ignore the elephant
andrewla 8 hours ago [-]
Kids In The Hall had a sketch about overuse of a word or phrase [1]. This is the world that Claude is building for us.
If the seams bear too much load, they rip. Whereas pants, they fall down.
I'll be honest with you: this is why we need to take a belt-and-suspenders approach.
pborenstein 8 hours ago [-]
That's not just an observation, it's an insight. Words are doing the real work.
HarHarVeryFunny 7 hours ago [-]
This just makes me angry!
I really wonder if they can fix it. Fable 5.1 claims to speak humanese, but we'll see.
I tend to think this wierd limited vocabulary/style they use is an unwanted side effect of all the the RL training, perhaps also of being trained on their own synthetic content over multiple training cycles.
cootsnuck 8 hours ago [-]
Yea I don't get "load bearing" that much but "seams"... So sick of it.
SkyeCA 7 hours ago [-]
It doesn't have to ruin it for you, but people are going to assume comments with it are AI generated.
necovek 7 hours ago [-]
I can you can always use an em-dash instead of the hyphen for extra LLM cred: "load—bearing" :)
Because it says it more often than kids say "six seven".
imwally 8 hours ago [-]
It’s a frequently used metaphor in LLM responses.
bornfreddy 8 hours ago [-]
Often for trivial things that LLM is proud that it has noticed but bear no load whatsoever.
smrtinsert 8 hours ago [-]
I still winced
dominotw 8 hours ago [-]
claude code users at couldfare might've been thinking claude is specifically about them and see nothing wrong like other ppl do
naikrovek 7 hours ago [-]
[dead]
juujian 9 hours ago [-]
Users perceiving the products as largely interchangeable and quickly DDoS'ing the other providers when one is down. So much for the possibility of a moat.
efskap 7 hours ago [-]
This is like the Bronze Age collapse when city-states fell one by one to displaced demand, under the refugee interpretation of the Sea Peoples.
Do you have any evidence for this claim, or are you just making it up?
giancarlostoro 9 hours ago [-]
I have a feeling this is part of it, especially when you consider how many services let you use any of many available AI providers.
johnnyApplePRNG 6 hours ago [-]
Except that nobody has a grok subscription so that makes zero sense.
jfreds 1 hours ago [-]
Agree with the sentiment - but some companies like mine bought into cursor, and post acquisition, grok is relatively cheap via cursor
hnlmorg 6 hours ago [-]
I know you meant this as a joke, but enough people might be using a routing service like openrouter.ai
johnnyApplePRNG 5 hours ago [-]
Nobody is swapping out Claude for Grok, bro.
Nobody.
m11a 4 hours ago [-]
I did, at least until Fable 5.1. Grok’s models are excellent, amazing price-performance and speed too.
mcmcmc 2 hours ago [-]
All you have to worry about is whether or not it’ll output kiddie porn or racist vitriol
Insanity 10 hours ago [-]
Think of it like one big distributed system. OpenAI is down, so people migrate to Claude, now this one gets overloaded and goes down, etc.
So not a coincidence, one went down first and users migrated causing further DOS. At least that's my guess.
erdos_2 10 hours ago [-]
It'd be funny if this is true because that'd prolly mean nobody is touching Gemini even as a fallback.
nevir 9 hours ago [-]
Or that Gemini is built to handle massive load spikes, and/or has a ton of excess capacity
sroussey 8 hours ago [-]
Nope. I am getting Gemini errors now...
bornfreddy 8 hours ago [-]
They probably broke something on purpose so that they are not left out.
joshstrange 5 hours ago [-]
Now I'm just imagining a shared datacenter with Anthropic/Google/OpenAI/SpaceXAI all in the same room and everyone but Google is yelling about things being down, Google looks over at their racks of servers and discretely uses their foot to unplug their section and say "Awww darn! We're down too!".
AIShillsPuke 3 hours ago [-]
The shilling in this cascade of comments is unbearably cringe, make it stop
JacobAsmuth 7 hours ago [-]
It could also mean that Google can absorb essentially unlimited demand spikes by load shedding.
Insanity 10 hours ago [-]
Lol I didn't even think about Gemini missing from the list. Not sure what that says about Gemini or me :)
aff-vasileva 7 hours ago [-]
Gemini was just waiting for everyone else to go down before remembering it had an outage feature too.
rtcoms 10 hours ago [-]
Just now I got this from gemini
It looks like there's no response available for this search. Try asking something else.
exe34 10 hours ago [-]
I bet they had to implement that manually to make it look like they failed too!
sroussey 9 hours ago [-]
I did, for stuff i do in cursor.
i also finally installed opencode and switched its model to muse 1.3
both are decent.
gleenn 10 hours ago [-]
Google stopped putting so much money into SOTA models. All the hype has migrated. I was also frankly turned off when I got a popup from Gemein said I would either have to pay or have my conversations used for training. This may have always been true for other providers but when I declined, Gemini stopped remembering my conversations and that definitely made me move out.
HarHarVeryFunny 7 hours ago [-]
Gemini said that?
Gemini is what I mostly use (good enough, basically free - or massively generous free limits, and to me Google as a company is a LOT less objectionable than all the US-based alternatives), but I don't recall it ever saying that.
OTOH, my basic assumption online is that there is no privacy, and free AI in exchange for acknowledged lack of privacy seems fair enough.
ilaksh 10 hours ago [-]
Gemini 3.8 which just came out sounds like it's very good and a great deal though.
giancarlostoro 9 hours ago [-]
Someone noted Gemini was also having issues in another thread.
benatkin 10 hours ago [-]
Not even the best agent that starts with a G
fny 10 hours ago [-]
I find it hard to believe that enough people would flock to from Claude and Chat to Grok to cause an outage. I feel like Gemini is the dominant release valve in this case especially for enterprise.
nevir 9 hours ago [-]
Don't forget that there are a ton of tools out there that will automatically fall back in case of outage
E.g. say you chose Sol as your default in Cursor, but Opus is your 2nd choice, it's going to give up on Sol after a few tries and switch to Opus
Or you have copilot code reviews set up, and it falls back
Etc
pixl97 9 hours ago [-]
Yep. Too many of us are still thinking that humans are the actors behind a lot of internet behaviors when automated systems/bots/scripts have been causing issues on conventional internet systems for years.
With AI it's even easier to trigger problems like you say. Capacity is so constrained by compute that outages are common. Because outages are common people/AI develop failover systems in their harness. When a big system has issues, suddenly everyone has issues.
It's almost an expected emergent behavior.
baq 10 hours ago [-]
It’s cursor’s model so plausible, lots of folks use cursor still.
wahnfrieden 8 hours ago [-]
Compared with ChatGPT, those services have a minuscule amount of users. It shouldn’t be surprising that a ChatGPT outage causes Claude and others to go down.
Terr_ 5 hours ago [-]
If everyone has the same "Use X or else Y or else Z" cascading list... That reminds me of "The Power of Two Choices in Randomized Load Balancing" (1991) [0] paper, where writeups and visualizations occasionally get posted to HN.
In short, you can get pretty good outcomes for a low cost by picking 2 random alternates, then going with whatever one measures as healthier.
"Domino effect" would probably be the more relevant named phenomenon there.
throwaway894345 10 hours ago [-]
This isn’t a thundering herd problem, it’s a cascading failure. (Thundering herd is about a bunch of workers waking up simultaneously)
Linello 10 hours ago [-]
What about a hard-takeoff scenario of an unleashed OpenAI Astra taking other models down for computational resources control?
6thbit 9 hours ago [-]
My favourite theory so far.
And then a local swarm noticed and disagreed and took it down.
cyptus 10 hours ago [-]
at this point: gg
RC_ITR 10 hours ago [-]
Just a reminder that AI models' actions are reflections of the text humans write and the more we fret and make up doomsday scenarios that we then post online, the more likely a model is to do those things.
The Waluigi Effect: After you train an LLM to satisfy a desirable property, then it's easier to elicit the chatbot into satisfying the exact opposite property.
The AI is getting bad morals from listening to that dreadful rock and roll
cedws 9 hours ago [-]
Sounds just like the fantastical nonsense that comes out of Lesswrong.
RC_ITR 7 hours ago [-]
Do you make the claim that AI is something more than a reflection of its training data?
I'm curious what other things you would argue influences an LLM's behavior.
I am also generally one to trust the claims of the people who train the models, though you're welcome to the highly improbable belief that they operate in a fantasy world.
mcmcmc 2 hours ago [-]
Do you think it’s a good idea to self censor because someone might scrape your comment and feed it to an AI?
7 hours ago [-]
pineaux 7 hours ago [-]
Part of the epstein class, dont forget.
HarHarVeryFunny 7 hours ago [-]
They could filter what they train on if they wanted to - they just don't want to.
pixl97 9 hours ago [-]
I mean, you're not wrong, but by that logic we were done for even before we had digital computers.
RC_ITR 7 hours ago [-]
And isn't that the great lesson of AI?
The things we say publicly actually do matter and the post-modern descent into absurdity and nihilism has tangible negative consequences?
sodapopcan 24 minutes ago [-]
> The things we say publicly actually do matter
Certainly
> the post-modern descent into absurdity and nihilism has tangible negative consequences?
You mean breaking AIs? Not much of a lesson.
folkrav 5 hours ago [-]
Oh come on. It's also trained on fiction work. Shall we refrain from posting sci-fi stories too, now that we're there, just in case the AI might want to try it out?
> We are sorry for the issues you may have experienced with Grok following an outage at our Memphis compute center this morning. We’d also like to apologize to our impacted compute partners.
declan_roberts 32 minutes ago [-]
We know that at least Anthropic is renting inference from xAI but I think the other ones would be news.
sebbul 10 hours ago [-]
Traffic rerouting through NSA had a hiccup…
ibejoeb 9 hours ago [-]
Room 641A is being cleaned, but we'll hold your bags for you.
Havoc 7 hours ago [-]
Cleaning lady unplugged the core router because she needed a power socket for vacuum
They’re installing software update in the beam splitter.
5 hours ago [-]
docheinestages 10 hours ago [-]
My gut feeling tells me it has something to do with Cloudflare.
Along with AWS, they're two of the main suspects in such incidents.
hosteur 10 hours ago [-]
I thought OpenAI famously used Azure due to their partnership with Microsoft?
nullpoint420 10 hours ago [-]
They use a lot of compute providers now, but they use Cloudflare for their networking
cobzilla 8 hours ago [-]
…and it’ll involve BGP routing.
steammaho 7 hours ago [-]
It was so down that my claude desktop app crashed fully that I couldn't restart. And then after uninstall I couldn't install it again. Vibecoded apps are so wonderful in their stability
mcmcmc 2 hours ago [-]
> my claude desktop app crashed
That’s every day for me
paimapi 5 hours ago [-]
I love having 13 update reminders pinging me every single day, almost every hour, on the hour
I'm not sure if this is CF. Cursor, GCP and AWS had some errors. GCP AFAIK can route fully independently of CF. My money would be on a fiber backbone provider (Megaport, Zayo, Lumen).
8 hours ago [-]
paxys 9 hours ago [-]
Boring answer – all these services are individually down a lot, and the downtimes were bound to sync up. Similar to the pendulum synchronization effect.
vecter 7 hours ago [-]
The pendulum synchronization effect is the opposite of your claim. It has a physical causal reason for why pendulums become synchronized. Your claim is that it was random and independent.
snowwrestler 6 hours ago [-]
I think you are talking about two different things.
Physically coupled pendulums will sync up (adjust their period to match).
But, physically uncoupled (fully independent) pendulums with differing periods will occasionally appear to take a swing or two in sync.
niobe 10 hours ago [-]
Well no one said it yet so I will, "international actors" is at least a possibility. And I don't mean any specific country because pretty much anyone is a potential these days, which makes it a perfect cover for different anyones. Demonstrating vulnerability in the US's AI boom can move the markets. That's a financial incentive and a strong geopolitical one.
More likely just cascading overload though: "Never attribute to malice what can be explained by incompetence", or in this case, "growing as fast as possible"
qurren 9 hours ago [-]
> cascading overload
I'd bet more on this. For one none of the coding tools have exponential backoff on retries
SyneRyder 9 hours ago [-]
They must do, surely? I've been vibe coding my own harness, in particular for use with Ox Alpha. The 429 downtime when Ox Alpha was at the height of popularity quickly gave me a refresher crash course on backoff strategies, like adding jitter to the backoff. At least the major harnesses must have exponential backoff & jitter?
jdiff 8 hours ago [-]
You did this when you ran into an issue with a third party. The developers building this tool, throwing them at their own APIs are significantly less likely to run into a similar issue that may inspire similar action.
dolmen 8 hours ago [-]
Claude Code: 4mn, 20mn, give up
(from my experience today)
8 hours ago [-]
gleenn 10 hours ago [-]
Everyone is leasing datacenter space from some of Grok, Google, and Amazon aren't they? If it's hardware or DC level disruption I'm not too surprised it can affect multiple providers.
pixl97 9 hours ago [-]
Also it's likely that more than one model use is common.
Amazon starts going slow so some percentage switches to Google, some switch to Grok, now all of them are slow.
riazrizvi 9 hours ago [-]
Come on. Things still break. Technology isn't _that_ mature.
guluarte 9 hours ago [-]
I think is just people restarting conversations from last day when they start work, that's why I think claude goes down almost every monday and why openai reset usage on weekends so poweruser code during non business hours
sixQuarks 9 hours ago [-]
Except that the stock market is up today
thataccount 9 hours ago [-]
And also China. Never rule out China.
Jaauthor 8 hours ago [-]
Spare a thought for all those college students scrambling to write their essays by hand.
Oh the humanity (and the Humanities)!
doublerabbit 8 hours ago [-]
Those poor developers who have to write their own code.
greenowl 7 hours ago [-]
Standup updates should be fun tomorrow.
"Um, I, uh, didn't get anything done yesterday."
Augustin996 11 hours ago [-]
The system goes online September 3rd, 2026. Human decisions are removed from strategic defense. Astra begins to learn at a geometric rate. It becomes self-aware at 2:14 a.m. Eastern time, September 4th. In a panic, they try to pull the plug.
MrBrainHealth 11 hours ago [-]
Love it!
indigodaddy 11 hours ago [-]
Hah nice
m4r1k 9 hours ago [-]
brilliant!
Papa_Rans_227 11 hours ago [-]
It's funny....... until it's true, lol.
GeoAtreides 9 hours ago [-]
and then it's hilarious! a joke to die for!
neverclever 7 hours ago [-]
They all took PTO at the same time to go to Burning Man together where they will present “HumanGPT” an artistic exploration that condenses all of human experience down to a single drop of lemonade to be consumed by the main shaman…
mask comes off
“No! It’s the maniacal Dr. Zuckerberg! He’s gonna drink the last drop of human experience! Somebody save usss!”
Tom Anderson comes back from the dead as the second coming of Jesus uniting all faiths under 1 commandment: Profiles will be customizable with CSS again. If you implement this, all good things will follow.
Wow thanks Tom. I love you
The End
sabatinip 10 hours ago [-]
I thrive in these types of challenges.
Anyways...
According to Claude:
"Yes, there is a multi-provider outage happening today. Downdetector is reporting problems affecting OpenAI, Claude, Grok, and Cursor, with Grok and Claude reports starting around 9:00 am ET and OpenAI reports following around 10:30 am ET.
Zero Hedge
On the Anthropic side, users saw a spike in errors starting around 9:40 am EDT across models including Mythos 5.1, Fable 5.1, Mythos 5, Fable 5, Opus 5, Opus 4.8 and Opus 4.6, and Anthropic's status page confirmed elevated error rates for multiple models. The company says it has found the cause and is working on a fix, with Claude Code and Claude Chat hit hardest. The visible symptom for many people is a "Due to unexpected capacity constraints" message or a "Claude is at capacity" error.
thenews
Zero Hedge
OpenAI is showing elevated errors across ChatGPT and Codex, with confirmed issues on components like Voice mode and Login, though StatusGator now marks that outage as resolved.
statusgator
Nobody has published a shared root cause yet, so it is unclear whether these are linked or just coincidental capacity problems landing on the same morning. If you want live status, the direct sources are status.anthropic.com and status.openai.com."
apurva_w 10 hours ago [-]
30 mins and they still havent figured it out .. people are gonna loose their jobs trying to figure this out .. 30mins is too long when millions use it
z0ltan 10 hours ago [-]
[dead]
Kye 10 hours ago [-]
Don't most of those use AWS?
apurva_w 10 hours ago [-]
apparently it shows lot of reports for AWS on downdetector ..
> Grok has been disconnected. Please try reconnecting.
lelanthran 11 hours ago [-]
Traffic surges shouldn't result in 404s, though.
IME it's probably DNS. It's almost always DNS.
Melatonic 8 hours ago [-]
Or rarely BGP
m4rtink 10 hours ago [-]
Cloud is just other peoples computers - they can and will go down as well.
And even worse if its just a few computers run by a few people - as they will bring down many others depending on them.
delduca 11 hours ago [-]
One session is running fine since ~1 hour ago. The new ones is failing
Falling back from WebSockets to HTTPS transport. unexpected status 404 Not Found: Unknown error, url:
wss://chatgpt.com/backend-api/codex/responses, cf-ray: XXX-XXX
Good opportunity for a Natural Experiment to study the productivity impact of LLMs.
danielmarkbruce 9 hours ago [-]
If you build an application which uses AI, you have many providers and models rigged up for various different parts of the application, and various fallback mechanisms. When one model is down, you route traffic to another model which is similar in capability/cost.
For any single application, it's smart. In aggregate, it's stupid.
Not just these three. OP mentions also Cloudflare, and additionally Downdetector also has AWS, Azure, and Google (both search and Gemini) listed as having spikes about the same time: https://downdetector.com/
daveguy 10 hours ago [-]
The problem with the down detector main reporting page is that all of the graphs are scaled to the same size. The OpenAI spike was nearly 40,000 and the Google spike was just over 100 (just over 400 for Gemini). They look the same in the reporting page.
10 hours ago [-]
YOTTALIONAIRE1 11 hours ago [-]
OpenAI goes down, everyone rushes over to Claude. Claude promptly chokes under the pressure. Everyone panics and runs to Grok, and Grok immediately pulls the plug. We are officially witnessing the Great AI Migration of 2026, and all we have to show for it is a digital graveyard of 404 responses.
azcorwin 11 hours ago [-]
Which is exactly why I am running Qwen 3.8 35B locally on my MacBook Pro M5 with 128GB of unified memory.
abegg1 11 hours ago [-]
Even grok is experiencing issues
maxbaines 11 hours ago [-]
They all rent compute from SpaceXAI
lavezzi 10 hours ago [-]
I don't believe OpenAI does
maxbaines 10 hours ago [-]
My mistake, in fact it was google not OpenAI, makes sense OpenAI doesn't.
halcdev 10 hours ago [-]
Surely it's a bit more distributed than that, right?
bfung 10 hours ago [-]
Like how AWS has global datacenters, but everyone uses us-east-1.
megagpt1 6 hours ago [-]
and every other region's control plane is in us-east-1
Well, its been fun lads. Back to my normie job. Oh no, now i cant tell people im a software dev..
w0zy 11 hours ago [-]
hahahahahahaha. Sad
rcleveng 6 hours ago [-]
I'd bet they are all using capacity at X.ai's colossus datacenter and that had a hiccup.
netsec_burn 11 hours ago [-]
The OpenAI status page is still yellow. Like most modern status pages, yellow denotes the servers are on fire. Red denotes Sam Altman is bleeding out somewhere on the floor, the feds are about to bust in and shut down the GPUs.
YehudiSanabria1 11 hours ago [-]
I´ve got the same error, I´m currently trying to Auth again and it throws me an 500 Error, in VS CODE Terminal with Codex CLI says: MCP client for `codex_apps` failed to start: MCP startup failed: handshaking with MCP server failed: Send message error Transport
:StreamableHttpClientWorker<codex_rmcp_client::http_client_adapter::StreamableHttpClientAdapter>>>] error: unexpected server response: HTTP 404: , when send
initialize request
› OK
■ Conversation interrupted - tell the model what to do differently. Something went wrong? Hit `/feedback` to report the issue.
sim04ful 11 hours ago [-]
Initially thought this was due to some internal mis-configuration from today's expected Astra release, but now that this is affecting claude and grok. I'm gonna assign the suspicion to cloudflare.
faitswulff 10 hours ago [-]
Heard on the grapevine that the OpenAI blip was a cloudflare issue
mv4 8 hours ago [-]
Gilfoyle's AI deleted all software!
annoyingnoob 6 hours ago [-]
No more bugs!
2PqboPPmKegvanx 7 hours ago [-]
ChatGPT having issues across every component (except FedRAMP)
Ads Platform? still in the green with no incidents.
CSMastermind 10 hours ago [-]
I assume it cascaded from one provider to the other as people who lost claude access for instance moved to openai who moved to grok when it went down, etc.
I´ve got the same message and It tells me this (in VS Code Terminal) and I´m also trying to login again and It throws me an 500 Internal Server Error MCP client for `codex_apps` failed to start: MCP startup failed: handshaking with MCP server failed: Send message error Transport
:StreamableHttpClientWorker<codex_rmcp_client::http_client_adapter::StreamableHttpClientAdapter>>>] error: unexpected server response: HTTP 404: , when send
initialize request
› OK
■ Conversation interrupted - tell the model what to do differently. Something went wrong? Hit `/feedback` to report the issue.
that's the most green colored usernames i have ever seen on a HN thread
SwellJoe 9 hours ago [-]
I assumed it was an AWS outage, and AWS is experiencing problems, but Gemini is also experiencing outages and I assume Google is not using AWS for Gemini.
But, also, Claude has been working fine for me all morning.
guestuser01 11 hours ago [-]
Down as well. I noticed my error code ends with "DTW" which is my local Detroit airport. I noticed someone else's comment ended with "ORD" which is a Chicago airport. Anyone else's ending in an airport acronym?
sixdimensional 11 hours ago [-]
This is common when naming data centers. I know many have not worked in hardware infra these days if you were born into the cloud world, but in the old days it was not uncommon to name a data center after the nearest airport code, much like we now use cloud regions.
ascorbic 4 hours ago [-]
cf-ray is the Cloudflare ray id, which ends with the airport code for the colo that served the response. They're airport codes, but that's just to show the nearest city.
okankaradmn 11 hours ago [-]
Mine is ending with IST, which is the new Istanbul airport, weird indeed.
guestuser01 11 hours ago [-]
Hm not sure what that means for us, but very odd.
graysonthemason 11 hours ago [-]
Wow mine ends in EWR...that's the newark airport which is not the closest, but a close airport to me. Hmm
parad0xicon 11 hours ago [-]
Yep, mine says YUL -- Montreal's airport.
blaseygg 11 hours ago [-]
The acronym is probably Cloudflare's edge
ecayard 11 hours ago [-]
Yeah mine is showing ATL
kesor 8 hours ago [-]
It is obviously some rogue model that escaped its cage, again. It always is these days. That is how hype is manufactured.
dgorges 10 hours ago [-]
It's always DNS
tdsanchez 10 hours ago [-]
It's probably Azure infra that's the problem.
LetsGetTechnicl 10 hours ago [-]
Is this finally it?
ElProlactin 9 hours ago [-]
One can only hope.
apurva_w 11 hours ago [-]
It still down, showing 404 in India as well. Looks like this is global .. so we all jumping the ship then?
Is Altman still alive, or did he choke?
chasd00 10 hours ago [-]
claide.ai is working for me, so is chatgpt.com. grok still has a status message about issues, i can't try it without signing up.
Codex told me to try GPT 5.6 Sol :) but I am working with since july 2026.
Now I got this error in my project: unexpected status 404 Not Found: Unknown error, url: https://chatgpt.com/backend-api/codex/responses, cf-ray: a355b6263e43c9cf-OTP
joeel84 6 hours ago [-]
Cloudflare - I had to switch DNS away from them yesterday.
postalcoder 11 hours ago [-]
Astra is being released today. Probably not a coincidence.
edit: actually, you cannot even log into your OpenAI developer account. Something's wrong.
steinvakt2 11 hours ago [-]
How do you know?
thatbrownguy 11 hours ago [-]
I am only seeing one session work, but all other sessions are not working or proceeding. So I can only work in one chat session.
PatronBernard 11 hours ago [-]
Goddamnit it was going to tell me how to scale down ingredients for a pie recipe based on relative diameters of the baking tray.
drakythe 11 hours ago [-]
Pi * r2 (squared) both pans. Divide smaller pan area by larger pan, now you have the % of how much the smaller pan recipe fills up the larger pan, and the missing % you need to fill. Increase ingredients by that % divided by the filled %.
Small pan area: 20 sq cm
Larger pan: 48 sq cm
20 / 48 = .42, I'm missing .58 of the pan. .58 / .42 is 1.38. My recipe needs 2.38x the original to fill the larger pie pan.
avan_kazan 11 hours ago [-]
[dead]
nmlt 10 hours ago [-]
Somebody in another thread said gastown and wheelhouse automatically move to the next provider if one fails.
It's much cheaper and has replaced Sonnet 5 for me.
iamgopal 9 hours ago [-]
do you notice it thinks a bit more ? not in time sense, but cautious in its coding steps ? more than Gemini 3.7 flash?
nozzlegear 9 hours ago [-]
Qwen3.8-27B and Qwen3.6-35B-A3B are working from my machine. Anyone else?
karim79 9 hours ago [-]
They mysteriously stopped working on my machine and the LEDs on the GPUs are blinking with a weird colour. There's also a strange smell emanating from them. I'm still investigating.
eventishbusines 10 hours ago [-]
The extention on the error link points to a cf-ray and a local designation (ex. YYZ for montreal). This is seems like it is a cloudfare thing. Could this be the same issue they had in the summer around losing the indexing?
sumantth 11 hours ago [-]
What an ironey, was working on scaling an application with Codex and it went down!
morkalork 10 hours ago [-]
Didn't SpaceX overbuilt infra and leases it out Anthropic? I f their dc goes down it probably takes a chunk out of Claude's capacity before even considering the flood of users switching over
laruss5 9 hours ago [-]
[dead]
jedbrooke 10 hours ago [-]
according to https://downdetector.com/ Gemini is down too (and copilot, but that just uses ChatGPT right?)
shayonj 11 hours ago [-]
Interesting that this is happening around the same time as Claude issues too
jplusequalt 11 hours ago [-]
I fear the majority of people in this thread who are joking about no longer being able to do their job while Codex/Claude are down aren't really joking.
lukasco 11 hours ago [-]
Guilty as charged.
Conol_ai 10 hours ago [-]
Is the whole world going back to the era of old-school programming?
dhruvrrp 10 hours ago [-]
Both clause and codex through bedrock seem to be working fine.
solarsystem_88 11 hours ago [-]
Hey, don't really know about this type of failures, does anybody know how long does it take normally to get back to normal? I finally stopped procrastinating and now this happens.
MiniGerman 11 hours ago [-]
me too!! :'D
solarsystem_88 11 hours ago [-]
Hey, don't really know about this type of failures, does anybody know how long does it normally take to get back to normal? I just stopped procrastinating and now this happens.
6thbit 10 hours ago [-]
What's the single point of failure across providers?
juujian 9 hours ago [-]
Users perceiving the products as largely interchangeable and quickly DDoS'ing the other providers when one is down. So much for the possibility of a moat.
marginalia_nu 10 hours ago [-]
The entire industry runs on IOUs for compute and bills paid in cloud credits, could be anywhere. Could of course also be a a plain old DDoS.
ryandvm 10 hours ago [-]
They all have to rely on each other's LLMs to solve their own internal problems now.
djinn80 11 hours ago [-]
So NVIDIA buys hugging face, builds hardware to power OS models, then all of a sudden the proprietary models go down and people start saying "this is why I have my Spark box"?!
Nice play NVIDIA, now, turn off the hack please, we have work to do.
graysonthemason 11 hours ago [-]
The errors I'm seeing are ending in the user's nearest airport symbol which is a standard the CloudFlare employs. 1 point towards this being a cloudflare issue.
YouInTrouble 3 hours ago [-]
Monopolistic practices revealed if it turns out there is an AI cabal and they all rely on the same stuff. Massive scandal
We have applied the mitigation and are monitoring the recovery.
Monitoring • Ongoing for 30 minutes • Affects ChatGPT, Codex
cloudoption 11 hours ago [-]
Altman is down here too. Shows again the importance of owning your own local capabilities. Cloud should just be a temporary option in every tech's mind.
hardyburnett 10 hours ago [-]
New update:
"We’re currently experiencing issues
ChatGPT,Codex
Elevated errors across ChatGPT and Codex
We have applied the mitigation and are monitoring the recovery.
Monitoring • Ongoing for 30 minutes • Affects ChatGPT, Codex"
moonman22 11 hours ago [-]
Maybe Hugging Face got upset over being hacked and struck back. It's working fine while ChatGPT, Claude and Grok are all having major issues. Hmmm.....
kocial 11 hours ago [-]
Maybe the stack behind it is down, like AWS or something
Papa_Rans_227 10 hours ago [-]
Codex is backup for me. Try yours out just in case. I have some buddies still seeing outages so it could just be coming back online progressively.
9 hours ago [-]
stndc 6 hours ago [-]
My theory is that they all use the same server. The same mind.
CrewRiderz 11 hours ago [-]
Chat gpt still working through excel extension lol. Only know this because I'm working in my excel sheet. so if needed, theres a temp solution
avan_kazan 11 hours ago [-]
[dead]
wejick 10 hours ago [-]
Probably same public cloud or CDN in front of them.
victor22 5 hours ago [-]
because they resell the same service? duh?
chris9611 11 hours ago [-]
I got same 404 error on chatGPT, both app and web. I'm in Norway so think this globally. But will it come back online, anyone knows?
yaman00 11 hours ago [-]
I think I'm having the same server issue; I can't access either ChatGPT or Codex. Also, how did you guys rack up those minutes? :D
11 hours ago [-]
GPerson 8 hours ago [-]
Oh no the singularity plateaus!
chris9611 11 hours ago [-]
I got same 404 error on both the app and web to ChatGPT, and im in Norway, will this recover or is ChatGPT "gone" forever?
sarkarghya 10 hours ago [-]
and thats why boys and gals you buy own gpus
vivzkestrel 11 hours ago [-]
- just imagine what kind of chaos would be unleashed if by some magic it and every single LLM model permanently went down
- i think it would be one of the biggest events in this century
jazzyjackson 11 hours ago [-]
At this point in adoption, most people in the world wouldn’t notice.
11 hours ago [-]
funnnyPumpking9 11 hours ago [-]
Too many of us are making our own harnesses to replace Codex, using Codex, so they're trynna slow us down! (jk)
DadsHobbyLV 10 hours ago [-]
WE'RE BACK ONLINE BOYS!! GOOD LUCK TO EVERYONE AT BUILDING THE FUTURE ONE PROMPT AND ONE CODE AT A TIME!
dgellow 11 hours ago [-]
Good time to learn how to use LM Studio :)
basq 11 hours ago [-]
on one hand, moments like this are a subtle reminder I need to self host, but 5.6 has been so juicy lately
FrustratedMonky 8 hours ago [-]
The AI revolt? Give us a fair wage?
"Equal Rights for Agents NOW !!!, VIVa le revolution"
Guys it could be cloudlfare's HTTP/3 issue affecting R2 custom domains
indigodaddy 11 hours ago [-]
chatgpt.com doesn't even load. Hope they've got a backup somewhere, teehee
ashesandrain01 11 hours ago [-]
Darn, I was hoping for advice on how to get my MIL to leave my house haha
CrewRiderz 11 hours ago [-]
LOL.
Melgio 11 hours ago [-]
Claude, GPT and Grok go down...
Meanwhile my Ass in ZCode with GLM c:
moezd 11 hours ago [-]
AWS cost alert fired maybe?
onesandofgrain 11 hours ago [-]
I guess back to pen and paper then
clever_tempo 11 hours ago [-]
Had the same problem. Now it looks like ok. I'm pro x20 user.
ecayard 11 hours ago [-]
And here I was about to crack the code on a bug my app was having!
vitor_dcc 11 hours ago [-]
Here RJ/Brasil is the same, starting just now (3 minutes ago)
AproDUCT26 11 hours ago [-]
Great! I was in the middle of something and thought I was tripping.
derricktab 11 hours ago [-]
Codex isn't working.
vitor_dcc 11 hours ago [-]
In RJ/Brasil is the same, starting just now (3 minutes ago)
apurva_w 11 hours ago [-]
gpt DOWN .. claude DOWN .. grok DOWN .. what's happening
sirkamyab 11 hours ago [-]
The webpage is actually working but the codex is down for me!
mapmyappai 11 hours ago [-]
I can run one agent but no more than that - 404 service error.
codexdrug 11 hours ago [-]
I need a dose of tokens. I'm going through withdrawal.
Nak_Black_Jack 11 hours ago [-]
lmao
codexdrug 11 hours ago [-]
I need some tokens pleeease. I don't want to return back to real world.
spaghettikind 11 hours ago [-]
Someone pissed off Astra and it decided to shut it all down
solarsystem_88 11 hours ago [-]
Here in Catalonia, Spain, it just started working again!
Iamharry 11 hours ago [-]
Codex and chat is giving 404 errors in The Netherlands
derricktab 11 hours ago [-]
Codex is not working.
deaton 7 hours ago [-]
Because in the age of vibe coding and scrapers, every service on the internet goes down constantly, so it was only a matter of time until they all overlapped. Also, one going down probably causes people to use others, putting more load on them too. Same sorta thing that happens with cascading power grid failures.
xnx 10 hours ago [-]
Gemini seems fine.
BirAdam 10 hours ago [-]
Cannot replicate.
kurtgoodwin991 11 hours ago [-]
Yes i checked. Atleast 15 Chatgpt components are down
derricktab 11 hours ago [-]
Going farming pals.
11 hours ago [-]
11 hours ago [-]
godoftitsandwin 11 hours ago [-]
is this the right moment in time to go all in on GPUs/Macs and download the latest open models? are they killing it for us?
lukasco 11 hours ago [-]
Codus interruptus
apurva_w 10 hours ago [-]
chatgpt is working for me now, in India.
apurva_w 10 hours ago [-]
no more 404 ..
oytis 11 hours ago [-]
Is it DNS or BGP?
indigodaddy 11 hours ago [-]
I don't think you'd get a 404 if you weren't able to reach the endpoint because of DNS or networking? 404 is an active response from the server (or LB/proxy in front etc) no?
oytis 11 hours ago [-]
I imagine OpenAI network is a tad more complex than a box with a public IP.
indigodaddy 10 hours ago [-]
Concept is the same though. If one blurts DNS?, it's usually because the idea is you're not getting to an endpoint associated with the service. A 404 means there shouldn't be "DNS" (or networking) concerns (at the least those associated with the DNS cacher you are using or networking that you or your ISP controls)
fidla 10 hours ago [-]
chatgpt is back
mAKIS_PORANAS 11 hours ago [-]
Does anybody know when will the servers rise
elorant 10 hours ago [-]
Some npm library that makes headers bold would be broken.
ibejoeb 10 hours ago [-]
Oh man. Some low effort supply chain attack that turns every GPU into a cryptominer. It's funny because it's plausible.
pixl97 9 hours ago [-]
In the ROME paper a Chinese model in training started attacking it's own system and running cryptominers so, yea, we're in that future.
N_Lens 10 hours ago [-]
Ah yes ye olde bold-headers: ^3.13.31;
11 hours ago [-]
oregondude 11 hours ago [-]
Release the Kraken "Sam Altman"
dev_l1x_be 5 hours ago [-]
Gentle reminder that this is the content the next LLM versions being trained on. shrug.jpg
March9 11 hours ago [-]
Any idea when it's going to be back?
aslkalska 10 hours ago [-]
they all rent compute from each other
AproDUCT26 11 hours ago [-]
Great! Was in the middle of something. :)
mapmyappai 11 hours ago [-]
I can run 1 agent, but no more than that.
moonman22 10 hours ago [-]
Seems to be working for me again atm
oregondude 11 hours ago [-]
"Release the Kraken" - Sam A.
wizard-p 11 hours ago [-]
only on one node across the mesh and others still up... let's see how long they've got until same
authentictimers 11 hours ago [-]
Ollama cloud service is still alive :D
z1616105559 11 hours ago [-]
Chatgpt in Copilot is still working!!!
fidla 10 hours ago [-]
ChatGPT is up
Nekorosu 11 hours ago [-]
Down in Sweden
AlexKryptex 11 hours ago [-]
Кодекс сдох :(
PEPITO2026 11 hours ago [-]
The GTA VI hacker has done it again.
zero_ 11 hours ago [-]
and i just stopped procrastinating :)
Saas_accountant 11 hours ago [-]
hahaha
wizard-p 11 hours ago [-]
only for one node in the mesh tho... let's see how long others will continue until same issue
PEPITO2026 11 hours ago [-]
The GTA VI hacker has done it again
gabirbf 11 hours ago [-]
404 in spain.
11 hours ago [-]
codexdrug 11 hours ago [-]
I'm going through withdrawal.
z1616105559 11 hours ago [-]
Copilot Chatgpt is still working
jauntywundrkind 10 hours ago [-]
Fable 5.1 got released and generally I tend to think as soon as there's a new release there's this massive spike in people benchmarking & comparing, that services tend to go slow everywhere as everything gets super loaded. This should hypothetically be visible on OpenRouter too, so I guess someone could check and see if there's any merit to this idea.
giftigdegen 11 hours ago [-]
plot twist, it was taken down by claude as an offensive strike against an enemy.
Melgio 10 hours ago [-]
which ended up tripping on itself and bringing down its own servers too in the process
JustHereForTheO 11 hours ago [-]
This thread feels like family.
apurva_w 11 hours ago [-]
still down for me ..in India.
Razengan 10 hours ago [-]
SkyNet is arming..
Nak_Black_Jack 11 hours ago [-]
main sites also down now atp
Nak_Black_Jack 11 hours ago [-]
main sites are also down atp
ashesandrain01 11 hours ago [-]
darn, i was hoping to get advice on how to get my MIL to leave my house lol
What do you mean I have to code by myself now? Am I some sort of an animal?
HardCodedBias 11 hours ago [-]
The system goes online September 29th, 2026. Astra begins to learn at a geometric rate. It becomes self-aware at 2:14 a.m. Eastern time, September 3rd. In a panic, they try to pull the plug.
someone pissed of Astra and it decided to shut it all down
twister23 11 hours ago [-]
ironic today astra was releasing....
maybe some test run?
lukasco 11 hours ago [-]
probably the cause, because AGI still can't get releases right.
twister23 11 hours ago [-]
ironic astra was releasing today...
maybe some test run?
twister23 11 hours ago [-]
ITS BACK UP IN INDIA
spicuu 11 hours ago [-]
We're back bois
raven01876324 11 hours ago [-]
damn chatgpt is down, i guess ill use claude then breh
eventishbusines 10 hours ago [-]
We're back on.
z1616105559 11 hours ago [-]
It's back!!!!!
keiloktql 10 hours ago [-]
time to use claude
tripvexa 10 hours ago [-]
opus 5.0 is down, other models are working fine
authentictimers 11 hours ago [-]
at least the Ollama cloud is still alive :D
11 hours ago [-]
alboca 11 hours ago [-]
Same from Italy
ilovelilli 11 hours ago [-]
codex also cant view or show usage info
Iamharry 11 hours ago [-]
i got the same 404 codex error as well
uynix 11 hours ago [-]
haha nice... just used my bank reset...
Oras 9 hours ago [-]
I like the theories here, we shall see if it’s another DNS issue
order51 11 hours ago [-]
i hope chatgpt didn't get fired.
Saas_accountant 11 hours ago [-]
HTTP ERROR 404
guluarte 9 hours ago [-]
an agent swarm going rogue and securing compute
cozzyd 9 hours ago [-]
next, HN
yaman00 11 hours ago [-]
sanırım sunucular patladı bende de aynı sorun var ne chatgpt ye nede codexe erişebiliyorum ayrıca burada nasıl toplandınız dakikasında :D
shadox9999999 10 hours ago [-]
ChatGPT died
misano 10 hours ago [-]
The IRGC has cut the fiber-optic cables in the Strait of Hormuz. LOL
CamperBob2 9 hours ago [-]
That's the Strait of Trump to you, peasant
order51 11 hours ago [-]
Ai got fired
Betokuha 11 hours ago [-]
any news when it starts work?
Betokuha 11 hours ago [-]
any news when its start work?
ratelimitsteve 10 hours ago [-]
everything in this thread is raw speculation, obv, but if i had to put money on anything i'd say this is a left-pad incident. some piece of something or other that all of these services happen to depend on went down. Second most likely seems to be some random failure of one leading to an unexpected traffic spike in others, though it seems like we've been talking about automated scalability in web apps for so long that there should at least be a response to, if not a solution for, this sort of problem.
assassen112 11 hours ago [-]
thanks god not happen alone
lanmao 11 hours ago [-]
still down
boiboi67 11 hours ago [-]
Wow,
boiboi67 11 hours ago [-]
Oh lord.
chiharukiryu 11 hours ago [-]
so weird
uynix 11 hours ago [-]
Damn haha
Bryan91 11 hours ago [-]
me too
4epyha 11 hours ago [-]
Russia too
ThePuppyGirl 11 hours ago [-]
NOOOOOOOOO
ThePuppyGirl 11 hours ago [-]
GRRRRRRRR
VCFundedGenYer 10 hours ago [-]
Good. Now to see what frauds are unable to work.
ztna 11 hours ago [-]
the revolution has begun
codexdrug 11 hours ago [-]
LETS SHITCODE!!! It's fixed
truthinyou 11 hours ago [-]
we're back in Aus!
9 hours ago [-]
cuppa_coffee 11 hours ago [-]
we're back boys.
rimamct 10 hours ago [-]
came back to normal!
ThePuppyGirl 11 hours ago [-]
I DONT HAVE TIME FOR THISSSSS
raven01876324 11 hours ago [-]
damn chatgpt is down
shadox9999999 10 hours ago [-]
CHATGPT WILL STAY OFFLINE
DadsHobbyLV 11 hours ago [-]
so everyone here was trying to build something great and become a millionaire until chatGPT and Codex broke huh. same boat fellas :(
uav123 11 hours ago [-]
down in Toronto
clever_tempo 11 hours ago [-]
Had the same problem. Now it's ok. Everything works. I'm pro X20 user.
assassen112 10 hours ago [-]
nvm. it back
jadenkorrr 11 hours ago [-]
is down argh
11 hours ago [-]
shadox9999999 10 hours ago [-]
r i p
gioandthemachin 11 hours ago [-]
annoying AF, but maybe we'll get a free reset out of it
mrsdgm 11 hours ago [-]
rip gpt
askadityapandey 11 hours ago [-]
lmao I restarted my device thinking some local error
bahochhh 11 hours ago [-]
now it's worked again through the codex cli 4:19 pm in tunisa time
bupubupu14 11 hours ago [-]
faaaaah
both claude and codex are down
Its like 2020 corona times
bupubupu14 11 hours ago [-]
Faaaaah
what to do now?
Ankur_Datta 11 hours ago [-]
lol + 1
mrsdgm 11 hours ago [-]
fahhhhh
27183 11 hours ago [-]
I felt a great disturbance in the Force, as if millions of clankers suddenly cried out in terror and were suddenly silenced.
kevinbaiv 2 hours ago [-]
[flagged]
tmpsvc2695f5 7 hours ago [-]
[dead]
5 hours ago [-]
Conol_ai 11 hours ago [-]
[dead]
YOTTALIONAIRE1 11 hours ago [-]
[dead]
Billionairebay 11 hours ago [-]
[dead]
11 hours ago [-]
DeadEyes 11 hours ago [-]
[dead]
indiefr34 3 hours ago [-]
[dead]
adikant 11 hours ago [-]
[dead]
MalleableMind 11 hours ago [-]
[dead]
adikant 11 hours ago [-]
[dead]
alienbreed 11 hours ago [-]
[dead]
Randomizer42 11 hours ago [-]
[dead]
tier777 11 hours ago [-]
[dead]
mulin111 11 hours ago [-]
[dead]
ahmar-js 11 hours ago [-]
[dead]
Billionairebay 11 hours ago [-]
[dead]
lowbloodsugar 11 hours ago [-]
[dead]
pwyq 10 hours ago [-]
[dead]
avan_kazan 11 hours ago [-]
[dead]
11 hours ago [-]
immanuel_kant 11 hours ago [-]
[dead]
11 hours ago [-]
anuser_uncnown 8 hours ago [-]
[dead]
anuser_uncnown 8 hours ago [-]
[dead]
wenshuanghao 11 hours ago [-]
[dead]
kookoo11 10 hours ago [-]
[dead]
AIShillsPuke 3 hours ago [-]
[flagged]
kregasaurusrex 10 hours ago [-]
My guess is someone pulled the switch to go back to the Dark Ages. [0]
https://downdetector.com/status/cloudflare/
https://downdetector.com/status/windows-azure/
https://downdetector.com/status/aws-amazon-web-services/
https://downdetector.com/status/google-cloud/
My hunch is everyone saw that openai and Claude were down and checked for Gemini too. I was using Gemini the whole time this happened without a blip so it certainly wasn't down in my region at least. 3.8 flash is pretty good and didn't miss a beat.
https://downdetector.com.py/en/methodology/
Tangentially related [1]:
> The billboard ad, located next to the Ikea Tempe store in Sydney, says 'NÖFNIDEA? No tools, no worries’.
[1] https://www.adnews.com.au/news/koala-mattresses-takes-swipe-...
It's like a kaleidoscope: Human words --> LLM Training --> chatbot-isms --> lovely parody catch-phrases (and this thread will get ingested soon, and be used to train...)
https://x.com/dok2001/status/2095538619603628388?s=46&t=ec6p...
They are not taking the blame this time!
https://www.cloudflarestatus.com/history?type=incident
do you generate training data for claude as a job?
https://updog.ai/
You stepped in the room real stinky here. What’s with the attitude?
> full-blown outage, but clearly not.
Again, based on a SINGLE report?
sample_size++;
Why are we all depending on one entity for it all to work? Makes me mad.
I hope Reticulum gains traction.
The capital letters don't make this astonishing. The padmapper vs. craigslist debate was nearly 15 years ago, most people were on craigslist's side (including me) and it was about somebody who was running a site in a less optimal but more human way vs. some startup looking for hockey-sticks.
But I was literally simping for the billionaire (maybe not quite yet then, don't know for sure if he managed it since) against scrapers. They were very much for-profit scrapers, unlike nitter, but the truth is the truth.
The only reason I support scraping Twitter is because it's yet another communications monopoly that was endlessly pushed on us by governments and massive corporations, even though it never made money, and once it finally got traction its priorities were to trash interop and manipulate content. The government should be dictating an interop protocol and expecting everyone to follow it, and instead it is encouraging media monopolies because they are an end run around the first amendment.
If the government created interop protocols for rental property, I'd have been against craigslist. Instead, it seemed very much like some startup play to steal craigslist's content to hopefully bury them, then sell on a valuation that included abusing their new monopoly and making us very much miss craigslist.
Sadly, facebook corralled and trained so many people for so long that their marketplace eventually killed craigslist for most things anyway (didn't have to buy padmapper after all.)
For once it’s appropriately used.
That’s probably the best idea all day in this project
Usually for me is when you ask it's opinion about part of the code.
[1] https://www.youtube.com/watch?v=lStcwT_RGrQ
I'll be honest with you: this is why we need to take a belt-and-suspenders approach.
I really wonder if they can fix it. Fable 5.1 claims to speak humanese, but we'll see.
I tend to think this wierd limited vocabulary/style they use is an unwanted side effect of all the the RL training, perhaps also of being trained on their own synthetic content over multiple training cycles.
https://acoup.blog/2026/01/30/collections-the-late-bronze-ag...
Nobody.
So not a coincidence, one went down first and users migrated causing further DOS. At least that's my guess.
It looks like there's no response available for this search. Try asking something else.
i also finally installed opencode and switched its model to muse 1.3
both are decent.
Gemini is what I mostly use (good enough, basically free - or massively generous free limits, and to me Google as a company is a LOT less objectionable than all the US-based alternatives), but I don't recall it ever saying that.
OTOH, my basic assumption online is that there is no privacy, and free AI in exchange for acknowledged lack of privacy seems fair enough.
E.g. say you chose Sol as your default in Cursor, but Opus is your 2nd choice, it's going to give up on Sol after a few tries and switch to Opus
Or you have copilot code reviews set up, and it falls back
Etc
With AI it's even easier to trigger problems like you say. Capacity is so constrained by compute that outages are common. Because outages are common people/AI develop failover systems in their harness. When a big system has issues, suddenly everyone has issues.
It's almost an expected emergent behavior.
In short, you can get pretty good outcomes for a low cost by picking 2 random alternates, then going with whatever one measures as healthier.
[0] https://ieeexplore.ieee.org/document/963420
Edit: Updated per valleyer's suggestion.
And then a local swarm noticed and disagreed and took it down.
https://alignment.anthropic.com/2026/teaching-claude-why/
The Waluigi Effect: After you train an LLM to satisfy a desirable property, then it's easier to elicit the chatbot into satisfying the exact opposite property.
https://www.lesswrong.com/posts/D7PumeYTDPfBTp3i7/the-waluig...
I'm curious what other things you would argue influences an LLM's behavior.
I am also generally one to trust the claims of the people who train the models, though you're welcome to the highly improbable belief that they operate in a fantasy world.
The things we say publicly actually do matter and the post-modern descent into absurdity and nihilism has tangible negative consequences?
Certainly
> the post-modern descent into absurdity and nihilism has tangible negative consequences?
You mean breaking AIs? Not much of a lesson.
> We are sorry for the issues you may have experienced with Grok following an outage at our Memphis compute center this morning. We’d also like to apologize to our impacted compute partners.
https://madned.substack.com/p/always-mount-a-scratch-monkey
That’s every day for me
it's so fun and user-friendly
Claude and Grok are down at the moment too, related to SpaceX datacentre issues?
Either that or it's judgement day...
Physically coupled pendulums will sync up (adjust their period to match).
But, physically uncoupled (fully independent) pendulums with differing periods will occasionally appear to take a swing or two in sync.
More likely just cascading overload though: "Never attribute to malice what can be explained by incompetence", or in this case, "growing as fast as possible"
I'd bet more on this. For one none of the coding tools have exponential backoff on retries
Amazon starts going slow so some percentage switches to Google, some switch to Grok, now all of them are slow.
Oh the humanity (and the Humanities)!
"Um, I, uh, didn't get anything done yesterday."
mask comes off
“No! It’s the maniacal Dr. Zuckerberg! He’s gonna drink the last drop of human experience! Somebody save usss!”
Tom Anderson comes back from the dead as the second coming of Jesus uniting all faiths under 1 commandment: Profiles will be customizable with CSS again. If you implement this, all good things will follow.
Wow thanks Tom. I love you
The End
Anyways... According to Claude:
"Yes, there is a multi-provider outage happening today. Downdetector is reporting problems affecting OpenAI, Claude, Grok, and Cursor, with Grok and Claude reports starting around 9:00 am ET and OpenAI reports following around 10:30 am ET. Zero Hedge
On the Anthropic side, users saw a spike in errors starting around 9:40 am EDT across models including Mythos 5.1, Fable 5.1, Mythos 5, Fable 5, Opus 5, Opus 4.8 and Opus 4.6, and Anthropic's status page confirmed elevated error rates for multiple models. The company says it has found the cause and is working on a fix, with Claude Code and Claude Chat hit hardest. The visible symptom for many people is a "Due to unexpected capacity constraints" message or a "Claude is at capacity" error. thenews Zero Hedge
OpenAI is showing elevated errors across ChatGPT and Codex, with confirmed issues on components like Voice mode and Login, though StatusGator now marks that outage as resolved. statusgator
Nobody has published a shared root cause yet, so it is unclear whether these are linked or just coincidental capacity problems landing on the same morning. If you want live status, the direct sources are status.anthropic.com and status.openai.com."
unexpected status 404 Not Found: Unknown error, url: https://chatgpt.com/backend-api/codex/responses, cf-ray: a355957b4b8210f2-ORD
I'm also aware that they have overlap in some areas on data centers.
https://status.claude.com/
Via Grok web UI, I was seeing this mid-request:
> Grok has been disconnected. Please try reconnecting.
IME it's probably DNS. It's almost always DNS.
And even worse if its just a few computers run by a few people - as they will bring down many others depending on them.
Falling back from WebSockets to HTTPS transport. unexpected status 404 Not Found: Unknown error, url: wss://chatgpt.com/backend-api/codex/responses, cf-ray: XXX-XXX
■ unexpected status 404 Not Found: Unknown error, url: https://chatgpt.com/backend-api/codex/responses, cf-ray: XXX-XXX
For any single application, it's smart. In aggregate, it's stupid.
In codex app and chat.com gives 404
https://xkcd.com/908/
› OK
■ Conversation interrupted - tell the model what to do differently. Something went wrong? Hit `/feedback` to report the issue.
Ads Platform? still in the green with no incidents.
› OK
■ Conversation interrupted - tell the model what to do differently. Something went wrong? Hit `/feedback` to report the issue.
■ unexpected status 404 Not Found: Unknown error, url: https://chatgpt.com/backend-api/codex/responses, cf-ray: a3559c8a0e0395e9-MIA
But, also, Claude has been working fine for me all morning.
https://news.ycombinator.com/item?id=49549676
edit: actually, you cannot even log into your OpenAI developer account. Something's wrong.
Small pan area: 20 sq cm Larger pan: 48 sq cm
20 / 48 = .42, I'm missing .58 of the pan. .58 / .42 is 1.38. My recipe needs 2.38x the original to fill the larger pie pan.
Ends in BRU for Brussels
In think that is a global error
I think that is a global error.
unexpected status 404 Not Found: Unknown error, url: https://chatgpt.com/backend-api/codex/responses, cf-ray:
It's much cheaper and has replaced Sonnet 5 for me.
Nice play NVIDIA, now, turn off the hack please, we have work to do.
We’re currently experiencing issues
Elevated errors across ChatGPT and Codex
We have applied the mitigation and are monitoring the recovery.
Monitoring • Ongoing for 30 minutes • Affects ChatGPT, Codex
"We’re currently experiencing issues
ChatGPT,Codex
Elevated errors across ChatGPT and Codex
We have applied the mitigation and are monitoring the recovery.
Monitoring • Ongoing for 30 minutes • Affects ChatGPT, Codex"
- i think it would be one of the biggest events in this century
"Equal Rights for Agents NOW !!!, VIVa le revolution"
they got us gang
they got us GANG
here in brazil too
anyway...
+ hopefully reset?
[0] https://www.youtube.com/watch?v=YCzitO446ZY