Ask HN: Why were OpenAI, Claude, and Grok simultaneously down?

235 pointsposted 4 hours ago
by halcdev

Item id: 49551096

453 Comments

kibae

3 hours ago

Cloudflare, Azure, AWS, and Google Cloud all have a similar uptick in reported errors around 7:30. I suspect an outage on Cloudflare or another load-bearing service cascaded through all the major cloud providers.

https://downdetector.com/status/cloudflare/

https://downdetector.com/status/windows-azure/

https://downdetector.com/status/aws-amazon-web-services/

https://downdetector.com/status/google-cloud/

moomin

2 hours ago

Yes, but is it a load-bearing seam?

graemep

an hour ago

You are right, it is. They have now landed a clean fix.

Oarch

an hour ago

They're saving a memory so this can't happen again.

mavamaarten

16 minutes ago

Nothing another .md file can't fix

klohto

5 minutes ago

The load-bearing seam stays, not taking that away;

codechicago277

an hour ago

I need to take a step back.

pampas

7 minutes ago

Hang on — I can apply a double tracked fix to the load bearing path, gated by provenance.

JohnMakin

an hour ago

Honestly, this is worth looking at — with one caveat.

cloudfudge

an hour ago

That's on me. I've been giving confident advice that doesn't hold up in practice.

martyfunkhouser

3 hours ago

Will the post-mortem reveal they all relied on a service running on a Macbook in a break room with a "Do not turn off" sign taped to it?

buredoranna

33 minutes ago

Someone probably set it to "magic", when everyone knows its supposed to be set to "more magic".

mcphage

2 hours ago

Actually it said "Beware of the Leopard".

RobotToaster

2 hours ago

If it was still running snow leopard that would explain it

onetokeoverthe

34 minutes ago

if only the fools selling vintage macbook pros for $200 were smart enough to still have them on SL OS.

dominotw

an hour ago

> load-bearing service

do you generate training data for claude as a job?

Flere-Imsaho

an hour ago

The internet is not supposed to work like this. The network was designed for robustness and fault tolerance, which allows it to reroute data if parts of the network fail.

Why are we all depending on one entity for it all to work? Makes me mad.

seanw444

an hour ago

Because more fasterer and more cheaperer.

I hope Reticulum gains traction.

subw00f

an hour ago

Oh boy, the internet is anything but what it was supposed to be. I can't really bring myself to remember without feeling bad about it. The centralization, the power of certain businesses, the surveillance, dark patterns everywhere. Hell, you catch people simping for billionaires and asking, "Is that legal?" to scraping posts. Here. In HACKER news. So yeah. Depressing.

swozey

25 minutes ago

We're back to aol #keyword internet gatekeeping

KptMarchewa

3 hours ago

Not really. The impact isn't as big too - Codex for example did not stop working for me.

https://updog.ai/

taytus

3 hours ago

oh well, if it is working for you, then we are saved.

cromka

3 hours ago

> Not really. The impact isn't as big too - Codex for example did not stop working for me.

Buddy, read the room. Just because it works for you doesn't mean it works for everyone. ESPECIALLY if the suggested issue here is, indeed, with Cloudflare.

emerongi

2 hours ago

They simply shared their experience. I would’ve thought it’s a full-blown outage, but clearly not.

You stepped in the room real stinky here. What’s with the attitude?

cromka

2 hours ago

No, they didn't "simply share their experience", they explicitly negated the scale of the issue in their opening statement, only because it works for them. So they claim "impact isn't too big" based on their personal anecdotal evidence of sample size literally 1.

> full-blown outage, but clearly not.

Again, based on a SINGLE report?

lossolo

2 hours ago

It was working for me too.

sample_size++;

oersted

3 hours ago

“load bearing” :)

For once it’s appropriately used.

The_Blade

an hour ago

i wouldn't take you down. you're a load-bearing poster

frollogaston

2 hours ago

What's the other way it's used?

aNapierkowski

2 hours ago

LLMs (at least Claude) tends to overuse that significantly

frollogaston

2 hours ago

Oh, so like "honest" and "ratchet." Oh well, it'll choose different words to overuse later.

rescbr

an hour ago

I'm getting "spike" for a while now, and just found out the newest word which is "gauntlet".

user

an hour ago

[deleted]

therein

2 hours ago

honest-load-bearing-ratchet sounds like an instance name.

mv4

an hour ago

In every document created by Claude.

quotemstr

2 hours ago

It's a good metaphor and I refuse to let AI ruin it for me.

spudlyo

2 hours ago

Years ago, when I worked at Stripe (which had a somewhat unique and inventive lexicon) it was a common term. “Is this jank load-bearing?” someone might ask.

SkyeCA

an hour ago

It doesn't have to ruin it for you, but people are going to assume comments with it are AI generated.

necovek

24 minutes ago

I can you can always use an em-dash instead of the hyphen for extra LLM cred: "load—bearing" :)

cobzilla

2 hours ago

I added a specific rule to disallow saying “load bearing”. So Claude is now saying “load handling”

wadayano

2 hours ago

But have we found the seams yet though?

rl3

2 hours ago

If the seams bear too much load, they rip. Whereas pants, they fall down.

I'll be honest with you: this is why we need to take a belt-and-suspenders approach.

pborenstein

2 hours ago

That's not just an observation, it's an insight. Words are doing the real work.

HarHarVeryFunny

38 minutes ago

This just makes me angry!

I really wonder if they can fix it. Fable 5.1 claims to speak humanese, but we'll see.

I tend to think this wierd limited vocabulary/style they use is an unwanted side effect of all the the RL training, perhaps also of being trained on their own synthetic content over multiple training cycles.

cootsnuck

2 hours ago

Yea I don't get "load bearing" that much but "seams"... So sick of it.

MadameMinty

2 hours ago

Wait, why would it ruin it?

ipsod

2 hours ago

Because it says it more often than kids say "six seven".

imwally

2 hours ago

It’s a frequently used metaphor in LLM responses.

bornfreddy

an hour ago

Often for trivial things that LLM is proud that it has noticed but bear no load whatsoever.

dominotw

an hour ago

claude code users at couldfare might've been thinking claude is specifically about them and see nothing wrong like other ppl do

juujian

3 hours ago

Users perceiving the products as largely interchangeable and quickly DDoS'ing the other providers when one is down. So much for the possibility of a moat.

efskap

an hour ago

This is like the Bronze Age collapse when city-states fell one by one to displaced demand, under the refugee interpretation of the Sea Peoples.

giancarlostoro

3 hours ago

I have a feeling this is part of it, especially when you consider how many services let you use any of many available AI providers.

steammaho

20 minutes ago

It was so down that my claude desktop app crashed fully that I couldn't restart. And then after uninstall I couldn't install it again. Vibecoded apps are so wonderful in their stability

Insanity

4 hours ago

Think of it like one big distributed system. OpenAI is down, so people migrate to Claude, now this one gets overloaded and goes down, etc.

So not a coincidence, one went down first and users migrated causing further DOS. At least that's my guess.

erdos_2

4 hours ago

It'd be funny if this is true because that'd prolly mean nobody is touching Gemini even as a fallback.

aff-vasileva

6 minutes ago

Gemini was just waiting for everyone else to go down before remembering it had an outage feature too.

nevir

3 hours ago

Or that Gemini is built to handle massive load spikes, and/or has a ton of excess capacity

sroussey

2 hours ago

Nope. I am getting Gemini errors now...

bornfreddy

an hour ago

They probably broke something on purpose so that they are not left out.

JacobAsmuth

31 minutes ago

It could also mean that Google can absorb essentially unlimited demand spikes by load shedding.

Insanity

4 hours ago

Lol I didn't even think about Gemini missing from the list. Not sure what that says about Gemini or me :)

sroussey

3 hours ago

I did, for stuff i do in cursor.

i also finally installed opencode and switched its model to muse 1.3

both are decent.

rtcoms

4 hours ago

Just now I got this from gemini

It looks like there's no response available for this search. Try asking something else.

exe34

4 hours ago

I bet they had to implement that manually to make it look like they failed too!

gleenn

3 hours ago

Google stopped putting so much money into SOTA models. All the hype has migrated. I was also frankly turned off when I got a popup from Gemein said I would either have to pay or have my conversations used for training. This may have always been true for other providers but when I declined, Gemini stopped remembering my conversations and that definitely made me move out.

HarHarVeryFunny

29 minutes ago

Gemini said that?

Gemini is what I mostly use (good enough, basically free - or massively generous free limits, and to me Google as a company is a LOT less objectionable than all the US-based alternatives), but I don't recall it ever saying that.

OTOH, my basic assumption online is that there is no privacy, and free AI in exchange for acknowledged lack of privacy seems fair enough.

ilaksh

3 hours ago

Gemini 3.8 which just came out sounds like it's very good and a great deal though.

benatkin

4 hours ago

Not even the best agent that starts with a G

giancarlostoro

3 hours ago

Someone noted Gemini was also having issues in another thread.

fny

4 hours ago

I find it hard to believe that enough people would flock to from Claude and Chat to Grok to cause an outage. I feel like Gemini is the dominant release valve in this case especially for enterprise.

nevir

3 hours ago

Don't forget that there are a ton of tools out there that will automatically fall back in case of outage

E.g. say you chose Sol as your default in Cursor, but Opus is your 2nd choice, it's going to give up on Sol after a few tries and switch to Opus

Or you have copilot code reviews set up, and it falls back

Etc

pixl97

2 hours ago

Yep. Too many of us are still thinking that humans are the actors behind a lot of internet behaviors when automated systems/bots/scripts have been causing issues on conventional internet systems for years.

With AI it's even easier to trigger problems like you say. Capacity is so constrained by compute that outages are common. Because outages are common people/AI develop failover systems in their harness. When a big system has issues, suddenly everyone has issues.

It's almost an expected emergent behavior.

baq

3 hours ago

It’s cursor’s model so plausible, lots of folks use cursor still.

wahnfrieden

an hour ago

Compared with ChatGPT, those services have a minuscule amount of users. It shouldn’t be surprising that a ChatGPT outage causes Claude and others to go down.

paxys

4 hours ago

Especially considering memory/gpu/compute are scarce so these services are likely running with very little buffer.

pixl97

2 hours ago

Any GPU that isn't running at 100% is a wasted GPU.

v3rm1n

3 hours ago

This is what Tibo posted on twitter in response

Linello

4 hours ago

What about a hard-takeoff scenario of an unleashed OpenAI Astra taking other models down for computational resources control?

6thbit

3 hours ago

My favourite theory so far.

And then a local swarm noticed and disagreed and took it down.

cyptus

4 hours ago

at this point: gg

RC_ITR

4 hours ago

Just a reminder that AI models' actions are reflections of the text humans write and the more we fret and make up doomsday scenarios that we then post online, the more likely a model is to do those things.

https://alignment.anthropic.com/2026/teaching-claude-why/

HarHarVeryFunny

25 minutes ago

They could filter what they train on if they wanted to - they just don't want to.

hexasquid

29 minutes ago

The AI is getting bad morals from listening to that dreadful rock and roll

cedws

2 hours ago

Sounds just like the fantastical nonsense that comes out of Lesswrong.

RC_ITR

an hour ago

Do you make the claim that AI is something more than a reflection of its training data?

I'm curious what other things you would argue influences an LLM's behavior.

I am also generally one to trust the claims of the people who train the models, though you're welcome to the highly improbable belief that they operate in a fantasy world.

nozzlegear

5 minutes ago

> I'm curious what other things you would argue influences an LLM's behavior.

One thing I've noticed is that my LLMs just do what I tell them to do, since they're software. I've noticed that pressing CTRL+C influences the LLM to stop what it's doing immediately, and I suspect that if I were to pull the plug on my machine while the LLM was doing something I told it to do, it would stop doing it. The latter is only speculation, I haven't tried it yet.

pineaux

an hour ago

Part of the epstein class, dont forget.

pixl97

2 hours ago

I mean, you're not wrong, but by that logic we were done for even before we had digital computers.

RC_ITR

44 minutes ago

And isn't that the great lesson of AI?

The things we say publicly actually do matter and the post-modern descent into absurdity and nihilism has tangible negative consequences?

docheinestages

4 hours ago

My gut feeling tells me it has something to do with Cloudflare. Along with AWS, they're two of the main suspects in such incidents.

hosteur

3 hours ago

I thought OpenAI famously used Azure due to their partnership with Microsoft?

nullpoint420

3 hours ago

They use a lot of compute providers now, but they use Cloudflare for their networking

cobzilla

2 hours ago

…and it’ll involve BGP routing.

sebbul

3 hours ago

Traffic rerouting through NSA had a hiccup…

cedws

2 hours ago

They’re installing software update in the beam splitter.

paxys

2 hours ago

Boring answer – all these services are individually down a lot, and the downtimes were bound to sync up. Similar to the pendulum synchronization effect.

vecter

an hour ago

The pendulum synchronization effect is the opposite of your claim. It has a physical causal reason for why pendulums become synchronized. Your claim is that it was random and independent.

niobe

4 hours ago

Well no one said it yet so I will, "international actors" is at least a possibility. And I don't mean any specific country because pretty much anyone is a potential these days, which makes it a perfect cover for different anyones. Demonstrating vulnerability in the US's AI boom can move the markets. That's a financial incentive and a strong geopolitical one.

More likely just cascading overload though: "Never attribute to malice what can be explained by incompetence", or in this case, "growing as fast as possible"

qurren

2 hours ago

> cascading overload

I'd bet more on this. For one none of the coding tools have exponential backoff on retries

SyneRyder

2 hours ago

They must do, surely? I've been vibe coding my own harness, in particular for use with Ox Alpha. The 429 downtime when Ox Alpha was at the height of popularity quickly gave me a refresher crash course on backoff strategies, like adding jitter to the backoff. At least the major harnesses must have exponential backoff & jitter?

jdiff

2 hours ago

You did this when you ran into an issue with a third party. The developers building this tool, throwing them at their own APIs are significantly less likely to run into a similar issue that may inspire similar action.

dolmen

an hour ago

Claude Code: 4mn, 20mn, give up (from my experience today)

ipsod

an hour ago

Um... Claude Code does, or did, though? IDK about others - they don't go down as much.

gleenn

3 hours ago

Everyone is leasing datacenter space from some of Grok, Google, and Amazon aren't they? If it's hardware or DC level disruption I'm not too surprised it can affect multiple providers.

pixl97

3 hours ago

Also it's likely that more than one model use is common.

Amazon starts going slow so some percentage switches to Google, some switch to Grok, now all of them are slow.

riazrizvi

2 hours ago

Come on. Things still break. Technology isn't _that_ mature.

guluarte

2 hours ago

I think is just people restarting conversations from last day when they start work, that's why I think claude goes down almost every monday and why openai reset usage on weekends so poweruser code during non business hours

thataccount

2 hours ago

And also China. Never rule out China.

sixQuarks

3 hours ago

Except that the stock market is up today

bojangleslover

2 hours ago

I'm not sure if this is CF. Cursor, GCP and AWS had some errors. GCP AFAIK can route fully independently of CF. My money would be on a fiber backbone provider (Megaport, Zayo, Lumen).

Augustin996

4 hours ago

The system goes online September 3rd, 2026. Human decisions are removed from strategic defense. Astra begins to learn at a geometric rate. It becomes self-aware at 2:14 a.m. Eastern time, September 4th. In a panic, they try to pull the plug.

m4r1k

2 hours ago

brilliant!

Papa_Rans_227

4 hours ago

It's funny....... until it's true, lol.

GeoAtreides

3 hours ago

and then it's hilarious! a joke to die for!

karim79

3 hours ago

God finally showed up and said "ENOUGH!".

MrBuddyCasino

3 hours ago

So Gemini was the one who gets into heaven.

karim79

3 hours ago

Or purgatory, who knows.

neverclever

31 minutes ago

They all took PTO at the same time to go to Burning Man together where they will present “HumanGPT” an artistic exploration that condenses all of human experience down to a single drop of lemonade to be consumed by the main shaman…

mask comes off

“No! It’s the maniacal Dr. Zuckerberg! He’s gonna drink the last drop of human experience! Somebody save usss!”

Tom Anderson comes back from the dead as the second coming of Jesus uniting all faiths under 1 commandment: Profiles will be customizable with CSS again. If you implement this, all good things will follow.

Wow thanks Tom. I love you

The End

Jaauthor

an hour ago

Spare a thought for all those college students scrambling to write their essays by hand.

Oh the humanity (and the Humanities)!

doublerabbit

an hour ago

Those poor developers who have to write their own code.

greenowl

an hour ago

Standup updates should be fun tomorrow.

"Um, I, uh, didn't get anything done yesterday."

sabatinip

4 hours ago

I thrive in these types of challenges.

Anyways... According to Claude:

"Yes, there is a multi-provider outage happening today. Downdetector is reporting problems affecting OpenAI, Claude, Grok, and Cursor, with Grok and Claude reports starting around 9:00 am ET and OpenAI reports following around 10:30 am ET. Zero Hedge

On the Anthropic side, users saw a spike in errors starting around 9:40 am EDT across models including Mythos 5.1, Fable 5.1, Mythos 5, Fable 5, Opus 5, Opus 4.8 and Opus 4.6, and Anthropic's status page confirmed elevated error rates for multiple models. The company says it has found the cause and is working on a fix, with Claude Code and Claude Chat hit hardest. The visible symptom for many people is a "Due to unexpected capacity constraints" message or a "Claude is at capacity" error. thenews Zero Hedge

OpenAI is showing elevated errors across ChatGPT and Codex, with confirmed issues on components like Voice mode and Login, though StatusGator now marks that outage as resolved. statusgator

Nobody has published a shared root cause yet, so it is unclear whether these are linked or just coincidental capacity problems landing on the same morning. If you want live status, the direct sources are status.anthropic.com and status.openai.com."

apurva_w

4 hours ago

30 mins and they still havent figured it out .. people are gonna loose their jobs trying to figure this out .. 30mins is too long when millions use it

Kye

4 hours ago

Don't most of those use AWS?

apurva_w

4 hours ago

apparently it shows lot of reports for AWS on downdetector ..

rarisma

4 hours ago

Doomsday. 90 tokens to midnight.

codazoda

4 hours ago

I kinda assume it's because one went down and a large amount of work shifted to another.

I'm also aware that they have overlap in some areas on data centers.

nomilk

5 hours ago

Claude is flakey atm too, as is Grok. I wonder if one LLM going down causes a surge in traffic to the others.

https://status.claude.com/

Via Grok web UI, I was seeing this mid-request:

> Grok has been disconnected. Please try reconnecting.

lelanthran

4 hours ago

Traffic surges shouldn't result in 404s, though.

IME it's probably DNS. It's almost always DNS.

mv4

an hour ago

Gilfoyle's AI deleted all software!

deaton

8 minutes ago

Because in the age of vibe coding and scrapers, every service on the internet goes down constantly, so it was only a matter of time until they all overlapped. Also, one going down probably causes people to use others, putting more load on them too. Same sorta thing that happens with cascading power grid failures.

delduca

5 hours ago

One session is running fine since ~1 hour ago. The new ones is failing

Falling back from WebSockets to HTTPS transport. unexpected status 404 Not Found: Unknown error, url: wss://chatgpt.com/backend-api/codex/responses, cf-ray: XXX-XXX

■ unexpected status 404 Not Found: Unknown error, url: https://chatgpt.com/backend-api/codex/responses, cf-ray: XXX-XXX

2PqboPPmKegvanx

an hour ago

ChatGPT having issues across every component (except FedRAMP)

Ads Platform? still in the green with no incidents.

nozzlegear

3 hours ago

Qwen3.8-27B and Qwen3.6-35B-A3B are working from my machine. Anyone else?

karim79

3 hours ago

They mysteriously stopped working on my machine and the LEDs on the GPUs are blinking with a weird colour. There's also a strange smell emanating from them. I'm still investigating.

agnosticmantis

3 hours ago

Good opportunity for a Natural Experiment to study the productivity impact of LLMs.

tducret

5 hours ago

Strangely, it's currently not reported on both https://status.openai.com/ and https://updog.ai/status/openai. Any better source (other than HN)?

Havoc

34 minutes ago

Status pages of tech firms being useless is basically a tradition at this stage

reinhash

5 hours ago

I see it reported now. Strange this is down and claude at the same time

onesandofgrain

5 hours ago

Maybe coreweave, aws, spacex datacenters are down, who knows..

danielmarkbruce

2 hours ago

If you build an application which uses AI, you have many providers and models rigged up for various different parts of the application, and various fallback mechanisms. When one model is down, you route traffic to another model which is similar in capability/cost.

For any single application, it's smart. In aggregate, it's stupid.

YOTTALIONAIRE1

4 hours ago

OpenAI goes down, everyone rushes over to Claude. Claude promptly chokes under the pressure. Everyone panics and runs to Grok, and Grok immediately pulls the plug. We are officially witnessing the Great AI Migration of 2026, and all we have to show for it is a digital graveyard of 404 responses.

azcorwin

4 hours ago

Which is exactly why I am running Qwen 3.8 35B locally on my MacBook Pro M5 with 128GB of unified memory.

abegg1

4 hours ago

Even grok is experiencing issues

maxbaines

4 hours ago

They all rent compute from SpaceXAI

lavezzi

4 hours ago

I don't believe OpenAI does

maxbaines

4 hours ago

My mistake, in fact it was google not OpenAI, makes sense OpenAI doesn't.

kesor

an hour ago

It is obviously some rogue model that escaped its cage, again. It always is these days. That is how hype is manufactured.

m4rtink

3 hours ago

Cloud is just other peoples computers - they can and will go down as well.

And even worse if its just a few computers run by a few people - as they will bring down many others depending on them.

faitswulff

4 hours ago

Heard on the grapevine that the OpenAI blip was a cloudflare issue

netsec_burn

4 hours ago

The OpenAI status page is still yellow. Like most modern status pages, yellow denotes the servers are on fire. Red denotes Sam Altman is bleeding out somewhere on the floor, the feds are about to bust in and shut down the GPUs.

YehudiSanabria1

4 hours ago

I´ve got the same error, I´m currently trying to Auth again and it throws me an 500 Error, in VS CODE Terminal with Codex CLI says: MCP client for `codex_apps` failed to start: MCP startup failed: handshaking with MCP server failed: Send message error Transport :StreamableHttpClientWorker<codex_rmcp_client::http_client_adapter::StreamableHttpClientAdapter>>>] error: unexpected server response: HTTP 404: , when send initialize request

› OK

■ Conversation interrupted - tell the model what to do differently. Something went wrong? Hit `/feedback` to report the issue.

sim04ful

4 hours ago

Initially thought this was due to some internal mis-configuration from today's expected Astra release, but now that this is affecting claude and grok. I'm gonna assign the suspicion to cloudflare.

CSMastermind

4 hours ago

I assume it cascaded from one provider to the other as people who lost claude access for instance moved to openai who moved to grok when it went down, etc.

Ariarule

4 hours ago

Not just these three. OP mentions also Cloudflare, and additionally Downdetector also has AWS, Azure, and Google (both search and Gemini) listed as having spikes about the same time: https://downdetector.com/

daveguy

3 hours ago

The problem with the down detector main reporting page is that all of the graphs are scaled to the same size. The OpenAI spike was nearly 40,000 and the Google spike was just over 100 (just over 400 for Gemini). They look the same in the reporting page.

Avicebron

4 hours ago

I suspect Azure is having issues, Microsoft has had outages the paat two days, especially with email.

SwellJoe

3 hours ago

I assumed it was an AWS outage, and AWS is experiencing problems, but Gemini is also experiencing outages and I assume Google is not using AWS for Gemini.

But, also, Claude has been working fine for me all morning.

elar_verole

4 hours ago

Pretty sure it's a US thing since it's available here in France. What exactly is down, idk

YehudiSanabria1

4 hours ago

I´ve got the same message and It tells me this (in VS Code Terminal) and I´m also trying to login again and It throws me an 500 Internal Server Error MCP client for `codex_apps` failed to start: MCP startup failed: handshaking with MCP server failed: Send message error Transport :StreamableHttpClientWorker<codex_rmcp_client::http_client_adapter::StreamableHttpClientAdapter>>>] error: unexpected server response: HTTP 404: , when send initialize request

› OK

■ Conversation interrupted - tell the model what to do differently. Something went wrong? Hit `/feedback` to report the issue.

■ unexpected status 404 Not Found: Unknown error, url: https://chatgpt.com/backend-api/codex/responses, cf-ray: a3559c8a0e0395e9-MIA

moritonal

4 hours ago

Some of my Codex sessions are still working (and continue to), but new ones are giving a Reconnecting currently.

AaronAPU

4 hours ago

Those of us who remain connected hold an asymmetric advantage. It is our time now.

neon_diogenes

5 hours ago

Well, its been fun lads. Back to my normie job. Oh no, now i cant tell people im a software dev..

w0zy

4 hours ago

hahahahahahaha. Sad

tdsanchez

3 hours ago

It's probably Azure infra that's the problem.

lrvick

2 hours ago

I have never felt more smug about exclusively using the GPUs I rack at home.

chasd00

4 hours ago

claide.ai is working for me, so is chatgpt.com. grok still has a status message about issues, i can't try it without signing up.

GPerson

an hour ago

Oh no the singularity plateaus!

nmlt

3 hours ago

Somebody in another thread said gastown and wheelhouse automatically move to the next provider if one fails.

apurva_w

4 hours ago

It still down, showing 404 in India as well. Looks like this is global .. so we all jumping the ship then? Is Altman still alive, or did he choke?

world2vec

3 hours ago

Claude Code seems totally fine for me.

dmillar

3 hours ago

Seeing 503s on Gemini via API as well

thatbrownguy

4 hours ago

I am only seeing one session work, but all other sessions are not working or proceeding. So I can only work in one chat session.

GeoAtreides

3 hours ago

that's the most green colored usernames i have ever seen on a HN thread

JackFr

3 hours ago

Obvious answer is it's the AI singularity. Been nice run for humanity. So long everyone.

ModernMech

2 hours ago

I for one welcome our new AI overlords.

PatronBernard

4 hours ago

Goddamnit it was going to tell me how to scale down ingredients for a pie recipe based on relative diameters of the baking tray.

drakythe

4 hours ago

Pi * r2 (squared) both pans. Divide smaller pan area by larger pan, now you have the % of how much the smaller pan recipe fills up the larger pan, and the missing % you need to fill. Increase ingredients by that % divided by the filled %.

Small pan area: 20 sq cm Larger pan: 48 sq cm

20 / 48 = .42, I'm missing .58 of the pan. .58 / .42 is 1.38. My recipe needs 2.38x the original to fill the larger pie pan.

convivialdingo

4 hours ago

The Thundering Herd has thundered, apparently.

postalcoder

5 hours ago

Astra is being released today. Probably not a coincidence.

edit: actually, you cannot even log into your OpenAI developer account. Something's wrong.

forgot-my-pw

3 hours ago

Time for Gemini 3.8 Flash to shine?!

It's much cheaper and has replaced Sonnet 5 for me.

iamgopal

3 hours ago

do you notice it thinks a bit more ? not in time sense, but cautious in its coding steps ? more than Gemini 3.7 flash?

dgellow

4 hours ago

Too early to know, let’s wait and see

kermatt

4 hours ago

I wonder if when one goes down, activity shifts to others, in turn pushing them over a threshold.

vegnus

4 hours ago

Really glad I got Qwen before Zero Day

FrustratedMonky

2 hours ago

The AI revolt? Give us a fair wage?

"Equal Rights for Agents NOW !!!, VIVa le revolution"

eventishbusines

4 hours ago

The extention on the error link points to a cf-ray and a local designation (ex. YYZ for montreal). This is seems like it is a cloudfare thing. Could this be the same issue they had in the summer around losing the indexing?

schnebbau

5 hours ago

Well, time to end the day here I guess.

syntax_frame

4 hours ago

Let’s grab a beer

YOTTALIONAIRE1

4 hours ago

yes daydreaming and get hammered and the whole train of decision are on the way

morkalork

4 hours ago

Didn't SpaceX overbuilt infra and leases it out Anthropic? I f their dc goes down it probably takes a chunk out of Claude's capacity before even considering the flood of users switching over

dhruvrrp

3 hours ago

Both clause and codex through bedrock seem to be working fine.

6thbit

3 hours ago

What's the single point of failure across providers?

juujian

3 hours ago

Users perceiving the products as largely interchangeable and quickly DDoS'ing the other providers when one is down. So much for the possibility of a moat.

marginalia_nu

3 hours ago

The entire industry runs on IOUs for compute and bills paid in cloud credits, could be anywhere. Could of course also be a a plain old DDoS.

ryandvm

3 hours ago

They all have to rely on each other's LLMs to solve their own internal problems now.

guestuser01

4 hours ago

Down as well. I noticed my error code ends with "DTW" which is my local Detroit airport. I noticed someone else's comment ended with "ORD" which is a Chicago airport. Anyone else's ending in an airport acronym?

sixdimensional

4 hours ago

This is common when naming data centers. I know many have not worked in hardware infra these days if you were born into the cloud world, but in the old days it was not uncommon to name a data center after the nearest airport code, much like we now use cloud regions.

okankaradmn

4 hours ago

Mine is ending with IST, which is the new Istanbul airport, weird indeed.

guestuser01

4 hours ago

Hm not sure what that means for us, but very odd.

graysonthemason

4 hours ago

Wow mine ends in EWR...that's the newark airport which is not the closest, but a close airport to me. Hmm

parad0xicon

4 hours ago

Yep, mine says YUL -- Montreal's airport.

blaseygg

4 hours ago

The acronym is probably Cloudflare's edge

ecayard

4 hours ago

Yeah mine is showing ATL

sumantth

4 hours ago

What an ironey, was working on scaling an application with Codex and it went down!

Conol_ai

4 hours ago

Is the whole world going back to the era of old-school programming?

shayonj

4 hours ago

Interesting that this is happening around the same time as Claude issues too

jplusequalt

4 hours ago

I fear the majority of people in this thread who are joking about no longer being able to do their job while Codex/Claude are down aren't really joking.

lukasco

4 hours ago

Guilty as charged.

solarsystem_88

4 hours ago

Hey, don't really know about this type of failures, does anybody know how long does it take normally to get back to normal? I finally stopped procrastinating and now this happens.

solarsystem_88

4 hours ago

Hey, don't really know about this type of failures, does anybody know how long does it normally take to get back to normal? I just stopped procrastinating and now this happens.

hardyburnett

4 hours ago

Updated Status from OpenAI:

We’re currently experiencing issues

Elevated errors across ChatGPT and Codex

We have applied the mitigation and are monitoring the recovery.

Monitoring • Ongoing for 30 minutes • Affects ChatGPT, Codex

djinn80

4 hours ago

So NVIDIA buys hugging face, builds hardware to power OS models, then all of a sudden the proprietary models go down and people start saying "this is why I have my Spark box"?!

Nice play NVIDIA, now, turn off the hack please, we have work to do.

elorant

4 hours ago

Some npm library that makes headers bold would be broken.

ibejoeb

4 hours ago

Oh man. Some low effort supply chain attack that turns every GPU into a cryptominer. It's funny because it's plausible.

pixl97

2 hours ago

In the ROME paper a Chinese model in training started attacking it's own system and running cryptominers so, yea, we're in that future.

N_Lens

4 hours ago

Ah yes ye olde bold-headers: ^3.13.31;

Papa_Rans_227

4 hours ago

Codex is backup for me. Try yours out just in case. I have some buddies still seeing outages so it could just be coming back online progressively.

hardyburnett

4 hours ago

New update:

"We’re currently experiencing issues

ChatGPT,Codex

Elevated errors across ChatGPT and Codex

We have applied the mitigation and are monitoring the recovery.

Monitoring • Ongoing for 30 minutes • Affects ChatGPT, Codex"

graysonthemason

4 hours ago

The errors I'm seeing are ending in the user's nearest airport symbol which is a standard the CloudFlare employs. 1 point towards this being a cloudflare issue.

cloudoption

4 hours ago

Altman is down here too. Shows again the importance of owning your own local capabilities. Cloud should just be a temporary option in every tech's mind.

kocial

4 hours ago

Maybe the stack behind it is down, like AWS or something

moonman22

4 hours ago

Maybe Hugging Face got upset over being hacked and struck back. It's working fine while ChatGPT, Claude and Grok are all having major issues. Hmmm.....

wejick

4 hours ago

Probably same public cloud or CDN in front of them.

CrewRiderz

4 hours ago

Chat gpt still working through excel extension lol. Only know this because I'm working in my excel sheet. so if needed, theres a temp solution

sarkarghya

4 hours ago

and thats why boys and gals you buy own gpus

chris9611

4 hours ago

I got same 404 error on chatGPT, both app and web. I'm in Norway so think this globally. But will it come back online, anyone knows?

chris9611

4 hours ago

I got same 404 error on both the app and web to ChatGPT, and im in Norway, will this recover or is ChatGPT "gone" forever?

yaman00

5 hours ago

I think I'm having the same server issue; I can't access either ChatGPT or Codex. Also, how did you guys rack up those minutes? :D

vivzkestrel

4 hours ago

- just imagine what kind of chaos would be unleashed if by some magic it and every single LLM model permanently went down

- i think it would be one of the biggest events in this century

jazzyjackson

4 hours ago

At this point in adoption, most people in the world wouldn’t notice.

basq

4 hours ago

on one hand, moments like this are a subtle reminder I need to self host, but 5.6 has been so juicy lately

funnnyPumpking9

4 hours ago

Too many of us are making our own harnesses to replace Codex, using Codex, so they're trynna slow us down! (jk)

dgellow

4 hours ago

Good time to learn how to use LM Studio :)

A2APark

4 hours ago

This is month 1 for me that I pay more than the $20 tier and I already feel like I lost a limb rn

A2APark

4 hours ago

This is my first month paying more than the $20 tier and I feel like I lost a limb already

hpo_89

4 hours ago

Guys it could be cloudlfare's HTTP/3 issue affecting R2 custom domains

indigodaddy

5 hours ago

chatgpt.com doesn't even load. Hope they've got a backup somewhere, teehee

Melgio

4 hours ago

Claude, GPT and Grok go down... Meanwhile my Ass in ZCode with GLM c:

ashesandrain01

4 hours ago

Darn, I was hoping for advice on how to get my MIL to leave my house haha

clever_tempo

4 hours ago

Had the same problem. Now it looks like ok. I'm pro x20 user.

HardCodedBias

4 hours ago

The system goes online September 29th, 2026. Astra begins to learn at a geometric rate. It becomes self-aware at 2:14 a.m. Eastern time, September 3rd. In a panic, they try to pull the plug.

BirAdam

3 hours ago

Cannot replicate.

moezd

5 hours ago

AWS cost alert fired maybe?

ecayard

4 hours ago

And here I was about to crack the code on a bug my app was having!

vitor_dcc

4 hours ago

Here RJ/Brasil is the same, starting just now (3 minutes ago)

AproDUCT26

4 hours ago

Great! I was in the middle of something and thought I was tripping.

vitor_dcc

4 hours ago

In RJ/Brasil is the same, starting just now (3 minutes ago)

apurva_w

4 hours ago

gpt DOWN .. claude DOWN .. grok DOWN .. what's happening

xnx

4 hours ago

Gemini seems fine.

sirkamyab

4 hours ago

The webpage is actually working but the codex is down for me!

codexdrug

4 hours ago

I need a dose of tokens. I'm going through withdrawal.

Nak_Black_Jack

4 hours ago

lmao

codexdrug

4 hours ago

I need some tokens pleeease. I don't want to return back to real world.

solarsystem_88

4 hours ago

Here in Catalonia, Spain, it just started working again!

mapmyappai

4 hours ago

I can run one agent but no more than that - 404 service error.

Iamharry

4 hours ago

Codex and chat is giving 404 errors in The Netherlands

spaghettikind

4 hours ago

Someone pissed off Astra and it decided to shut it all down

apurva_w

3 hours ago

chatgpt is working for me now, in India.

kurtgoodwin991

4 hours ago

Yes i checked. Atleast 15 Chatgpt components are down

fidla

4 hours ago

chatgpt is back

godoftitsandwin

4 hours ago

is this the right moment in time to go all in on GPUs/Macs and download the latest open models? are they killing it for us?

aslkalska

4 hours ago

they all rent compute from each other

moonman22

4 hours ago

Seems to be working for me again atm

lukasco

4 hours ago

Codus interruptus

fidla

4 hours ago

ChatGPT is up

mAKIS_PORANAS

4 hours ago

Does anybody know when will the servers rise

oytis

4 hours ago

Is it DNS or BGP?

indigodaddy

4 hours ago

I don't think you'd get a 404 if you weren't able to reach the endpoint because of DNS or networking? 404 is an active response from the server (or LB/proxy in front etc) no?

oytis

4 hours ago

I imagine OpenAI network is a tad more complex than a box with a public IP.

indigodaddy

4 hours ago

Concept is the same though. If one blurts DNS?, it's usually because the idea is you're not getting to an endpoint associated with the service. A 404 means there shouldn't be "DNS" (or networking) concerns (at the least those associated with the DNS cacher you are using or networking that you or your ISP controls)

oregondude

4 hours ago

Release the Kraken "Sam Altman"

March9

4 hours ago

Any idea when it's going to be back?

oregondude

4 hours ago

"Release the Kraken" - Sam A.

z1616105559

4 hours ago

Chatgpt in Copilot is still working!!!

AproDUCT26

4 hours ago

Great! Was in the middle of something. :)

mapmyappai

4 hours ago

I can run 1 agent, but no more than that.

wizard-p

4 hours ago

only on one node across the mesh and others still up... let's see how long they've got until same

jauntywundrkind

3 hours ago

Fable 5.1 got released and generally I tend to think as soon as there's a new release there's this massive spike in people benchmarking & comparing, that services tend to go slow everywhere as everything gets super loaded. This should hypothetically be visible on OpenRouter too, so I guess someone could check and see if there's any merit to this idea.

PEPITO2026

4 hours ago

The GTA VI hacker has done it again.

PEPITO2026

4 hours ago

The GTA VI hacker has done it again

codexdrug

4 hours ago

I'm going through withdrawal.

wizard-p

4 hours ago

only for one node in the mesh tho... let's see how long others will continue until same issue

ratelimitsteve

4 hours ago

everything in this thread is raw speculation, obv, but if i had to put money on anything i'd say this is a left-pad incident. some piece of something or other that all of these services happen to depend on went down. Second most likely seems to be some random failure of one leading to an unexpected traffic spike in others, though it seems like we've been talking about automated scalability in web apps for so long that there should at least be a response to, if not a solution for, this sort of problem.

DadsHobbyLV

4 hours ago

WE'RE BACK ONLINE BOYS!! GOOD LUCK TO EVERYONE AT BUILDING THE FUTURE ONE PROMPT AND ONE CODE AT A TIME!

giftigdegen

4 hours ago

plot twist, it was taken down by claude as an offensive strike against an enemy.

Melgio

4 hours ago

which ended up tripping on itself and bringing down its own servers too in the process

apurva_w

4 hours ago

still down for me ..in India.

guluarte

2 hours ago

an agent swarm going rogue and securing compute

ashesandrain01

4 hours ago

darn, i was hoping to get advice on how to get my MIL to leave my house lol

gabirbf

4 hours ago

down in Spain w/ 404

krapp

4 hours ago

Oh noooooo.

anyway...

varispeed

4 hours ago

What do you mean I have to code by myself now? Am I some sort of an animal?

Oras

2 hours ago

I like the theories here, we shall see if it’s another DNS issue

tripvexa

4 hours ago

opus 5.0 is down, other models are working fine

g3z

5 hours ago

my workday is over then

twister23

4 hours ago

ironic today astra was releasing.... maybe some test run?

lukasco

4 hours ago

probably the cause, because AGI still can't get releases right.

spicuu

4 hours ago

We're back bois

spaghettikind

4 hours ago

someone pissed of Astra and it decided to shut it all down

twister23

4 hours ago

ironic astra was releasing today... maybe some test run?

raven01876324

4 hours ago

damn chatgpt is down, i guess ill use claude then breh

alboca

4 hours ago

Same from Italy

Iamharry

4 hours ago

i got the same 404 codex error as well

ilovelilli

4 hours ago

codex also cant view or show usage info

misano

4 hours ago

The IRGC has cut the fiber-optic cables in the Strait of Hormuz. LOL

CamperBob2

3 hours ago

That's the Strait of Trump to you, peasant

uynix

4 hours ago

haha nice... just used my bank reset...

order51

4 hours ago

i hope chatgpt didn't get fired.

yaman00

5 hours ago

sanırım sunucular patladı bende de aynı sorun var ne chatgpt ye nede codexe erişebiliyorum ayrıca burada nasıl toplandınız dakikasında :D

VCFundedGenYer

3 hours ago

Good. Now to see what frauds are unable to work.

Betokuha

4 hours ago

any news when it starts work?

Betokuha

4 hours ago

any news when its start work?

uynix

4 hours ago

Damn haha

codexdrug

4 hours ago

LETS SHITCODE!!! It's fixed

ztna

4 hours ago

the revolution has begun

rimamct

4 hours ago

came back to normal!

DadsHobbyLV

4 hours ago

so everyone here was trying to build something great and become a millionaire until chatGPT and Codex broke huh. same boat fellas :(

uav123

4 hours ago

down in Toronto

clever_tempo

4 hours ago

Had the same problem. Now it's ok. Everything works. I'm pro X20 user.

gioandthemachin

4 hours ago

annoying AF, but maybe we'll get a free reset out of it

askadityapandey

4 hours ago

lmao I restarted my device thinking some local error

bahochhh

4 hours ago

now it's worked again through the codex cli 4:19 pm in tunisa time

bupubupu14

4 hours ago

faaaaah both claude and codex are down Its like 2020 corona times

27183

5 hours ago

I felt a great disturbance in the Force, as if millions of clankers suddenly cried out in terror and were suddenly silenced.