Live data from Hacker News

Ask HN: Why were OpenAI, Claude, and Grok simultaneously down?

news.ycombinator.com

711–720 of 780 posts

Re: Ask HN: Why were OpenAI, Claude, and Grok simultaneously down?

#712

Neither of these companies have stellar uptime records. Their downtime episodes overlapped in this instance. In this case, it was a partial downtime for both. Also, OpenAI is saying what caused it: > "A routing error starting around 7:43 am PT on Thursday, September 3, made ChatGPT and Codex unavailable for some users across platforms" Anthropic stated their issue started earlier: > "The company began alerting about…

Do you know the probability of all these companies being down at precisely the same time?

"precisely the same time" meaning 80 minutes apart?

Re: Ask HN: Why were OpenAI, Claude, and Grok simultaneously down?

#713

It’s probably the thing that everyone thinks it is. OpenAI, Anthropic and SpaceXAI are all routed through something that we’re not supposed to know exists and that thing had a whoopsie.

Why would that be probable? People on average are fantastically bad at getting probabilities right.

Historical precedent repeated over and over, and the US ties to DoW work, and the national security implications. It's actually the Occam's Razor explanation if you know the history.

Re: Ask HN: Why were OpenAI, Claude, and Grok simultaneously down?

#714

Earlier quoted context omitted.

Could you expand on this more? It's not clear to me what this is implying.

https://en.wikipedia.org/wiki/33_Thomas_Street https://theintercept.com/2016/11/16/the-nsas-spy-hub-in-new-...

This is the building from Control and the fact that 1. it's a real building and 2. it's an actual spy building is even crazier to me:

https://control.fandom.com/wiki/Oldest_House

https://en.wikipedia.org/wiki/Control_(video_game)

Re: Ask HN: Why were OpenAI, Claude, and Grok simultaneously down?

#715

Neither of these companies have stellar uptime records. Their downtime episodes overlapped in this instance. In this case, it was a partial downtime for both. Also, OpenAI is saying what caused it: > "A routing error starting around 7:43 am PT on Thursday, September 3, made ChatGPT and Codex unavailable for some users across platforms" Anthropic stated their issue started earlier: > "The company began alerting about…

Do you know the probability of all these companies being down at precisely the same time?

Yesterday it was 100%

Re: Ask HN: Why were OpenAI, Claude, and Grok simultaneously down?

#716

I work at OpenAI and I was the Incident Commander for yesterday's outage. We had a routing error within our infra that caused issues for some of our products. It was not related to the Astra launch. We don't comment on other providers' outages.

Did you negotiate extra hard for the 'Incident Commander' title? I'm a bit jealous to be honest.

Re: Ask HN: Why were OpenAI, Claude, and Grok simultaneously down?

#717

Earlier quoted context omitted.

https://en.wikipedia.org/wiki/Room_641A

Does the NSA even have the ability to monitor all AI traffic like this? Wouldn’t that require tons of data centers that there literally hasn’t been time to build yet? I really have no idea, maybe the asymmetry of the compute required to monitor is way lower than the compute required to serve inference?

https://en.wikipedia.org/wiki/Utah_Data_Center

Re: Ask HN: Why were OpenAI, Claude, and Grok simultaneously down?

#718

I work at OpenAI and I was the Incident Commander for yesterday's outage. We had a routing error within our infra that caused issues for some of our products. It was not related to the Astra launch. We don't comment on other providers' outages.

[dead]

Re: Ask HN: Why were OpenAI, Claude, and Grok simultaneously down?

#719

I work at OpenAI and I was the Incident Commander for yesterday's outage. We had a routing error within our infra that caused issues for some of our products. It was not related to the Astra launch. We don't comment on other providers' outages.

Did you negotiate extra hard for the 'Incident Commander' title? I'm a bit jealous to be honest.

It's actually a standard term for the person who plays this role during incident response https://www.pagerduty.com/resources/incident-management-resp...

Re: Ask HN: Why were OpenAI, Claude, and Grok simultaneously down?

#720

I work at OpenAI and I was the Incident Commander for yesterday's outage. We had a routing error within our infra that caused issues for some of our products. It was not related to the Astra launch. We don't comment on other providers' outages.

Did you negotiate extra hard for the 'Incident Commander' title? I'm a bit jealous to be honest.

[dead]
Post reply on HN