Live data from Hacker News

Google Cloud Is Down

news.ycombinator.com

461–470 of 630 posts

Re: Google Cloud Is Down

#461
post #428

Earlier quoted context omitted.

Back when S3 failures would take town Reddit, parts of Twitter .. Netflix survived because they had additional availability zones. I can remember the bigger names started moving more stuff to their own data centers. AWS tries to lock people in to specific services now which makes it really difficult to migrate. It also takes a while before you get to the tipping point where hosting your own is more financially viable…

I think you’re misremembering about Twitter, which still doesn’t use AWS except for data analytics and cold storage last I heard (2 months ago).

Avatars were hosted on S3 for a long time, IIRC.

Re: Google Cloud Is Down

#462
post #27

Anyone using both AWS and GCP that can form an opinion on availability of both? As a GCP customer I am not very happy with theirs.

I've heard from people who have worked with both AWS and GCP that AWS has far better availability.

I've also heard similar from a teammate who previously worked with GCP. That said I know several folks who work for GCP and they are expending significant resources to improve the product and add features.

Re: Google Cloud Is Down

#463
post #20

GCP status page is worthless as it's always happy and green when production systems are down and then they might acknowledge something an hour later

Was noticing massive issues earlier and thought that maybe my account was blocked due to breaching from TOS as I was heavily playing with Cloud Run. Then I noticed gitlab was also acting up but my Chinese internet was still surprisingly responsive. Tried the status page which said everything was fine and searched Twitter for "google cloud" and also found nobody talking about it. Typically Twitter is the single source of truth for service outages as people start talking about it

Re: Google Cloud Is Down

#464

Earlier quoted context omitted.

It's not so much AWS vs. in-house. But AWS (or GCP/DO/etc.) vs. multi/hybrid solutions. The latter of which would presumably have lower downtime.

Why would you think that self-managed has lower downtime than AWS using multiple datacenters/regions?

Actually, I imagine that if you could go multi-regional then your self-managed solution may be directly competitive in terms of uptime. The idea that in-house can't be multi-regional is a bit old fashioned in 2019.

Re: Google Cloud Is Down

#465
post #425

Earlier quoted context omitted.

The discount seems way too small. I would pay a premium for a cloud provider happy to give 100 percent discount for the month for 10 minutes downtime, and 100 percent discount for the year for an hour's downtime.

Any cloud provider offering those terms would go out of business VERY quickly. Outages happen, all providers are incentivized to minimize the frequency and severity of disruptions - not just from the financial hit of breaching SLA (which for something like this will be significant), but for the reputational damage which can be even more impactful.

How often does amazon or google go down for ten minutes?

But let's work backwards from the goal instead.

If you charge twice as much, and then 20-30% of months are refunded by the SLA, you make more money and you have a much stronger motivation to spend some of that cash on luxurious safety margins and double-extra redundancy.

So what thresholds would get us to that level of refunding?

Re: Google Cloud Is Down

#466
post #337

Earlier quoted context omitted.

Google is well known for not caring about small shops, only if you are a multi million dollar customer with dedicated account manager you can expect reasonable support. That's been the case forever with them.

Does Amazon treat smaller customers any better? I am genuinely asking, as I have no clue.

Definitely. A previous small company I worked in had some S3 Snafu and AWS Support was super helpful.

Re: Google Cloud Is Down

#467

Earlier quoted context omitted.

us-west-1 is Northern California (Bay area). us-west-2 is Oregon (Boardman).

Incorrect. GCE us-west1 is the Dalles, Oregon and us-west2 is Los Angeles.

What I said is correct for AWS. In retrospect I guess the context was a bit ambiguous.

(I will note that I was technically more right in the most obnoxiously pedantic sense since the hyphenation style you used is unique to AWS - `us-west-1` is AWS-style while `us-west1` is GCE-style :P)

Re: Google Cloud Is Down

#469
post #91

Earlier quoted context omitted.

That's one region, not the multiple region that OP mentioned

Services in other regions depended transitively on us-east-1, so it was a multiple region outage.

Which services in other regions? I remember that day well, but I had my eyes on us-east-1 so I don't remember what else (other than status reporting) was affected elsewhere.

Re: Google Cloud Is Down

#470
post #82
post #63

Disclosure: I work on Google Cloud (but disclaimer, I'm on vacation and so not much use to you!). We're having what appears to be a serious networking outage. It's disrupting everything, including unfortunately the tooling we usually use to communicate across the company about outages. There are backup plans, of course, but I wanted to at least come here to say: you're not crazy, nothing is lost (to those concerns do…

>nothing is lost except time

and reputation.
Post reply on HN