Earlier quoted context omitted.
Trying high for that 2-nines reliability. You just can't get that level of reliability if you do it yourself, no matter how hard you try.
We won't do it ourselves, but we also won't do it with a provider that has accumulated 50+ hours of downtime in less than a month all the while having no communication or support. That's barely clearing the one nine availability for the last 30 days (93%) for our particular stack on CF, this is insane. Mind you last time we were hit by a 22h outage on Oct. 9 we didn't get so much as an email from CF either during or…
Cloudflare outage – 24 hours now
51–60 of 76 posts
Re: Cloudflare outage – 24 hours now
#52Earlier quoted context omitted.
In this case it is not. A power outage in a critical data-center is the root cause here: https://www.cloudflarestatus.com/incidents/hm7491k53ppg
What I'm confused by is we had "power partially restored" 22 hours ago, and no news from PDX02 since. I assume both Clouflare and Flexential are on DEFCON 1 right now, but I'm wondering if it might be more than just the building going dark. There's something about a failover than was attempted and crashed halfway through, but unclear if that's what's causing the 24h+ situation.
Re: Cloudflare outage – 24 hours now
#53Re: Cloudflare outage – 24 hours now
#54I'm really looking forward to the post-mortem to this.
What should our expectations be? The best assumption could be that this is the new normal.
Re: Cloudflare outage – 24 hours now
#55Wonder if this is related to the many product launches recently? Even though my general impression is that they test rigorously with long-running alpha and beta test phases.
It is the consequences of a power outage in Flexential PDX02 data center.
Re: Cloudflare outage – 24 hours now
#56Hmm... Who just changed their dns vs. riding it out?
The only feature I need to research in new providers is: access to Whois ASN numbers, which I insert into HTTP request headers. I use this to tailor my site for .gov and .edu users.
Re: Cloudflare outage – 24 hours now
#57Earlier quoted context omitted.
In this case it is not. A power outage in a critical data-center is the root cause here: https://www.cloudflarestatus.com/incidents/hm7491k53ppg
If you can't cope with the loss of a data centre you're not really running a resilient system.
Re: Cloudflare outage – 24 hours now
#58Earlier quoted context omitted.
Trying high for that 2-nines reliability. You just can't get that level of reliability if you do it yourself, no matter how hard you try.
We won't do it ourselves, but we also won't do it with a provider that has accumulated 50+ hours of downtime in less than a month all the while having no communication or support. That's barely clearing the one nine availability for the last 30 days (93%) for our particular stack on CF, this is insane. Mind you last time we were hit by a 22h outage on Oct. 9 we didn't get so much as an email from CF either during or…
Re: Cloudflare outage – 24 hours now
#59BTW Cloudflare tunnels are not working (for the at least the last 16 hours), but it says "Operational" and "restored" on the ticket. Since Shopify's CLI uses Cloudflare tunnels by default to load local resources, Shopify partners are affected by this outage by unable to develop apps, unless they use another tunnel: [0] https://github.com/Shopify/cli/issues/3065 [1] https://github.com/Shopify/cli/issues/3060
Re: Cloudflare outage – 24 hours now
#60Looking forward to a more decentralised global Internet, with packets being routed through alternative paths, so outages like these become a non-event. I understand we do not have the technology for that just yet, and DevOps able to configure TLS terminators on their own are worth their weight in gold. Hard to imagine how the Internet could ever exist without Cloudflare.
It's not just packet routing though, many of their other products seem to be affected as well.