AWS multiple services outage in us-east-1
481–490 of 1001 posts
Re: AWS multiple services outage in us-east-1
#482I think no matter how hard you try to avoid it, in the end there's always a massive dependency chain for modern digital infrastructure[^2].
[1]: https://itsfoss.community/uploads/default/optimized/2X/a/ad3...
Re: AWS multiple services outage in us-east-1
#483Every week or so we interview a company and ask them if they have a fall-back plan in case AWS goes down or their cloud account disappears. They always have this deer-in-the-headlights look. 'That can't happen, right?' Now imagine for a bit that it will never come back up. See where that leads you. The internet got its main strengths from the fact that it was completely decentralized. We've been systematically erodin…
Re: AWS multiple services outage in us-east-1
#484Signal was also down.
Re: AWS multiple services outage in us-east-1
#485Re: AWS multiple services outage in us-east-1
#486Maybe this is the event to get everyone off of piling everything onto us-east-1 and hoping for the best, but the last few outages didn’t, so I don’t expect this one to, either.
Re: AWS multiple services outage in us-east-1
#487AWS doesn’t talk about that much publicly, but if you press them they will admit in private that there are some pretty nasty single points of failure in the design of AWS that can materialize if us-east-1 has an issue. Most people would say that means AWS isn’t truly multi-region in some areas.
Not entirely clear yet if those single points of failure were at play here, but risk mitigation isn’t as simple as just “don’t use us-east-1” or “deploy in multiple regions with load balancing failover.”
Re: AWS multiple services outage in us-east-1
#488Now, I may well be naive - but isn't the point of these systems that you fail over gracefully to another data centre and no-one notices?
Re: AWS multiple services outage in us-east-1
#489Every week or so we interview a company and ask them if they have a fall-back plan in case AWS goes down or their cloud account disappears. They always have this deer-in-the-headlights look. 'That can't happen, right?' Now imagine for a bit that it will never come back up. See where that leads you. The internet got its main strengths from the fact that it was completely decentralized. We've been systematically erodin…
Planning for an AWS outage is a complete waste of time and energy for most companies. Yes it does happen but very rarely to the tune of a few hours every 5-10 years. I can almost guarantee that whatever plans you have won’t get you fully operational faster than just waiting for AWS to fix it.
You know that’s not true, is-east-1 last one was 2 years ago? But other services have bad days and foundational one drag others a long
Re: AWS multiple services outage in us-east-1
#490Perhaps for the internet as a whole, but for each individual service it underscores the risk of not hosting your service in multiple zones or having a backup