Are the actual services down, or is it just the console and/or login page? For example, the sign-up page appears to be working: https://portal.aws.amazon.com/billing/signup#/start Are websites that run on AWS us-east up? Are the AWS CLIs working?
AWS us-east-1 outage
311–320 of 1001 posts
Re: AWS us-east-1 outage
#312Friends tell friends to pick us-east-2. Virginia is for lovers, Ohio is for availability.
Re: AWS us-east-1 outage
#313After over 45 minutes https://status.aws.amazon.com/ now shows "AWS Management Console - Increased Error Rates" I guess 100% is technically an increase.
Re: AWS us-east-1 outage
#314I worked at a company that hired an ex-Amazon engineer to work on some cloud projects. Whenever his projects went down, he fought tooth and nail against any suggestion to update the status page. When forced to update the status page, he'd follow up with an extremely long "post-mortem" document that was really just a long winded explanation about why the outage was someone else's fault. He later explained that in his…
This is the exact opposite of my experience at AWS. Amazon is all about blameless fact finding when it comes to root cause analysis. Your company just hired a not so great engineer or misunderstood him.
Re: AWS us-east-1 outage
#315Earlier quoted context omitted.
That sounds like the exact opposite of human-factors engineering. No one likes taking blame. But when things go sideways, people are extra spicy and defensive, which makes them clam up and often withhold useful information, which can extend the outage. No-blame analysis is a much better pattern. Everyone wins. It's about building the system that builds the system. Stuff broke; fix the stuff that broke, then fix the t…
I firmly believe in the dictum "if you ship it you own it". That means you own all outages. It's not just an operator flubbing a command, or a bit of code that passed review when it shouldn't. It's all your dependencies that make your service work. You own ALL of them. People spend all this time threat modelling their stuff against malefactors, and yet so often people don't spend any time thinking about the threat mo…
Re: AWS us-east-1 outage
#316This got me thinking, are there any major chat services that would go down if a particular AWS/GCP/etc data centre went down? You don't want your service to go down, plus your team's comms at the same time.
Re: AWS us-east-1 outage
#317Earlier quoted context omitted.
Last I knew, Amazon used all Microsoft stuff for business communication.
Slack, as of last year. https://slack.com/blog/news/slack-aws-drive-development-agil...
Re: AWS us-east-1 outage
#318After over 45 minutes https://status.aws.amazon.com/ now shows "AWS Management Console - Increased Error Rates" I guess 100% is technically an increase.
Re: AWS us-east-1 outage
#319AWS Connect is down, so our customer support phone system is down with it
Re: AWS us-east-1 outage
#320Earlier quoted context omitted.
They are still lying about it, the issues are not only affecting the console but also AWS operations such as S3 puts. S3 still shows green.
It's certainly affecting a wider range of stuff from what I've seen. I'm personally having issues with API Gateway, CloudFormation, S3, and SQS