Inside the longest Atlassian outage
11–20 of 772 posts
Re: Inside the longest Atlassian outage
#12Re: Inside the longest Atlassian outage
#13Re: Inside the longest Atlassian outage
#14The fact it's been so long and they still haven't revealed and explained the root cause of the outage is going to make it hard to regain trust on their buggy, slow tools. The bright side of the incident is that competitors that somewhat care about users have a unique opportunity to stand out.
They did last night: https://www.atlassian.com/engineering/april-2022-outage-upda...
Re: Inside the longest Atlassian outage
#15Re: Inside the longest Atlassian outage
#16I mean, I thought they were text.
5 days to restore text?
They must be generated by a huge complex deep learning voodoo.
Atlassian is working on the bleeding edge of technology. This outage is understandable...
Re: Inside the longest Atlassian outage
#17Selectively restoring data only for certain rows is super hard. But the communications by Atlassian has been the worst I have ever seen in the industry.
Re: Inside the longest Atlassian outage
#18> However, if they [restore backups], while the impacted ~400 companies would get back all their data, everyone else would lose all data committed since that point OK, so you restore backups to a separate system, and selectively copy the stomped accounts data back to production. Simple concepts aren't that simple at their scale, sure, but I suspect this is skimping details on some truly horrendous monolithic architec…
Re: Inside the longest Atlassian outage
#19The fact it's been so long and they still haven't revealed and explained the root cause of the outage is going to make it hard to regain trust on their buggy, slow tools. The bright side of the incident is that competitors that somewhat care about users have a unique opportunity to stand out.
> The fact it's been so long and they still haven't revealed and explained the root cause of the outage They did last night: https://www.atlassian.com/engineering/april-2022-outage-upda...
Ouch. I hope no one person got the blame. This is a systemic failure. Regardless, my regards to the engineers involved.
Re: Inside the longest Atlassian outage
#20> However, if they [restore backups], while the impacted ~400 companies would get back all their data, everyone else would lose all data committed since that point OK, so you restore backups to a separate system, and selectively copy the stomped accounts data back to production. Simple concepts aren't that simple at their scale, sure, but I suspect this is skimping details on some truly horrendous monolithic architec…