Earlier quoted context omitted.
> Monolithic architecture. This particular problem had nothing to do with a monolithic architecture. Your app can be a monolith, but that still doesn't mean your BI team can't have a separate data warehouse or at least separate read replicas to run queries against.
If your website crashes because a single person ran a query, your system is too monolithic. You can have thousands of little microservices running all over the place, but a single query causing a fault proves that a vital system is running without redundancy or load sharing and that other systems cannot handle the situation. You have too many aspects of your service all tied together within a single system. It is too…
Twitter was down
491–500 of 544 posts
Re: Twitter was down
#492Earlier quoted context omitted.
Pretending that junior engineers is the problem, is the problem.
Just checking what your objection is. Is it that you think experience is overrated, or is it just that he was speculating without any evidence?
Re: Twitter was down
#493Earlier quoted context omitted.
Rule one of having interns and retaining your sanity is that interns get their own branch to muck around in.
All changes should be in a new branch.
Re: Twitter was down
#494Earlier quoted context omitted.
Just checking what your objection is. Is it that you think experience is overrated, or is it just that he was speculating without any evidence?
Can't speak for OP, but I can tell you what mine is. If you have an intern or a Junior Engineer, they should have a more senior engineer to monitor and mentor them. In the situation where a Junior Engineer gets blamed for a screw up: 1. The Senior Engineer failed in their responsibility. 2. The Senior Engineer failed in their responsibility. A Junior Engineer should be expected to write bad code, but not put it into…
(I agree with the premise that an intern or junior eng is supposed to be mentored, and their mistakes caught. How else should they learn?)
Re: Twitter was down
#495Earlier quoted context omitted.
lol yes, whats the quote on "Don't assume bad intention when incompetence is to blame"? After seeing how people write code in the real world, I'm actually surprised there aren't more outages.
Well we have an entire profession of SRE/Systems Eng roles out there that are mostly based on limiting impact for bad code. Some of the places I've worked with the worst code/stacks had the best safety nets. I spent a while shaking my head wondering how this shit ran without an outage for so long until I realized that there was a lot of code and process involved in keeping the dumpster fire in the dumpster.
Re: Twitter was down
#496It wouldn’t be surprising if a large number of people, as of 2019, are secretly rooting for Twitter to permanently go away.
Re: Twitter was down
#497Earlier quoted context omitted.
#4 I work at Facebook. I worked at Twitter. I worked at CloudFlare. The answer is nothing other than #4. #1 has the right premise but the wrong conclusion. Software complexity will continue escalating until it drops by either commoditization or redefining problems. Companies at the scale of FAANG(+T) continually accumulate tech debt in pockets and they eventually become the biggest threats to availability. Not the ne…
since all of them happen in high profile business hours, i'd guess either #1 or #5. For #4 to be the actual cause, outages out of business hours would be more prevalent and longer.
https://twitter.com/internetarchive/status/11436045396956160...
https://twitter.com/internetarchive/status/11433789908260044...
Re: Twitter was down
#498Ok, this is too many high-profile, apparently unrelated outages in the last month to be completely a coincidence. Hypotheses: 1) software complexity is escalating over time, and logically will continue to until something makes it stop. It has now reached the point where even large companies cannot maintain high reliability. 2) internet volume is continually increasing over time, and periodically we hit a point where…
I (don't) like how you exclude Russia, China, Iran and somebody from your definition of 'us'.
Re: Twitter was down
#499Earlier quoted context omitted.
> We notice errors more now. Mistakes are instantly news. Heck, just look at Twitter itself from its original "Fail Whale" days where there was so much downtime, to now where even this relatively small amount of downtime is the top story on HN for hours.
So, when it went down, was there a Fail Whale displayed during this most recent incident?
I looked it up: in 2013, because they didn't want to be associated w/ outages.