Earlier quoted context omitted.
Really, downtime of google so far this year: More than 0 minutes. Downtime of my own server: 0 minutes. Of course I'm not trying to run a massively scalable service coping with millions of customers, because I don't need that.
You think that's how statistics works? Obviously it's possible to keep your own server running with 0 downtime. But you are exposed to a much higher risk of severe downtimes longer than the cloud provider would likely be. Be it hardware failure, grid, ISP, whatever.
Google Outage in Europe
141–150 of 186 posts
Re: Google Outage in Europe
#142Earlier quoted context omitted.
> Googler here, not speaking on the behalf of the company, my opinions are my own Why do employees at big tech names (FAANG et al.) are so often so cautious as to include this as a foreword everywhere? Twitter bios are full of that, for instance. It is crazy to me; who would expect anything else that our opinions being your own and nothing more? Who would expect that your word (with all due respect) is worth anything…
Can't speak for other companies, but this is covered in basic training at Google - if you're not authorized to speak on behalf of the company, you must make it clear when your writings may be mistaken or constructed to represent the company. Basically the company has specially trained people that speak on behalf of the company, and that message should not be confounded by personal opinions of other employees. For exa…
Though I almost got to be the official spokesperson for British Telecom responding on the alt.2600 news group about the the Met police VMB hack - press office was cool but the internal security was not.
Re: Google Outage in Europe
#143Earlier quoted context omitted.
>It’s still less downtime than your own server what makes you think so, actually?
In my experience, it's very easy to achieve high uptime through luck. If I only run a single server, and it has a 10% chance of failing in any given year, I have a 90% chance of achieving 100% uptime in a given year. In my experience it's also very easy to think you've got all your bases covered when actually you haven't. I'm protected against mains power failures by a UPS and a generator - but a UPS switchgear fault…
And an excellent corollary to this is that when you have lucky 100% uptime there is no incentive to optimize mean time to recovery.
Sure the raspberry pi in your closet has been running fine for years, 100% uptime, but then a component fails. Do you have a replacement on hand? Are you continuously monitoring it to know it went down? The component failed at 3am, did it page you? Did you hop right out of bed to rush to fix it?
Single systems can have really nice uptime until they don’t. Then you are hoping that the people on hand can repair what’s going on after months or years of never having to do that. Mean time to recovery might be a week while you wait for new hardware or a few hours while you google some error message you’ve never encountered.
People can run their own systems if they want to, but they shouldn’t confuse good luck with rigorous engineering.
Re: Google Outage in Europe
#144Earlier quoted context omitted.
>It’s still less downtime than your own server what makes you think so, actually?
In my experience, it's very easy to achieve high uptime through luck. If I only run a single server, and it has a 10% chance of failing in any given year, I have a 90% chance of achieving 100% uptime in a given year. In my experience it's also very easy to think you've got all your bases covered when actually you haven't. I'm protected against mains power failures by a UPS and a generator - but a UPS switchgear fault…
The server not failing is not the only outage mode.
Re: Google Outage in Europe
#145I do laugh at people that say "Using the cloud means less downtime than your own server", then things like this come along :D
Right? My private website on my RPI is running now 2 Years without a problem and only minimal downtime due to rebooting for the new kernels. It is amazing how much uptime you can achieve with a 5$ Computer in comparison to a 1730000000000$ (1,73 tera $) Company. Even if you compensate for dynamic content.
Was your home internet available all of the time? How many times did you reboot your modem?
Re: Google Outage in Europe
#146Earlier quoted context omitted.
Googler here, not speaking on the behalf of the company, my opinions are my own People do absolutely NOT get fired over incidents. Making mistakes is human. An incident will prompt a review of the systems and safeguards in place to prevent such an incident, much like an airline incident investigation - basically "somebody fat-fingered it" is never the answer, postmortems are always blameless EDIT: now that I think of…
> Googler here, not speaking on the behalf of the company, my opinions are my own Why do employees at big tech names (FAANG et al.) are so often so cautious as to include this as a foreword everywhere? Twitter bios are full of that, for instance. It is crazy to me; who would expect anything else that our opinions being your own and nothing more? Who would expect that your word (with all due respect) is worth anything…
Re: Google Outage in Europe
#147Earlier quoted context omitted.
139,995 employees at Google * 1,000,000 = $139,995,000,000 $140 billion dollars. On training. On the one hand... you know what, I'd love to work in an environment like that. Seriously. On the other hand... what's the argument you make to the CFO in support of this? Honest question, interested to hear answers.
I work part time in the Army. In the Army when you go from their equivalent of junior to mid-level they take you out of your job for eight months of dedicated personal development, before you start your first mid-level job. When you go to their equivalent of senior they take you out for a year . I don't know how much that costs all-in, including the salaries, instructors, facilities, but might be starting to approach…
Re: Google Outage in Europe
#148There were outages around the same time last year. Somebody in the HN thread commented back then that the employees evaluation and promotion window ends around december/eoy, thus more releases are made. https://en.m.wikipedia.org/wiki/Google_services_outages
Google's more interested in placating their primadonna engineers than solving customer problems.
Re: Google Outage in Europe
#149Earlier quoted context omitted.
More reason to promote for self-host, and decentralised systems like smtp, matrix.org, ActivityPub. The idea of having all the data and server concentrated in a few player such as google, amazon, ccp is not reliable for the digital operation of the planet.
The question is: should we trust small, underfunded, hobbyist servers more than large corporations that have a money-driven reason to maintain high quality of services?
for those with budgets to hire server administrators (or pay for third-party managed services)? yes.
second category includes almost anyone with a large following like institutional users.
hell the incumbent social media service operators can white-label their existing software and sell this as a service.
Re: Google Outage in Europe
#150Earlier quoted context omitted.
> Googler here, not speaking on the behalf of the company, my opinions are my own Why do employees at big tech names (FAANG et al.) are so often so cautious as to include this as a foreword everywhere? Twitter bios are full of that, for instance. It is crazy to me; who would expect anything else that our opinions being your own and nothing more? Who would expect that your word (with all due respect) is worth anything…
I work at a large tech company, and they do mention in the on boarding materials that we represent the company, so we should be careful in our social media profiles. My solution to this is to not associate my social media profiles with my employer. This is technically not really what we’re supposed to do, and I might have to change that approach at some point if I move high enough in the org to start getting attentio…