OVH Incident in Strasbourg
21–30 of 207 posts
Re: OVH Incident in Strasbourg
#22It started with all our SBG servers going down simultaneously. Approximately 1h later all our RBX servers went down as well including the OVH status page and all other OVH web applications. Either their SBG and RBX data centers are somehow connected or those are indeed two independent incidents.
Re: OVH Incident in Strasbourg
#23Earlier quoted context omitted.
It seems more likely that their data centers aren't quite as isolated as they thought they'd be. The outage also appears to be limited to their locations in Europe.
I always wondered why they promote having 6 datacentres in Roubaix when google maps shows that they're all within 50m. Can't be too much redundancy there.
Re: OVH Incident in Strasbourg
#24wow, yesterday I was playing with their public cloud because considering choosing them. I had some connection problem with my private networking there (deleted it more than once) and opened a ticket. If it was me... sorry, haha. Not good advertisement but it can happen to everyone.
Re: OVH Incident in Strasbourg
#25ovh.com looks down for me too.
You can check it's hosted by OVH:
$ whois $(dig sudokugarden.de +short)
Re: OVH Incident in Strasbourg
#26I imagine Mr Good Guy at OVH telling some others: "guys we have a single point of failure in our architecture with SBG, maybe we should... - naaah it's fine, we do not have time nor resources" Then shit happens. edit: I have no idea what is happening exactly, but OVH being what it is, it seems extremely weird that all datacenters "can" get down at the same time, and it looks like a serious architecture problem to me…
That seems a little uncharitable.
No expert on the field but that's the first time I can remember that a provider of that size loses connection to most of their data centres at once. That can happen with one product (eg S3 failure) but datacentre switches should work even if the rest is on fire.
Re: OVH Incident in Strasbourg
#27Anybody knows ETA?
Re: OVH Incident in Strasbourg
#28Some servers in GRA still appear to work if that's of any help. All data centres offline at once sounds more like an attack than a power failure in one location. According to them, there was a power failure in SBG but I don't see how that should affect routing in data centres several hundred miles away. https://twitter.com/olesovhcom/status/928541667283623936 EDIT: Maybe related to the Cisco issue? https://blogs.cisc…
> According to them, there was a power failure in SBG but I don't see how that should affect routing in data centres several hundred miles away. According to them, they do their routing in SBG, so it's plausible that it could lead to all of their network being down.
Re: OVH Incident in Strasbourg
#29Re: OVH Incident in Strasbourg
#30I moved away from OVH after I paid 3 months advance (~$300) for a server which burned down after 1 1/2 months. They did not issue any refunds (data, blood, sweat and tears were lost that day). I have been an OVH customer for 12 years. Today, I'm glad to have moved away all my production environments as well.
I'm still at OVH (support is reasonable and prices cheap) but would never trust one provider with all my infrastructure. In the end it always turns out that there is a single point of failure and if it's just the billing department. Using two providers protects you from that and if you chose some with good peering and free traffic, keeping both in sync is relatively easy. That's just the problem with services as AWS,…
Having one AWS account scares the crap out of me as well. It’s never a good thing if all your eggs are in one basket.
My money is on stuff spread across Bytemark, Linode and DigitalOcean with a DR plan involving mostly automatic recovery.
AWS doesn’t get a look in as it is extremely costly to port away from anything that isn’t bare metal and pipes.