Live data from Hacker News

Fly.io Postgres cluster down for 3 days, no word from them about it

webcache.googleusercontent.com

341–350 of 493 posts

Re: Fly.io Postgres cluster down for 3 days, no word from them about it

#341
post #99

It's really disappointing that they made this forum thread private, apparently in response to this HN thread blowing up. This is the first negative HN thread I've seen about them, it's not even really that bad because this kind of downtime is expected, and they can't get to every forum post, and their response that someone posted here is totally reasonable in my opinion. So why is the link to the thread 404ing and wh…

(a) Not even close to the first negative HN thread about us. (b) We definitely didn't make the thread private in response to HN. (c) It should be public again.

I wonder if there will ever be a wake up call to the arrogance of people at fly.io

At work when it came up in a meeting people went around with horror stories of broken elements while the status page wasn't updated, terrible communication and an overall attitude that nothing is wrong, even when servers go down for days at a time.

Re: Fly.io Postgres cluster down for 3 days, no word from them about it

#342

Holy hell, there are some hostile comments here! I've had service issues on Fly that I've escalated to support in the past, and given my experience it feels highly unlikely that they tried sweep this under the rug or somesuch. At the time we had deployed a small business workload (few 100$/mo in billings) and paid for their $29 support plan, so grain of salt there. We faced service issues and, while the service relia…

> I think the assumption of badfaith in this thread from Fly is pretty unprofessional.

Customers aren't supposed to show professionalism. Service providers are. I didn't see disrespectful comments here.

People here are just poiting this has happened many times and look like a pattern. If you don't fix a communication issue after multiple occurrences, you might not be ill intentioned, but at least careless.

Re: Fly.io Postgres cluster down for 3 days, no word from them about it

#343
post #124

Earlier quoted context omitted.

I'm using Digital Ocean App platform, which does pretty much everything for me. It's very simple to use. I can run my app as a single developer without caring about infrastructure for 99% of the time.

Do they offer authentication/authorization? This is the one thing I need in every app and don't want to do myself.

Would you consider a project like https://github.com/authcompanion/authcompanion2 for the authentication side? Missing anything?

Re: Fly.io Postgres cluster down for 3 days, no word from them about it

#344

Earlier quoted context omitted.

what's up with the status page?

Status pages are usually for marketing purposes. Why would anyone want to become a new customer if all they see is jumble of green, yellow and red? Green status pages attract business.

thanks for sharing your opinion, but I was looking for a reply from someone inside fly.io

Re: Fly.io Postgres cluster down for 3 days, no word from them about it

#345
post #166

I would use Fly for hosting servers, but never databases. For this very reason.

... and there are much cheaper places to host a server if you don't care about databases, like bare metal hosters and tier-2 VPSes with good reliability like Vultr and Digital Ocean.

Re: Fly.io Postgres cluster down for 3 days, no word from them about it

#346
post #78

This is why people should just run a small managed k8s cluster on GCP. This has worked super well for us. https://www.vmii.org/blog/2023/03/12/kubernetes/

I tried using kubernetes a while back for hosting a side project on a raspberry pi. I guess technically I was running microk8s on the pi and had to install kubectl locally to interact with it. I actually like some of the concepts, like pods and ingress, but one thing I noticed that I didn't like, as far as I remember, was that there's not really a good way in kubernetes to make your YAML more dynamic. Apparently you'…

Yes, the post I linked describes how to do that. It took time to learn. Nowadays I’d just do it with GPT!

Re: Fly.io Postgres cluster down for 3 days, no word from them about it

#347

Holy hell, there are some hostile comments here! I've had service issues on Fly that I've escalated to support in the past, and given my experience it feels highly unlikely that they tried sweep this under the rug or somesuch. At the time we had deployed a small business workload (few 100$/mo in billings) and paid for their $29 support plan, so grain of salt there. We faced service issues and, while the service relia…

> I think the assumption of badfaith in this thread from Fly is pretty unprofessional. Customers aren't supposed to show professionalism. Service providers are. I didn't see disrespectful comments here. People here are just poiting this has happened many times and look like a pattern. If you don't fix a communication issue after multiple occurrences, you might not be ill intentioned, but at least careless.

(Note: I edited my comment to make it clear I'm referring to badfaith from commenters, quote above is from pre-edit)

I'd argue the expectation goes both ways. I won't link to specific comments, but I think it's pretty clear that some of them cross the line to disrespectful.

Re: Fly.io Postgres cluster down for 3 days, no word from them about it

#348

Earlier quoted context omitted.

Does anyone know how Hetzner pricing is half of DO yet is profitable, while DO is loss making with 6% operating margin?

Simple, Hetzner mainly operates on Germany, the people are mostly Germans, and they automate the stuff to a point a small team could manage it well even if not remotely, so they have less cost on human resources.

I've wondered how they can host this cheap in Germany given their very high electricity prices.

Maybe that's not actually the dominant cost, or they've optimized everything else so well they can just eat the electric bill.

Re: Fly.io Postgres cluster down for 3 days, no word from them about it

#349

Earlier quoted context omitted.

[flagged]

Was the sarcasm an attempt to discredit criticism of Fly's operational processes by pointing out that another company also has issues in how they handle outage notifications?

You could have just answered my previous question with "No, I am not familiar with sarcasm".

Because you clearly don't understand sarcasm, I'll be blunt:

No, I'm not trying to discredit any criticism of this provider. I agree with the comment I replied to, that this kind of failure mode is fucking ridiculous. My response thus is not an attempt to normalise this, but to highlight the elephant in the room, which is that AWS - the gold standard for "hosting" services for many a startup and techbro - *also* has Rube Goldberg like levels of interdependence that cause cascading failures *every time* something goes wrong, and *also* have a status board so confidently green that it may as well be an ad for lawn care products.

Re: Fly.io Postgres cluster down for 3 days, no word from them about it

#350
post #247

Earlier quoted context omitted.

What do they do that makes them more efficient?

i'll guess they pick optimized components for it. like the longtime workhorse was a high performance skylake desktop cpu w/o ecc ram

The secret is in the cooling system. They have individual cooling systems for each server. Less heat = longer sustained loads
Post reply on HN