Live data from Hacker News

Hetzner is working on LLM Inference

sliplane.io

11–20 of 92 posts

Re: Hetzner is working on LLM Inference

#11
post #5
post #3

Good to see more developments in this space. I quite like this service, which is a little further than Hetzner and has several models to choose from: https://www.infomaniak.com/en/hosting/ai-services

Infomaniak is such a shitty company, I had to use them on a previous job I worked at and dealing with them was awful.

That makes it sound like dealing with Hetzner is easy. Is the case?

Re: Hetzner is working on LLM Inference

#12

This seems like a smart move, given their ability to host efficiently. I approve of efforts to make the cost of inference for smaller useful models slowly approach 'close to zero' and there are many good paths for getting there. It is useful for companies to get fast hosting for the class of smaller models they may end up hosting in house.

I really hope they dont stop at the small models though! The bigger ones that dont fit on a single GPU are more interesting I think

Re: Hetzner is working on LLM Inference

#13

It would certainly be interesting to have a highly respected EU-native inference provider, if only to make the regulatory gods happy

Scaleway have two separate (one fully managed one a bit less) services for that:

https://www.scaleway.com/en/generative-apis/

https://www.scaleway.com/en/inference/

Re: Hetzner is working on LLM Inference

#15
post #13

It would certainly be interesting to have a highly respected EU-native inference provider, if only to make the regulatory gods happy

Scaleway have two separate (one fully managed one a bit less) services for that: https://www.scaleway.com/en/generative-apis/ https://www.scaleway.com/en/inference/

yea, but no prompt caching right? This makes it unusable for my usecase at least, the cost would be insane

Re: Hetzner is working on LLM Inference

#17
Potentially interesting article ruined by AI slop hallucinations like

> For now, the API is fast, free, and fun to try. The next hardware announcement will tell us much more than another small model would.

Re: Hetzner is working on LLM Inference

#18
post #17

Potentially interesting article ruined by AI slop hallucinations like > For now, the API is fast, free, and fun to try. The next hardware announcement will tell us much more than another small model would.

what is the hallucination here? it is fast, free and fun to try. And I genuinely think that the hardware decision (if they get bigger gpus) decides if this will be a banger product or not?

Re: Hetzner is working on LLM Inference

#19
post #11
post #5

Earlier quoted context omitted.

Infomaniak is such a shitty company, I had to use them on a previous job I worked at and dealing with them was awful.

That makes it sound like dealing with Hetzner is easy. Is the case?

Depends on your mindset. If you are an Engineer that does not need hand-holding and you accept terse answers from support with the gratitude to the human on the other side, then go for it.

You will at least not need months of training and tough exams to be an expert in Hetzner Cloud. It is simple, but that's the point.

Re: Hetzner is working on LLM Inference

#20
post #11

Earlier quoted context omitted.

That makes it sound like dealing with Hetzner is easy. Is the case?

Depends on your mindset. If you are an Engineer that does not need hand-holding and you accept terse answers from support with the gratitude to the human on the other side, then go for it. You will at least not need months of training and tough exams to be an expert in Hetzner Cloud. It is simple, but that's the point.

I mostly use Hetzner baremetal servers and not cloud, but the support there is in my experience very competent, but also VERY german and direct. I can imagine if youre neither german or super technical that this can be intimidating or considered unfriendly
Post reply on HN