Live data from Hacker News

Hetzner introduces GPU server for AI training

hetzner.com

31–40 of 57 posts

Re: Hetzner introduces GPU server for AI training

#31
post #17

Earlier quoted context omitted.

Even when it does, GPUs with more VRAM than a flagship gaming model have always and will always come at a massive price premium. That ceiling is currently 24GB, if you need more than that it's going to cost you.

I don't see why. AMD could buy a lot of free community work just by putting out a 32 GB version of their next gen flagship at around 1k. Not chasing the margin for a generation would buy them a decent amount of community support - they need it if they want to compete with Nvidia at any point. Ahh I just saw OP was saying sub 100$ - yeah that's never going to happen.

> next gen flagship at around 1k

Have you seen GPU prices lately (last 5 years)? A RTX 4080 super with merely 16 GB of VRAM is at least 1k, there's no way a 32 GB next gen flagship would be released at this price range.

Re: Hetzner introduces GPU server for AI training

#33

Earlier quoted context omitted.

That's incorrect, dedicated Hetzner servers do not have hourly pricing. That only applies to Hetzner Cloud.

That's incorrect. Dedicated servers do now offer hourly pricing. See this page which lists the hourly pricing for the new GEX130 server: https://www.hetzner.com/dedicated-rootserver/matrix-gpu/ But, as far as I know (hence the sibling question), you still need to pay the setup fee each time you launch.

Thanks for the link, seems like a major flaw to combine hourly pricing with a setup fee. :')

Re: Hetzner introduces GPU server for AI training

#34

Earlier quoted context omitted.

That's incorrect, dedicated Hetzner servers do not have hourly pricing. That only applies to Hetzner Cloud.

That's incorrect. Dedicated servers do now offer hourly pricing. See this page which lists the hourly pricing for the new GEX130 server: https://www.hetzner.com/dedicated-rootserver/matrix-gpu/ But, as far as I know (hence the sibling question), you still need to pay the setup fee each time you launch.

Actually rechecking the terms of the GPU server, it states:

> Cancellation period: 30 days to the end of the month

To me that indicates it's not possible to order the server for only a few hours.

Re: Hetzner introduces GPU server for AI training

#35
post #25

I use runpod or vast for training my (small) models (a few million parameters) mostly using RTX4090 up to 4 GPUs. Training is a sporadic task. Is not worth it for me to book it monthly (at these prices)

Hetzner customer here. It’s a little hard to understand in the UX, but the price shown is the monthly max price. It is actually paid by the hour. The price per hour for this server is € 1.5980 more info: https://docs.hetzner.com/general/others/new-billing-model/

That's the second time I read this comment and I still don't believe it: it's listed as a "dedicated root server" (usually billed by the month) with no mention of typical cloud offers.

Could you please clarify?

Re: Hetzner introduces GPU server for AI training

#37

Earlier quoted context omitted.

That's incorrect. Dedicated servers do now offer hourly pricing. See this page which lists the hourly pricing for the new GEX130 server: https://www.hetzner.com/dedicated-rootserver/matrix-gpu/ But, as far as I know (hence the sibling question), you still need to pay the setup fee each time you launch.

Actually rechecking the terms of the GPU server, it states: > Cancellation period: 30 days to the end of the month To me that indicates it's not possible to order the server for only a few hours.

You’re right, the billing is hourly (until you reach the monthly cap) but you also have this cancellation period. That’s weird?

Re: Hetzner introduces GPU server for AI training

#38

Earlier quoted context omitted.

That's incorrect. Dedicated servers do now offer hourly pricing. See this page which lists the hourly pricing for the new GEX130 server: https://www.hetzner.com/dedicated-rootserver/matrix-gpu/ But, as far as I know (hence the sibling question), you still need to pay the setup fee each time you launch.

Actually rechecking the terms of the GPU server, it states: > Cancellation period: 30 days to the end of the month To me that indicates it's not possible to order the server for only a few hours.

That term appears on all servers, but seems to be in conflict with their billing FAQ: https://docs.hetzner.com/general/others/new-billing-model/

Either way... it's a mess. They added hourly billing for dedicated servers six months ago, there's not much excuse for still having contradictory information hanging around.

Re: Hetzner introduces GPU server for AI training

#39

What’s the most cost effective option for hosting an llm these days? I don’t need to train, I just want to use one of the llama models for inference to reduce my reliance on 3rd parties.

Consider also an online llama as a service like deepinfra. I have a local 3090 for playing around with the smaller models, but it's nice having the option of calling the 405b.

Re: Hetzner introduces GPU server for AI training

#40

What’s the most cost effective option for hosting an llm these days? I don’t need to train, I just want to use one of the llama models for inference to reduce my reliance on 3rd parties.

If you don't need a big model and are fine with hosting locally, an RTX3060 with 12GB VRAM is going to do just fine. Can be bought for about 200-300 USD.

I've been pleasantly surprised by what such a mediocre GPU and Llama3 8B can do for certain (simple) use cases. Ollama makes it all pretty easy.

Post reply on HN