Live data from Hacker News

How much cost would it take to run a LLMs locally

news.ycombinator.com

1–7 of 7 posts

Re: How much cost would it take to run a LLMs locally

#2
First, check what LLMs your system can actually handle based on your RAM/VRAM, then choose the model variant accordingly.

If you want to use the model through an API for your website, a good setup would be a dedicated machine/server to host it, since running LLMs locally can consume a lot of memory.

Depending on your hardware, you can look at smaller variants of Gemma or Qwen Coder.

Another option is to use a hosted inference provider. It’ll cost a little, but you’ll usually get much faster inference compared to running locally especially if your system isn’t high-end.

Re: How much cost would it take to run a LLMs locally

#4

First, check what LLMs your system can actually handle based on your RAM/VRAM, then choose the model variant accordingly. If you want to use the model through an API for your website, a good setup would be a dedicated machine/server to host it, since running LLMs locally can consume a lot of memory. Depending on your hardware, you can look at smaller variants of Gemma or Qwen Coder. Another option is to use a hosted…

What should be the minimum hardware requirements I got Nvidia processor with a high quality graphic card and 16 gb ram