Besides from privacy: I already making twice now.you own the hardware and the price had doubled since i bought. Almost tripled. You missed the opportunity and i have 4 of those awesome machines. Cry on. I sell those to business who need local air gapped requirments and I make a lot more money! I can run the alliterated models where none of the service prvoider even dare to provide. THose benefits outweights a few K.…
> And show me an api provider that allows me to run 10x agents concurrently for 5 days straights . Any of them on a Max/Pro plan as long as you are smart about model selection? That's my main objection to local inference, I'd need a whole rack of GPUs to do as many things in parallel that I can do for $400 a month. I do plan on setting up some local inference hardware, but...RAM and GPU prices alone are $$$$
no , all you need is one small DGXSPark with proper setup.