Earlier quoted context omitted.
I didn't find the answer there, that's why I asked. What hardware is needed, how much of it, cooling, and what does it all cost you? Or are you saying I can take my old desktop and serve Deepseek v3.2 to 10k users simultaneously and it would cost me about $1 per megatoken?
I'm simply saying this: there are third party hosters of Open Weight models like deepseek and they have been doing this for a while. Obviously they are not subsidised, do you disagree? If you agree, they have a way to price it at a point that people wanna pay for it and also they aren't losing money. So there's nothing inherent about inference that makes it too costly or whatever.
> Obviously they are not subsidised, do you disagree? If you agree, they have a way to price it at a point that people wanna pay for it and also they aren't losing money.
> So there's nothing inherent about inference that makes it too costly or whatever.
Do we have audited GAAP financial data for any of these companies? If we don't, all these are... vibes, man.