Earlier quoted context omitted.
> API that is likely a loss-leader to grab market share (hosted LLM cloud models). I don't think so, not anymore. If you look at API providers that host open-source models, you will see that they have very healthy margin between their API cost and inference hardware cost (this is, of course, not the only cost) [1]. And that does not take into account any proprietary inference optimizations they have. As for closed-mo…
I use whisper to transcribe long conversations, and deploying the model myself on vastai is ten times cheaper than OpenAI's API offer.
LLMs are cheap
311–319 of 319 posts
Re: LLMs are cheap
#312Earlier quoted context omitted.
I don’t completely disagree, but “assertion one” [1] [1] ~ you can obviously verify this yourself by doing it yourself and seeing how expensive it is. …is an enormously weak argument. You suppose. You guess. We guess. Let’s be honest, you can just stop at: > I don’t think so. Fair. I don’t either; but that’s about all we can really get at the moment afaik.
he's not wrong, if you can run a open weights model in any cloud, you can very straightforwardly estimate the cost of running the model. considering that these providers either use long-term contracts or maybe even buy their own hardware, this theoretical cloud deployment is itself an overestimate of the costs
But:
A) it makes absolutely no difference to the fact you have no idea what the big LLM providers are actually doing.
B) Just asserting some random thing and saying “anyone competent can verify this themselves” is a weak argument. Youre saying youve done the research, but failing to provide any evidence you actual have
If youve crunched the numbers then man up and post them.
If not, then stop at “I think…”
“This is based on my experience running production workloads…” is a nice way of saying “I dont have any data to backup what Im saying”.
If you did, you could just link to it.
…by not posting data you make your argument non-falisifyable.
It is just an oppinion.
Re: LLMs are cheap
#313Re: LLMs are cheap
#314Earlier quoted context omitted.
> You think AWS are going to subsidies your usage of somebody else’s models Yes >indefinitely? No, and that's the point.
Have they ever actually done this? I can't think of a time they've actually raised their prices ever that isn't the Route53 passing on registrar costs.
What company selling a primarily AI-based service right now is making a profit on that service?
Re: LLMs are cheap
#315The ecological cost is far more important.
Arguably, the impacts in employment, careers, theft of creative works, and other damage of inexpensive LLM bots are in the short term, and terms of impacts on mere human lives, more imminently and pressingly so.
Re: LLMs are cheap
#316Earlier quoted context omitted.
> LLMs are text in -> text out, and you can drop in a new LLM and replace them without changing anything else. If anything, newer LLMs will just have more capabilities. I don't mean to be too pointed here, but it doesn't sound like you have built anything at scale with LLMs. They are absolutely not plug n play from a behavior perspective. Yes, there is API compatibility (text in, text out) but that is not what matter…
What kind of quirks have you seen that the next model wasn't better at?
Re: LLMs are cheap
#317The entire comparison hinges on people only making simple factual searches ("what is the capital of USA") on both search engines and LLMs. I'm going to say that's far enough from the standard use case for both these sets of APIs to be entirely meaningless. - If I'm using a search engine, I want to search the web. Yes these engines are increasingly providing answers rather than just search results, but that's a UI/pro…
>If I'm using a search engine, I want to search the web. Yes these engines are increasingly providing answers rather than just search results, but that's a UI/product feature rather than an API one. This is a great point, lets hold onto that. >If I'm using an LLM, it is for parsing large amounts of input data, image recognition, complex analysis, deep thinking/reasoning, coding. Strongly disagree. Sometimes when goog…
Yes. The reason being that Google does not want you to ure other websites than Google.
Re: LLMs are cheap
#318Earlier quoted context omitted.
Sure. To clarify, I'm not asserting that in two years today's 4o will be 100x more expensive, but the sum of many core offerings from various companies will be. It won't be unheard of for people to spend $2k-$10k/yr between many AI services
It's not unheard of for people to spend $13,000 on a pure silver frying pan. Is it common? No.
Re: LLMs are cheap
#319:pop-corn: