With cache hit rates being effectively free, harnesses like Reasonix have let me do a month of work for less than 2 dollars. It's not even the subsidies making it cheap, American providers like Digital Ocean or Cloudflare host the same model with similar pricing.
How does caching help here? How much repetition is there in queries?
The unbearable cheapness of open weight models
81–90 of 195 posts
Re: The unbearable cheapness of open weight models
#82The giants knew this was coming, and soon 95% of AI tasks will be able to be done by open models (coding, research, cowork style work). So why pay a premium? Why use them at all? This leaves the labs with two options: 1) push the frontier in a way only massive scale can, and cash in on it (mythos level cyber security, recursive training, frontier science work). There’s big money for never before possible capabilities…
3) Buy all the RAM, increasing the barrier to entry to push back the tide a bit, in time for a juicy IPO.
Re: The unbearable cheapness of open weight models
#83Aren't these open models so cheap because they're (partially) chinese gov. sponsored, and because they're stealing and redistributing the IP that comes in?
Re: The unbearable cheapness of open weight models
#84Earlier quoted context omitted.
Won’t all they need to do is say “best in class, latest models, fastest” and wine and dine a few execs and those enterprise deals will be signed? In this case the people tasked with using the product won’t actually mind.
No one is getting fired for using SotA.
But that is months away, so not my problem?
Re: The unbearable cheapness of open weight models
#85Earlier quoted context omitted.
Mythos was outperformed by small, specific local models in multiple oss project.
i'd love to hear about this! do you have examples?
Harness example: https://github.com/evilsocket/audit
Re: The unbearable cheapness of open weight models
#86Let's imagine that Anthropic/OpenAI fail to manufacture scarcity by villainizing Open Weight models (a sincere probability). What is left for these corporations to prop up their prices, or any margin at all? I expect scaffolding around tool use, supporting bespoke implementation and driving risk down for institutional adoption. (They might even build an insurance tool to protect accountants/lawyers from errors in com…
> It seems plainly clear to me that information and information processing is commodifying (for the first time in human history?). Without the age-old bottlenecks at the top of the value chain, capital will surely flow downwards, right? Isn't this the thing people have said about every new technology since the printing press? And it has been mostly true, but it has also been the case that the incumbents have fought h…
Recall the notion of a bottleneck, and this distinction will become clear. Those prior technological changes never inverted a bottleneck, and this one does.
Re: The unbearable cheapness of open weight models
#87Re: The unbearable cheapness of open weight models
#88Earlier quoted context omitted.
3) Buy all the RAM, increasing the barrier to entry to push back the tide a bit, in time for a juicy IPO.
Buying all the RAM can't work forever. Scarcity increases prices, high prices increase supply, improves RAM R&D budgets, and forces users to find ways to economize around low RAM availability.
Re: The unbearable cheapness of open weight models
#89I don’t get it. So many here are saying open weight models will kill the frontier labs. But open source and similar have tried to beat private companies everywhere all the time, and people still buy the best products even if great open source alternatives are available. Why wouldn’t this be the case for AI too?
there is a larger appetite for something like open source AI mostly b/c of price. we all know these labs have not figured out their pricing model, and we're all holding our breath out of fear of what the prices could be.
also, if you consider that the only toll to knowledge work before was personal time, and now you need to pay $100s month just to keep up with the baseline speed. it makes sense people are looking for something that gets them back to a workflow where the price to do work is near $0.00.
I think for a smaller group though, it's more to do with a certain combination of principles. Some people don't want censorship, other's want ownership, some want the knowledge of working on LLMs to not be gate kept.