Live data from Hacker News

The unbearable cheapness of open weight models

jamesoclaire.com

161–170 of 195 posts

Re: The unbearable cheapness of open weight models

#161

Earlier quoted context omitted.

License the training corpus and encourage copyright suits against outputs from models trained on unlicensed corpora.

This won't work if the courts decide that training is fair use, which certainly seems the direction they are going.

Output is a separate issue from training. Courts will never decide that a identical copy spit out by an LLM is non-infringing simply because it went through an LLM stage. Copyright laundering is wishful thinking by tech folks.

Re: The unbearable cheapness of open weight models

#162

Earlier quoted context omitted.

> So why are they losing so much money? Mostly training. Claude didn't just get to be so good at coding by magic, it was suddenly so good because they did truly staggering amounts of RLHF and RLAIF on it. They are still doing that today, on any tasks they can figure out how to evaluate it on. This is capex for them. Their margins on inference are >90% today for tokens they sell (plans are hard to count, but still pro…

> Their margins on inference are >90% today for tokens they sell (plans are hard to count, but still profitable). That doesn’t make any sense, it doesn’t add up. Have you seen how much money they’re raising and burning? We know that training does not cost tens of billions. Brockman said OpenAI expects to spend $50 billion on compute this year. OpenAI’s revenue run rate is less than $50 billion for this year! For 90%…

Capacity for inference isn't a cost issue, it's an availability issue. There just isn't enough hardware out there.

Re: The unbearable cheapness of open weight models

#163
post #65

Earlier quoted context omitted.

Maybe you missed the part where starlink / orbiting datacenters don't really have to even make money as long as they partially fund rocket launch tests. Or maybe you don't take Elon seriously when he talks about Mars.

> Maybe you missed the part where starlink / orbiting datacenters don't really have to even make money as long as they partially fund rocket launch tests. I am only dismissing the orbital data centres, I do see a future for Starlink. One with competition, but a future nonetheless. I'm old enough to remember the dot.com bubble and "we lose money on each unit and make up for it in scale": If they don't make sense, they…

Ok so you're ignoring the entire thing. Sigh.

Re: The unbearable cheapness of open weight models

#164

Earlier quoted context omitted.

This is an example of common knowledge that is wrong. People look at their cash burn, assume that they spend this to subsidize inference, and get bonkers answers. Inference is not their largest expense. Inference is cheap. Anthropic is only drastically subsidizing their plans if you count their training expenses as part of their costs.

Are you an anthropic insider or something? Because if you are you should delete this comment. If you aren’t then you don’t know what the hell you’re talking about.

For one point, you can look at the costs of similarly sized open-source models from inference providers (which are only making money on the markup on the compute), and compare with anthropic's prices. There's a pretty big price difference there and it would be hard to believe that anthropic's models are that much more expensive to run than those models.

Re: The unbearable cheapness of open weight models

#165

> What worries me about this is that Anthropic and OpenAI seem to have backed themselves into a corner of high costs. Can they reasonably decrease their prices by 20-50x to compete with DeepSeek or Xiaomi’s Mimo? They have high prices, not high costs. They will obviously keep prices as high as they can for as long as they can, while keeping demand up. Once demand starts to fall, so will the prices. > Are these models…

What are you even talking about? Everyone knows that Anthropic is drastically subsidizing their plans. It's actually the exact opposite of what you're talking about. The costs are extremely high and the prices are actually what's being subsidized and cheap right now.

I dont think that’s accurate. I mean look at how much more expensive frontier closed source models are vs something like glm 5.2 which is just about as good. Serving glm is really cheap, and high margin. Obviously no one knows, just how much their inference costs, but if we assume that opus/gpt are maybe 15-20% more parameters than glm 5.2, then it makes no sense for them to charge almost much much more than glm

Re: The unbearable cheapness of open weight models

#166
post #150

Earlier quoted context omitted.

Surely the same can be said for the people saying the opposite?

I didn’t make a claim. The parent explicitly said it was a misconception that inference is not profitable. No one knows if it’s profitable or not so we’re left to speculate.

You can make a fairly decent assumption by calculating the margin on serving glm 5.2, and adding say 30% extra costs and it still leaves a healthy margin

Re: The unbearable cheapness of open weight models

#167

Earlier quoted context omitted.

I didn’t make a claim. The parent explicitly said it was a misconception that inference is not profitable. No one knows if it’s profitable or not so we’re left to speculate.

You can make a fairly decent assumption by calculating the margin on serving glm 5.2, and adding say 30% extra costs and it still leaves a healthy margin

Where are you getting 30% from

Re: The unbearable cheapness of open weight models

#168

Earlier quoted context omitted.

You can make a fairly decent assumption by calculating the margin on serving glm 5.2, and adding say 30% extra costs and it still leaves a healthy margin

Where are you getting 30% from

It was a rough heuristic for how much more opus/5.5 would presumably cost if you extrapolate from glm5.2 prices. In any case input tokens for 5.4/4.6 are 70-100% more expensive, cached about 3-100%, and output tokens anywhere from 60%-240% more as per all their current api pricing. I highly doubt 5.4/4.6 are that much more expensive to serve given how cheap and commoditized inference has become, and how comparable they are perfomance wise.

Re: The unbearable cheapness of open weight models

#169
post #23

Earlier quoted context omitted.

3) Buy all the RAM, increasing the barrier to entry to push back the tide a bit, in time for a juicy IPO.

4) Make it illegal to use anything but regulated models.

Why illegal, just pass these 3000 pages FAA-level certification, export controls and KYC. We're free country, after all!

Re: The unbearable cheapness of open weight models

#170
post #65

Earlier quoted context omitted.

> Maybe you missed the part where starlink / orbiting datacenters don't really have to even make money as long as they partially fund rocket launch tests. I am only dismissing the orbital data centres, I do see a future for Starlink. One with competition, but a future nonetheless. I'm old enough to remember the dot.com bubble and "we lose money on each unit and make up for it in scale": If they don't make sense, they…

Ok so you're ignoring the entire thing. Sigh .

On the contrary: I've paid a lot of attention, causing me to look at it closely and determine it is a terrible idea worthy of an illustrated 5,000 word blog post explaining exactly how terrible.

If you build the DC satellites as currently specified, you're strictly better off not launching them. That's how bad the idea is.

Post reply on HN