Earlier quoted context omitted.
I'd be very surprised if that were the case, unless they have severely devalued how much "100% usage" is worth – which, as I understand, they could at any point, given that they don't publicly specify how many tokens (or at least "credits" [1]) are included in each plan per month. It really reminds me of pay-to-win games at this point: Two currencies (credits, tokens), both with a floating, intransparent exchange rat…
Yeah, it's not even close. $20/m = appox API $700/m $100/m = appox API $3,500/m $200/m = appox API $14,000/m [0] https://www.reddit.com/media?url=https%3A%2F%2Fpreview.redd....
OpenAI: GPT 5.6 Sol price reduction (until at least Nov 21)
321–330 of 352 posts
Re: OpenAI: GPT 5.6 Sol price reduction (until at least Nov 21)
#322Earlier quoted context omitted.
Yeah, it's not even close. $20/m = appox API $700/m $100/m = appox API $3,500/m $200/m = appox API $14,000/m [0] https://www.reddit.com/media?url=https%3A%2F%2Fpreview.redd....
though of course, approaching that number requires being ready to continue using it every five hours, stopping and starting all hours of the day and every day including weekends and holidays. possible to set up a harness to do if you have one single expensive task, but for most spiky ask-a-query collaborative work not easy to use all of.
Re: OpenAI: GPT 5.6 Sol price reduction (until at least Nov 21)
#323The fact that AI models can be so easily distilled and replicated is such a stroke of luck. 10 or 15 years ago if one had asked me to envision a future where a private company invents artificial intelligence, I'd have thought for sure they'd have a massive moat, be very difficult to catch, and it would create an almost instant monopoly. Rather, it seems that selling intelligence might end up as a race to the bottom.…
It's a mistake to think only OpenAI and Anthropic are actually spending the big bucks on pretrain, and the others just distill that. The Chinese models are pretrained on large clusters just like OpenAI ones are. Yes, they use outputs of the frontier models to further improve the final model, but even without those outputs they'd still have very strong models. It's not like in a world without distillation things would…
Transformers are like just a step or two removed from being fancy convolutional neural networks. I guess I'm just surprised that it didn't turn out to require more 'special sauce' with extremely elaborate internal architectures, and less of a big-data approach.
Because the data is so central in building these LLMs, rather than some special insights or ideas in the model architecture, or very special hardware requirements, the field is much more open than I would have guessed some years ago. And it's the fact that the data is so central that makes distillation possible in the first place.
Re: OpenAI: GPT 5.6 Sol price reduction (until at least Nov 21)
#324The price difference to Deepseek models (deepseek-v4-flash, deepseek-v4-pro and deepseek-v4-flash-vision-exp) is still significant while the performance difference is not.
Performance difference it large by all benchmarks. DeepSeek fell behind. It's Kimi K3 or GLM-5.3 now.
Re: OpenAI: GPT 5.6 Sol price reduction (until at least Nov 21)
#325Earlier quoted context omitted.
though of course, approaching that number requires being ready to continue using it every five hours, stopping and starting all hours of the day and every day including weekends and holidays. possible to set up a harness to do if you have one single expensive task, but for most spiky ask-a-query collaborative work not easy to use all of.
Isn't the five hour limit effectively gone, at least on ChatGPT Plus? I haven't seen it in a while now.
Re: OpenAI: GPT 5.6 Sol price reduction (until at least Nov 21)
#326Re: OpenAI: GPT 5.6 Sol price reduction (until at least Nov 21)
#327Re: OpenAI: GPT 5.6 Sol price reduction (until at least Nov 21)
#328Earlier quoted context omitted.
Explain “exact same thing”? Like it load the Wikipedia page for Fort Worth, Texas, does it make Wikipedia spin up an editor to go write the article? I’m really not sure how you’re getting “exact” here.
Anthropic scraped the whole web, scanned every book, pirated every bit of media to feed in to their training. Distillers are doing essentially the same thing scraping all the knowledge from the LLM to create a training set for a new one. They are crying about theft after committing the largest theft in human history.
Re: OpenAI: GPT 5.6 Sol price reduction (until at least Nov 21)
#329Earlier quoted context omitted.
Nothing against what I said in there. Distillation is far from enough to get to the frontier. Its at best a ramp (the most efficient one used by everyone ) that shortcut and saves millions of rl runs before a model moves.
"No evidence" -> Large lab saying that they use distillation when training their models
Re: OpenAI: GPT 5.6 Sol price reduction (until at least Nov 21)
#330Earlier quoted context omitted.
It still occupies a place on the Pareto frontier.
IMO Pareto frontier is a lie. At any given moment there's only one "best" strategy and my guess would be in most cases it is just using the best model.
I'll be sure to tell the economists.
You're just price insensitive. The point of a Pareto frontier is mapping cost and utility. Not everyone can afford to run fable all day.