Note that they have not actually dropped the price yet: https://x.com/OpenAIDevs/status/1932463601119637532 > We’ll post to @openaidevs once the new pricing is in full effect. In $10… 9… 8… There is also speculation that they are only dropping the input price, not the output price (which includes the reasoning tokens).
I think that was a joke. New pricing is already in place: Input: $2.00 / 1M tokens Cached input: $0.50 / 1M tokens Output: $8.00 / 1M tokens https://openai.com/api/pricing/ Now cheaper than gpt-4o and same price as gpt-4.1 (!).
OpenAI dropped the price of o3 by 80%
21–30 of 518 posts
Re: OpenAI dropped the price of o3 by 80%
#22how do we know it's not a quantized version of o3? what's stopping these firms from announcing the full model to perform well on the benchmarks and then gradually quantizing it (first at Q8 so no one notices, then Q6, then Q4, ...). I have a suspicion that's how they were able to get gpt-4-turbo so fast. In practice, I found it inferior to the original GPT-4 but the company probably benchmaxxed the hell out of the tu…
Re: OpenAI dropped the price of o3 by 80%
#23Personally I've found these bigger models (o3/Claude 4 Opus) to be disappointing for coding.
Re: OpenAI dropped the price of o3 by 80%
#24...how? I'd understand a 20-30% price drop from infra improvements for a model as-is, but 80%? I wonder if "we quantized it lol" would classify as false advertising for modern LLMs.
Re: OpenAI dropped the price of o3 by 80%
#25how do we know it's not a quantized version of o3? what's stopping these firms from announcing the full model to perform well on the benchmarks and then gradually quantizing it (first at Q8 so no one notices, then Q6, then Q4, ...). I have a suspicion that's how they were able to get gpt-4-turbo so fast. In practice, I found it inferior to the original GPT-4 but the company probably benchmaxxed the hell out of the tu…
Quantization is a massive efficiency gain for near negligible drop in quality. If the tradeoff is quantization for an 80 percent price drop I would take that any day of the week.
Hmm, that's evidently and anecdotally wrong:
Re: OpenAI dropped the price of o3 by 80%
#26how do we know it's not a quantized version of o3? what's stopping these firms from announcing the full model to perform well on the benchmarks and then gradually quantizing it (first at Q8 so no one notices, then Q6, then Q4, ...). I have a suspicion that's how they were able to get gpt-4-turbo so fast. In practice, I found it inferior to the original GPT-4 but the company probably benchmaxxed the hell out of the tu…
I swear every time a new model is released it's great at first but then performance gets worse over time. I figured they were fine-tuning it to get rid of bad output which also nerfed the really good output. Now I'm wondering if they were quantizing it.
Re: OpenAI dropped the price of o3 by 80%
#27It's going to be a race to the bottom, they have no moat.
Once new MacBooks and iPhones have enough memory onboard this is going to be a disaster for OpenAI and other providers.
Re: OpenAI dropped the price of o3 by 80%
#28You know. because LLMs can only be built by corporations... but because they're so easy to build, I see the price going down massively thanks to competition. Consumers benefit because all the companies are trying to out run each other.
Re: OpenAI dropped the price of o3 by 80%
#29always seemed to me that efficient caching strategies could greatly reduce costs… wonder if they cooked up something new
Re: OpenAI dropped the price of o3 by 80%
#30Earlier quoted context omitted.
I swear every time a new model is released it's great at first but then performance gets worse over time. I figured they were fine-tuning it to get rid of bad output which also nerfed the really good output. Now I'm wondering if they were quantizing it.
[flagged]