OpenAI dropped the price of o3 by 80%
11–20 of 518 posts
Re: OpenAI dropped the price of o3 by 80%
#12how do we know it's not a quantized version of o3? what's stopping these firms from announcing the full model to perform well on the benchmarks and then gradually quantizing it (first at Q8 so no one notices, then Q6, then Q4, ...). I have a suspicion that's how they were able to get gpt-4-turbo so fast. In practice, I found it inferior to the original GPT-4 but the company probably benchmaxxed the hell out of the tu…
Re: OpenAI dropped the price of o3 by 80%
#13how do we know it's not a quantized version of o3? what's stopping these firms from announcing the full model to perform well on the benchmarks and then gradually quantizing it (first at Q8 so no one notices, then Q6, then Q4, ...). I have a suspicion that's how they were able to get gpt-4-turbo so fast. In practice, I found it inferior to the original GPT-4 but the company probably benchmaxxed the hell out of the tu…
Are there any benchmarks that track historical performance?
a proxy to that may be the anecdotal evidence of users who report back in a month that model X has gotten dumber (started with gpt-4 and keeps happening, esp. with Anthro and OpenAI models). I haven't heard such anecdotal stories about Gemini, R1, etc.
Re: OpenAI dropped the price of o3 by 80%
#14Personally I've found these bigger models (o3/Claude 4 Opus) to be disappointing for coding.
Re: OpenAI dropped the price of o3 by 80%
#15...how? I'd understand a 20-30% price drop from infra improvements for a model as-is, but 80%? I wonder if "we quantized it lol" would classify as false advertising for modern LLMs.
Re: OpenAI dropped the price of o3 by 80%
#16Re: OpenAI dropped the price of o3 by 80%
#17how do we know it's not a quantized version of o3? what's stopping these firms from announcing the full model to perform well on the benchmarks and then gradually quantizing it (first at Q8 so no one notices, then Q6, then Q4, ...). I have a suspicion that's how they were able to get gpt-4-turbo so fast. In practice, I found it inferior to the original GPT-4 but the company probably benchmaxxed the hell out of the tu…
Some users. For me the drop was so huge it became almost unusable for the things I had used it for.
Re: OpenAI dropped the price of o3 by 80%
#18how do we know it's not a quantized version of o3? what's stopping these firms from announcing the full model to perform well on the benchmarks and then gradually quantizing it (first at Q8 so no one notices, then Q6, then Q4, ...). I have a suspicion that's how they were able to get gpt-4-turbo so fast. In practice, I found it inferior to the original GPT-4 but the company probably benchmaxxed the hell out of the tu…
Re: OpenAI dropped the price of o3 by 80%
#19Note that they have not actually dropped the price yet: https://x.com/OpenAIDevs/status/1932463601119637532 > We’ll post to @openaidevs once the new pricing is in full effect. In $10… 9… 8… There is also speculation that they are only dropping the input price, not the output price (which includes the reasoning tokens).
Input: $2.00 / 1M tokens
Cached input: $0.50 / 1M tokens
Output: $8.00 / 1M tokens
https://openai.com/api/pricing/
Now cheaper than gpt-4o and same price as gpt-4.1 (!).