Live data from Hacker News

OpenAI O3-Mini

openai.com

131–140 of 944 posts

Re: OpenAI O3-Mini

#131

> While OpenAI o1 remains our broader general knowledge reasoning model, OpenAI o3-mini provides a specialized alternative for technical domains requiring precision and speed. I feel like this naming scheme is growing a little tired. o1 is for general knowledge reasoning, o3-mini replaces o1-mini but might be more specialized than o1 for certain technical domains...the "o" in "4o" is for "omni" (referring to its mult…

They should be calling it ChatGPT and ChatGPT-mini, with other models hidden behind some sort of advanced mode power user menu. They can roll out major and minor updates by number. The whole point of differentiating between models is to get users to self limit the compute they consume - rate limits make people avoid using the more powerful models, and if they have a bad experience using the less capable models, or if they're frustrated by hopping between versions without some sort of nuanced technical understanding, it's just a bad experience overall.

OpenAI is so scattered they haven't even bothered using their own state of the art AI to come up with a coherent naming convention? C'mon, get your shit together.

Re: OpenAI O3-Mini

#133

Earlier quoted context omitted.

Name is just a label. It's not supposed to mean anything.

Think how awesome the world would be if labels ALSO had meanings.

As someone else said in another thread, if you could derive the definition from a word, the word would be as long as the definition, which would defeat the purpose.

Re: OpenAI O3-Mini

#134

Earlier quoted context omitted.

You know you are running an extremely nerfed version of the model, right?

I did update my comment, but said that I am using the distilled version, so yes?

Even the full model scores below Claude on livebench so a distilled version will likely be even worse.

Re: OpenAI O3-Mini

#135

What is the comparison of this versus DeepSeek in terms of good results and cost?

Probably a good idea to wait for external benchmarks like Aider, but my guess is it'll be somewhere between DeepSeek V3 and R1 in terms of benchmarks — R1 trades blows with o1-high, and V3 is somewhat lower — but I'd expect o3-mini to be considerably faster. Despite the blog post saying paid users can access o3-mini today, I don't see it as an option yet in their UI... But IIRC when they announced o3-mini in December they claimed it would be similar to 4o in terms of overall latency, and 4o is much faster than V3/R1 currently.

Re: OpenAI O3-Mini

#137

Earlier quoted context omitted.

They really need someone in marketing. If the model is for technical stuff, then call it the technical model. How is anyone supposed to know what these model names mean? The only page of theirs attempting to explain this is a total disaster. https://platform.openai.com/docs/models

Yes, this $300Bn company generating +$3.4Bn in revenue needs to hire marketing expert. They can begin by sourcing ideas from us here to save their struggling business from total marketing disaster.

[dead]

Re: OpenAI O3-Mini

#139

"oh no DeepSeek copied our product it's not fair" > proceeds to release a product based on DeepSeek ah, alas the hypocrisy...

The thing they previewed back in December before the whole Deepseek kerfuffle this week?

Don't get me wrong, I'm laughing at OpenAI just like everyone else, but if they were really copying Deepseek, they'd be releasing a smaller model distilled from Deepseek API responses, and have it be open source to boot. This is neither

Re: OpenAI O3-Mini

#140

>developer messages looks like finally their threat model has been updated to take into account that the user might be too "unaligned" to be trusted with the ability to provide a system message of their own

...I'm pretty sure they just renamed the key...
Post reply on HN