Live data from Hacker News

OpenAI O3-Mini

openai.com

151–160 of 944 posts

Re: OpenAI O3-Mini

#151
im just glad it looks like o3-mini finally has internet access

the o1 models were already so niche that i never used them, but not being able to search the web made them even more useless

Re: OpenAI O3-Mini

#152
post #125

Earlier quoted context omitted.

You can run it locally, pause it when it thinks wrong and correct it's chain of thought.

Oh wow I did not know and dont have the hardware to run it locally unfortunately

You probably have the hardware to run the smallest distill, it runs even on my ancient laptop. It's not very smart but it still does the CoT and you can have fun editing it.

Re: OpenAI O3-Mini

#153

I’ll take the China Deluxe instead, actually. I’ve been incredibly pleased with DeepSeek this past week. Wonderful product, I love seeing its brain when it’s thinking.

I did a blind test and still prefer Gemini, Claude, and OpenAI to deepseek.

Re: OpenAI O3-Mini

#154

I’ll take the China Deluxe instead, actually. I’ve been incredibly pleased with DeepSeek this past week. Wonderful product, I love seeing its brain when it’s thinking.

Sometimes its thinking is more useful than the actual output.

Re: OpenAI O3-Mini

#157

Earlier quoted context omitted.

The -mini postfix makes perfect sense, probably even clearer than the old "turbo" wording. Naturally, the latest small model may be better than larger older models... but not always and not necessarily in everything. What you'd expect from a -mini model is exactly what is delivered. The non-reasoning line was also pretty straightforward. Newer base models get a larger prefix number and some postfixes like 'o' were ad…

> I wonder if we'll end up with both a 4o and o4... The perplexing thing is that someone has to have said that, right? It has to have been brought up in some meeting when they were brainstorming names that if you have 4o and o1 with the intention of incrementing o1 you'll eventually end up with an o4. Where they really went off the rails was not just bailing when they realized they couldn't use o2. In that moment the…

The obvious solution could be to just keep skipping the even numbers and go to o5.

Re: OpenAI O3-Mini

#158
post #50

It looks like a pretty significant increase on SWE-Bench. Although that makes me wonder if there was some formatting or gotcha that was holding the results back before. If this will work for your use case then it could be a huge discount versus o1. Worth trying again if o1-mini couldn't handle the task before. $4/million output tokens versus $60. https://platform.openai.com/docs/pricing I am Tier 5 but I don't believ…

Genuinely curious, What made you choose OpenAI as your preferred api provider? Its always been the least attractive to me.

Re: OpenAI O3-Mini

#159

> While OpenAI o1 remains our broader general knowledge reasoning model, OpenAI o3-mini provides a specialized alternative for technical domains requiring precision and speed. I feel like this naming scheme is growing a little tired. o1 is for general knowledge reasoning, o3-mini replaces o1-mini but might be more specialized than o1 for certain technical domains...the "o" in "4o" is for "omni" (referring to its mult…

They should be calling it ChatGPT and ChatGPT-mini, with other models hidden behind some sort of advanced mode power user menu. They can roll out major and minor updates by number. The whole point of differentiating between models is to get users to self limit the compute they consume - rate limits make people avoid using the more powerful models, and if they have a bad experience using the less capable models, or if…

"ChatGPT" (chatgpt-4o) is now its own model, distinct from gpt-4o.

As for self-limiting usage by non-power users, they're already doing that: ChatGPT app automatically picks a model depending on what capabilities you invoke. While they provide a limited ability to see and switch the model in use, they're clearly expecting regular users not to care, and design their app around that.

Re: OpenAI O3-Mini

#160
post #128

Earlier quoted context omitted.

Not really. They’re successful because they created one of the most interesting products in human history, not because they have any idea how to brand it.

If that were the case, they’d be neck and neck with Anthropic and Claude. But ChatGPT has far more market share and name recognition, especially among normies. Branding clearly plays a huge role.

That's first mover advantage.
Post reply on HN