Live data from Hacker News

OpenAI O3-Mini

openai.com

211–220 of 944 posts

Re: OpenAI O3-Mini

#211
> Testers preferred o3-mini's responses to o1-mini 56% of the time

I hope by this they don't mean me, when I'm asked 'which of these two responses do you prefer'.

They're both 2,000 words, and I asked a question because I have something to do. I'm not reading them both; I'm usually just selecting the one that answered first.

That prompt is pointless. Perhaps as evidenced by the essentially 50% response rate: it's a coin-flip.

Re: OpenAI O3-Mini

#212
post #43

Haven't used openai in a bit -- whyyy did they change "system" role (now basically an industry-wide standard) to "developer"? That seems pointlessly disruptive.

They mention in the model card, it's so that they can have a separate "system" role that the user can't change, and they trained the model to prioritise it over the "developer" role, to combat "jailbreaks". Thank God for DeepSeek.

They should have just created something above system and left as it was.

Re: OpenAI O3-Mini

#213
post #12
post #3

So far, it seems like this is the hierarchy o1 > GPT-4o > o3-mini > o1-mini > GPT-4o-mini o3 mini system card: https://cdn.openai.com/o3-mini-system-card.pdf

I think OpenAI really needs to rethink its product naming, especially now that they have a portfolio where there's no such clear hierarchy, but they have a place along different axis (speed, cost, reasoning, capabilities, etc). Your summary attempt e.g. also misses o3-mini vs o3-mini-high. Lots of trade-ofs.

They can still do models o3o, oo3 and 3oo. Mini-o3o-high, not to be confused with mini-O3o-high (the first o is capital).

Re: OpenAI O3-Mini

#214

The real heated contest here amongst the top AI labs is to see who can come up with the most confusing product names.

It's nice to see Google finally having competition in a space it used to really dominate (though they definitely still are holding their own with all the Gemini naming). I feel like it takes real effort to have product names be this confusing and capricious

Re: OpenAI O3-Mini

#215

I’ll take the China Deluxe instead, actually. I’ve been incredibly pleased with DeepSeek this past week. Wonderful product, I love seeing its brain when it’s thinking.

Using R1 with Perplexity has impressed me in a way that none of the previous models have, and I can't even figure out if it's actually R1, seems likely that its a 70B-llama distillation since that's what AWS offers on Bedrock but from what I can find Perplexity does have their own H100 cluster through Amazon so it's feasible they could be hosting the real thing? But I feel like they would brag about that achievement…

I played with their model, and I want able to make him follow any instructions, it looked like it just reads first message and ignore rest of the conversation. not sure if they is bug with oupenrouter or model, but I was highly disappointed.

from way how it thinks/responds looks like it's one of destinations , likely llama one I also suspect that many of free/cheap providers also serve llama instead of real R1

Re: OpenAI O3-Mini

#216

The real heated contest here amongst the top AI labs is to see who can come up with the most confusing product names.

Someone dropped the ball with Phi models. There is clearly an opportunity for XP and Ultimate and X/S editions.

I really think a "OpenAI Me" is what's needed.

Re: OpenAI O3-Mini

#218

The real heated contest here amongst the top AI labs is to see who can come up with the most confusing product names.

Someone dropped the ball with Phi models. There is clearly an opportunity for XP and Ultimate and X/S editions.

Personally waiting for the ME model. Should be great at jokes and humor.

Re: OpenAI O3-Mini

#219
First thing I noticed on API and Chat for it is THIS THING IS FAST. That alone makes it a huge upgrade to o1-pro (not really comparable I know, just saying). Can't imagine how much I'll get done with this type of speed.

Re: OpenAI O3-Mini

#220
post #117

Earlier quoted context omitted.

Yes, this $300Bn company generating +$3.4Bn in revenue needs to hire marketing expert. They can begin by sourcing ideas from us here to save their struggling business from total marketing disaster.

At the least they should care more about UX. I have no idea how to restore the sidebar on chatgpt on desktop lol

Click the 'open sidebar' icon in the top left corner of the screen.
Post reply on HN