Live data from Hacker News

OpenAI o3-pro

help.openai.com

71–80 of 209 posts

Re: OpenAI o3-pro

#71
post #4

So, upgrade to Teams and pay the $50? Plus more usage of o3. Seems like it might be a shot at the $100 claude max?

What do you mean with "pay the $50"? Also, does anybody know what limits o3-pro has under the team plan? I don't see it available in the model picker at all (on team).

I believe teams is $25/user with a 2 user minimum.

Re: OpenAI o3-pro

#72
post #71

Earlier quoted context omitted.

What do you mean with "pay the $50"? Also, does anybody know what limits o3-pro has under the team plan? I don't see it available in the model picker at all (on team).

I believe teams is $25/user with a 2 user minimum.

Ah, thanks for explaining!

Re: OpenAI o3-pro

#73

Earlier quoted context omitted.

I just can’t believe nobody at the company has enough courage to tell their leadership that their naming scheme is completely stupid and insane. Four is greater than three, and so four should be better than three. The point of a name is to describe something so that you don’t confuse your users, not to be cute.

The reason their naming scheme is so bad is because their initial attempts at GPT-5 failed in training. It was supposed to be done ~1 year ago. Because they'd promised that GPT-5 would be vastly more intelligent than GPT-4, they couldn't just name any random model "GPT-5", so they suddenly had to start naming things differently. So now there's GPT-4.5, GPT-4.1, the o-series, ...

Surely there's a less stupid way than naming two very different models o4 and 4o.

Re: OpenAI o3-pro

#75
post #36

Earlier quoted context omitted.

That's definitely not the case here. The new o3-pro is slow - it took two minutes just to draw me an SVG of a pelican riding a bicycle. o3-preview was much faster than that. https://simonwillison.net/2025/Jun/10/o3-pro/

Would you say this is the best cycling pelican to date? I don't remember any of the others looking better than this. Of course by now it'll be in-distribution. Time for a new benchmark...

I like the Gemini 2.5 Pro ones a little more: https://simonwillison.net/2025/Jun/6/six-months-in-llms/#ai-...

Re: OpenAI o3-pro

#76
post #36
post #28

The guys in the other thread who said that OpenAI might have quantized o3 and that's how they reduced the price might be right. This o3-pro might be the actual o3-preview from the beginning and the o3 might be just a quantized version. I wish someone benchmarks all of these models to check for drops in quality.

That's definitely not the case here. The new o3-pro is slow - it took two minutes just to draw me an SVG of a pelican riding a bicycle. o3-preview was much faster than that. https://simonwillison.net/2025/Jun/10/o3-pro/

Wow! pelican benchmark is now saturated

Re: OpenAI o3-pro

#77
post #15

Earlier quoted context omitted.

free users don't have this model selector, and probably don't care which model they get so 4o is good enough. paid users at 20$/month get more models which are better, like o3. paid users at 200$/month get the best models that are also costing OpenAI the most money, like o3-pro. I think they plan to unify them with GPT-5.

That doesn't help much when we're asymptotically approaching GPT-5. We're probably going to be at GPT-4.9999 soon.

Not necessarily true. GPT-4.1 was released after GPT-4.5-preview. Next model might be GPT-3.7.

Re: OpenAI o3-pro

#78

I'm really hoping GPT5 is a larger jump in metrics than the last several releases we've seen like Claude3.5 - Claude4 or o3-mini-high to o3-pro. Although I will preface that with the fact I've been building agents for about a year now and despite the benchmarks only showing slight improvement, I have seen that each new generation feels actively better at exactly the same tasks I gave the previous generation. It would…

I'm seeing big advances that arent shown in the benchmarks, I can simply build software now that I couldnt build before. The level of complexity that I can manage and deliver is higher.

mind telling examples?

Re: OpenAI o3-pro

#79

I'm really hoping GPT5 is a larger jump in metrics than the last several releases we've seen like Claude3.5 - Claude4 or o3-mini-high to o3-pro. Although I will preface that with the fact I've been building agents for about a year now and despite the benchmarks only showing slight improvement, I have seen that each new generation feels actively better at exactly the same tasks I gave the previous generation. It would…

It's hard to be 100% certain, but I am 90% certain that the benchmarks leveling off, at this point, should tell us that we are really quite dumb and simply not good very good at either using or evaluating the technology (yet?).

either that or the improvements aren't as large as before.

Re: OpenAI o3-pro

#80
post #13

here's a nice user review we published: https://www.latent.space/p/o3-pro sama's highlight[0]: > "The plan o3 gave us was plausible, reasonable; but the plan o3 Pro gave us was specific and rooted enough that it actually changed how we are thinking about our future." I kept nudging the team to go the whole way to just let o3 be their CEO but they didn't bite yet haha 0: https://x.com/sama/status/1932533208366608568

if o3 is so good why aren't they using it to replace management?
Post reply on HN