Live data from Hacker News

OpenAI O3-Mini

openai.com

181–190 of 944 posts

Re: OpenAI O3-Mini

#182

Earlier quoted context omitted.

R1 (70B-distill) itself is very uncensored, will give you full account of tiannanmen square from vague prompts. Asking R1 "what significant things happened in china in 1989" had it volunteering that "the death toll was in the hundreds or thousands and the exact number remains disputed to this day". The only thing that's censored is the web interface.

When asking it about the concept of human rights and the various forms in which it manifests (i.e. demographic equality under the law). I get a mixture of mundane nuance and bizarre answers that Xi Jingping himself could have written. With references to unity and the importance of social harmony over the "freedoms of the few". This tracks when considering that the model was trained on western model outputs and then t…

I definitely am not getting that, perhaps the 671b model is notably worse than the 70b llama distill in this respect. 70b seemed pretty happy to talk about the ethnic cleansing of the Uyghurs in Xinjiang by the CCP and Palestinians in Gaza by Israel, it did some both-sides ing but it generally seemed to provide a balanced-ish viewpoint. At least I think it provided a viewpoint that comports with my best guess of what the average person globally would consider balanced.

Re: OpenAI O3-Mini

#183

Earlier quoted context omitted.

They really need someone in marketing. If the model is for technical stuff, then call it the technical model. How is anyone supposed to know what these model names mean? The only page of theirs attempting to explain this is a total disaster. https://platform.openai.com/docs/models

> How is anyone supposed to know what these model names mean? Normies don't have to know - ChatGPT app focuses UX around capabilities and automatically picks the appropriate model for capabilities requested; you can see which model you're using and change it, but don't need to . As for the techies and self-proclaimed "AI experts" - OpenAI is the leader in the field, and one of the most well-known and talked about tec…

What if you use ASCII 234? Ω (edit: works!)

Re: OpenAI O3-Mini

#184
post #128

Earlier quoted context omitted.

Not really. They’re successful because they created one of the most interesting products in human history, not because they have any idea how to brand it.

If that were the case, they’d be neck and neck with Anthropic and Claude. But ChatGPT has far more market share and name recognition, especially among normies. Branding clearly plays a huge role.

ChatGPT is still benefitting from first mover advantage. Which they’ve leveraged to get to the position they’re at today.

Over time, competitors catch up and first mover advantage melts away.

I wouldn’t attribute OpenAI’s success to any extremely smart marketing moves. I think a big part of their market share grab was simply going (and staying) viral for a long time. Manufacturing virality is notoriously difficult (and based on the usability and poor UI of ChatGPT early versions, it feels like they got lucky in a lot of ways)

Re: OpenAI O3-Mini

#185

I’ll take the China Deluxe instead, actually. I’ve been incredibly pleased with DeepSeek this past week. Wonderful product, I love seeing its brain when it’s thinking.

Being able to see the thinking trace in R1 is so useful, as you can go back and see if it's getting stuck, making a wrong assumption, missing data, etc. To me that makes it materially more useful than the OpenAI reasoning models, which seem impressive, but are much harder to inspect/debug.

the fact that openai hides the reasoning tokens from us to begin with shows that what they are doing behind the scenes isnt all that impressive, and likely easily cloned (r1)

would be nice if they made them visible now

Re: OpenAI O3-Mini

#186
post #67

Earlier quoted context omitted.

OpenAI clearly states that they train on your data https://help.openai.com/en/articles/5722486-how-your-data-is...

By default, we do not train on any inputs or outputs from our products for business users, including ChatGPT Team, ChatGPT Enterprise, and the API. We offer API customers a way to opt-in to share data with us, such as by providing feedback in the Playground, which we then use to improve our models. Unless they explicitly opt-in, organizations are opted out of data-sharing by default. The business bit is confusing, I…

So for posterity, in this subthread we found that OpenAI indeed trains on user data and it isn't something that only DeepSeek does.

Re: OpenAI O3-Mini

#187
post #75

Earlier quoted context omitted.

Inscrutable naming is a proven strategy for muddying the waters.

Salesforce would like a word...

The USB-IF as well. Retroactively changing the name of a previous standard was particularly ridiculous. It's always been USB 3.1 Gen 1 like we've always been at war with Eastasia.

Re: OpenAI O3-Mini

#188
post #50

It looks like a pretty significant increase on SWE-Bench. Although that makes me wonder if there was some formatting or gotcha that was holding the results back before. If this will work for your use case then it could be a huge discount versus o1. Worth trying again if o1-mini couldn't handle the task before. $4/million output tokens versus $60. https://platform.openai.com/docs/pricing I am Tier 5 but I don't believ…

Tier 3 here and already see it on Limits page, so maybe the wait won't be long.

Re: OpenAI O3-Mini

#189

What is the comparison of this versus DeepSeek in terms of good results and cost?

Deepseek is the state of the art right now in terms of performance and output. It's really fast. The way it "explains" how it's thinking is remarkable.

Re: OpenAI O3-Mini

#190
post #50

It looks like a pretty significant increase on SWE-Bench. Although that makes me wonder if there was some formatting or gotcha that was holding the results back before. If this will work for your use case then it could be a huge discount versus o1. Worth trying again if o1-mini couldn't handle the task before. $4/million output tokens versus $60. https://platform.openai.com/docs/pricing I am Tier 5 but I don't believ…

Genuinely curious, What made you choose OpenAI as your preferred api provider? Its always been the least attractive to me.

Until recently they were the only game in town, so maybe they accrued significant spend back then?
Post reply on HN