Live data from Hacker News

OpenAI O3-Mini

openai.com

771–780 of 944 posts

Re: OpenAI O3-Mini

#771
post #9

Earlier quoted context omitted.

What about "o1 Pro mode". Is that just o1 but with more reasoning time, like this new o3-mini's different amount of reasoning options?

I have been paying $200 per month for 01-pro mode and I am very disappointed right now because they have completely replaced the model today. It used to think for 1-5 minutes and deliver an unbelievably useful one-shot answer. Now, it only thinks for 7 seconds just like the 03-mini model and I can't tell the difference in the answers. I hope this is just a day 1 implementation bug but I suspect they have just decided…

This is why I use the Azure-hosted versions (disclosure: I’m an MS FTE, but I use all sorts of 3rd party models for my own projects) - I _know_ which version is behind each endpoint and when they will be replaced (you can also pin versions within a support window that varies according to model), so I don’t have to rework all my prompts and throw work away at the drop of a hat.

Re: OpenAI O3-Mini

#772
post #682

Earlier quoted context omitted.

I put one of my own blog posts through NotebookLM soon after it became available, it hallucinated content I didn't write and missed out things I had written. Nice TTS, but otherwise I found it unimpressive.

I’m not talking about the TTS and podcast creation. I’m talking about just asking questions where it gives you the answer with citations.

Given what it got wrong was in the LLM part, that is a distinction without a difference.

Re: OpenAI O3-Mini

#773

Earlier quoted context omitted.

Amazon is already forcing this pattern on mobile users not logged in. If you want to see the reviews, all you get is an AI summary, the star rating, maybe 1 or 2 reviews, then you have to log in to see more.

The cure is to stop using Amazon.

It has already caused me to not further investigate purchases. Similar to yelp, if I have to log in to see things, you're dead to me.

Re: OpenAI O3-Mini

#775

Earlier quoted context omitted.

On both HN & Reddit, I find the comments more informative and less frustrating than reading the article usually. But I guess YMMV.

They definitely used to be, but haven't been much good for years. At least 8 years in the case of reddit, maybe 3 in the case of hackernews. Though at this point it's a habit I cannot quite bring myself to break...

It's worse, but there's nowhere better than here imo.

Except very niche topics maybe

Re: OpenAI O3-Mini

#776
post #669
post #306

I used o3-mini to summarize this thread so far. Here's the result: https://gist.github.com/simonw/09e5922be0cbb85894cf05e6d75ae... For 18,936 input, 2,905 output it cost 3.3612 cents. Here's the script I used to do it: https://til.simonwillison.net/llms/claude-hacker-news-themes...

Currently on the internet people skip the article and go straight to the comments. Soon people will skip the comments and go striaght to an AI summary reading neither the original article nor the comments.

I won't. I'm looking for real life human voices.

Re: OpenAI O3-Mini

#777
post #753

Earlier quoted context omitted.

A 12% margin is literally the opposite of a coin flip. Unless you have a really bad coin.

I wasn't expecting for my comment to be red so literally but ok. We're talking about the most cost-efficient model, the competition here is on price, not on a 12% incremental performance (which would make sense for the high end model). To my knowledge deepseek is the cheaper service which is what matters on the low-end (unless the increase in performance was in such magnitude that the extra-charge would be worth the…

What does deepseek have to do with a comparison between o1-mini and o3-mini?

Re: OpenAI O3-Mini

#778
post #669

Earlier quoted context omitted.

Currently on the internet people skip the article and go straight to the comments. Soon people will skip the comments and go striaght to an AI summary reading neither the original article nor the comments.

But then there will be no comments to summarize.

I went to this thought too but then I remembered the 90-9-1 rule. The AI summary is for some portion of the 90. The 9 are still going to comment. What they comment on and how they generate the comments might change though.

Re: OpenAI O3-Mini

#779

Earlier quoted context omitted.

I mean, do you think this is awful ? https://pastebin.com/Ja14mt6L

If R1 one-shotted this then I revise my opinion somewhat. This doesn't give me a gut "this is awful" emotional reaction (although I don't think it's good - it's pretty cliche, and I found my eyes glossing over pretty quickly). I was somewhat turned off of DeepSeek (the first few questions I gave it, it returned 100% hallucinated answers). But maybe I'll have to look into it more, thanks.

I regenerated a page a few times but yeah, I gave no other instructions besides that. Also I approached things page by page. That was 3 pages.

It's cliche but this was the prompt:

I want you to write the first book of a fantasy series. The novel will be over 450 pages long with 30 chapters. Each chapter should have between 15 to 18 pages. Write the first page of the first chapter of this novel. Do not introduce the elements of the synopsis or worldbuilding and story details too quickly. Weave in the world, characters, and plot naturally. Pace it out properly. That means that several elements of the story may not come into light for several chapters.

I had a lot of success with it coming up with decidedly not cliche world building elements after I arranged a sort of interview style interrogation (It asked me questions about what I was looking for generally and generated world building elements along the way).

However, once you start giving a lot of information about the world etc in the prompt as well then the pacing gets weird.

Re: OpenAI O3-Mini

#780
post #669

Earlier quoted context omitted.

Currently on the internet people skip the article and go straight to the comments. Soon people will skip the comments and go striaght to an AI summary reading neither the original article nor the comments.

But then there will be no comments to summarize.

Our digital twins will write the comments. They will be us, but with none of our flaws. They will never experience the shame of posting a dumb joke, getting flamed, and then deleting it, for they will have tested all ideas to prevent such an oversight. They will never experience the satisfaction-turned-to-puzzlement of posting an expertly crafted, well-researched comment that took 2 hours of the workday to draft - only to receive one upvote, for their research will be instantaneous and their outputs efficient. Of course they will never need to humbly reply, 'Ah, I missed that, good catch!' to a child comment indicating the entire premise of their question would be answered with a simple reading of the linked article - for they will have deeply and instantly read the article. Yes, our digital twins will be us, but better - and we will finally be free to play in the mud.
Post reply on HN