Live data from Hacker News

OpenAI O3-Mini

openai.com

921–930 of 944 posts

Re: OpenAI O3-Mini

#921
post #780

Earlier quoted context omitted.

But then there will be no comments to summarize.

Our digital twins will write the comments. They will be us, but with none of our flaws. They will never experience the shame of posting a dumb joke, getting flamed, and then deleting it, for they will have tested all ideas to prevent such an oversight. They will never experience the satisfaction-turned-to-puzzlement of posting an expertly crafted, well-researched comment that took 2 hours of the workday to draft - on…

Cripes [1].

[1] https://qntm.org/perso

Re: OpenAI O3-Mini

#923
post #848

Earlier quoted context omitted.

Sounds about right, as we are post-dead internet in public places. There was a thread about the US tariffs on Canada I was reading on a stock investment subreddit. The whole page was full of people complaining about Elon Musk, Donald Trump, "Buy Canadian" comments, moralizing about Alberta's conservative government and other unrelated noise. None of this was related to the topic; stocks and funds that seemed well-pla…

> zoomer internet church Stealing this.

Please do. Don't forget to appreciate the mantra-like repetition of Brian Taylor Cohen and Qasim Rashid Esq memes.

Re: OpenAI O3-Mini

#924
I tried to get it to build me a slightly challenging app to break out data from a fairly obscure file format for some PLC code, after having tried with Claude.

o3-mini produced volumes of code more quickly and more of it, but Claude still had greater insight in to the problem and decoded the format to a noticeably greater degree.

Whereas 03-mini quickly got to a certain point, it wasn't long before it was obvious it wasn't really going any further - like it's big cousin, but in it's own way, it was lazy and forgetful, seeming at times more interested in telling me what I might try than actually trying itself.

Interestingly, even when I gave it a copy of Claude's code it still wasn't able to get to the same depth of understanding.

Re: OpenAI O3-Mini

#926

Earlier quoted context omitted.

The problem with claims like these that models are not doing “actual reasoning” is that they are often hot takes and not thought through very well. For example, since reasoning doesn’t yet have any consensus definition that can be applied as a yes/no test - you have to explain what you specifically mean by it, or else the claim is hollow. Clarify your definition, give a concrete example under that definition of somet…

Explain this to me please: we don't have any consensus definition of _mathematics_ that can be applied as a yes/no test. Does that mean we don't know how to do mathematics, or that we don't know whether something, is, or, more importantly, isn't mathematics? For example, if I throw a bunch of sticks in the air and look at their patterns to divine the future- can I call that "mathematics" just because nobody has a "co…

Yeah sure there’s lots of research on reasoning. The papers I’ve seen that make claims about it are usually pretty precise about what it means in the context of that work and that specific claim, at least in the hard sciences listed.

Re: OpenAI O3-Mini

#927
post #513

Earlier quoted context omitted.

I haven’t tried o3, but one issue I struggle with in large context analysis tasks is the LLMs are never thorough. In a task like this thread summarization, I typically need to break the document down and loop through chunks to ensure it actually “reads” everything. I might have had to recurse into individual conversations with some small max-depth and leaf count and run inference on each, and then have some aggregati…

o1-pro is incredibly good at this. You'll be amazed

[dead]

Re: OpenAI O3-Mini

#929
post #737

Earlier quoted context omitted.

You do you, but hivemind thinking is a real thing. I have seen highly upvoted comments seemingly "debunk" an article where on closer examination it becomes clear they actually didn't read the article either. It quickly becomes this weird bubble of people just acting on what everything "thinks" the content is about without ever having looked at the content. I get that is easier, but intellectually you are doing yourse…

That is the biggest problem on this website, people want to feel smart by 'debunking' things they don't really understand. It leads to a lot of contrarian views with poor signal to noise ratio, especially when the topic is slightly outside of average user experience (midwit programming)

not to mention walls of text as people argue back and fourth

Re: OpenAI O3-Mini

#930

Earlier quoted context omitted.

Explain this to me please: we don't have any consensus definition of _mathematics_ that can be applied as a yes/no test. Does that mean we don't know how to do mathematics, or that we don't know whether something, is, or, more importantly, isn't mathematics? For example, if I throw a bunch of sticks in the air and look at their patterns to divine the future- can I call that "mathematics" just because nobody has a "co…

Yeah sure there’s lots of research on reasoning. The papers I’ve seen that make claims about it are usually pretty precise about what it means in the context of that work and that specific claim, at least in the hard sciences listed.

I'm complaining because I haven't seen any such papers. Which ones do you have in mind?
Post reply on HN