Earlier quoted context omitted.
But then there will be no comments to summarize.
Our digital twins will write the comments. They will be us, but with none of our flaws. They will never experience the shame of posting a dumb joke, getting flamed, and then deleting it, for they will have tested all ideas to prevent such an oversight. They will never experience the satisfaction-turned-to-puzzlement of posting an expertly crafted, well-researched comment that took 2 hours of the workday to draft - on…
OpenAI O3-Mini
921–930 of 944 posts
Re: OpenAI O3-Mini
#922Re: OpenAI O3-Mini
#923Earlier quoted context omitted.
Sounds about right, as we are post-dead internet in public places. There was a thread about the US tariffs on Canada I was reading on a stock investment subreddit. The whole page was full of people complaining about Elon Musk, Donald Trump, "Buy Canadian" comments, moralizing about Alberta's conservative government and other unrelated noise. None of this was related to the topic; stocks and funds that seemed well-pla…
> zoomer internet church Stealing this.
Re: OpenAI O3-Mini
#924o3-mini produced volumes of code more quickly and more of it, but Claude still had greater insight in to the problem and decoded the format to a noticeably greater degree.
Whereas 03-mini quickly got to a certain point, it wasn't long before it was obvious it wasn't really going any further - like it's big cousin, but in it's own way, it was lazy and forgetful, seeming at times more interested in telling me what I might try than actually trying itself.
Interestingly, even when I gave it a copy of Claude's code it still wasn't able to get to the same depth of understanding.
Re: OpenAI O3-Mini
#925Re: OpenAI O3-Mini
#926Earlier quoted context omitted.
The problem with claims like these that models are not doing “actual reasoning” is that they are often hot takes and not thought through very well. For example, since reasoning doesn’t yet have any consensus definition that can be applied as a yes/no test - you have to explain what you specifically mean by it, or else the claim is hollow. Clarify your definition, give a concrete example under that definition of somet…
Explain this to me please: we don't have any consensus definition of _mathematics_ that can be applied as a yes/no test. Does that mean we don't know how to do mathematics, or that we don't know whether something, is, or, more importantly, isn't mathematics? For example, if I throw a bunch of sticks in the air and look at their patterns to divine the future- can I call that "mathematics" just because nobody has a "co…
Re: OpenAI O3-Mini
#927Earlier quoted context omitted.
I haven’t tried o3, but one issue I struggle with in large context analysis tasks is the LLMs are never thorough. In a task like this thread summarization, I typically need to break the document down and loop through chunks to ensure it actually “reads” everything. I might have had to recurse into individual conversations with some small max-depth and leaf count and run inference on each, and then have some aggregati…
o1-pro is incredibly good at this. You'll be amazed
Re: OpenAI O3-Mini
#928Re: OpenAI O3-Mini
#929Earlier quoted context omitted.
You do you, but hivemind thinking is a real thing. I have seen highly upvoted comments seemingly "debunk" an article where on closer examination it becomes clear they actually didn't read the article either. It quickly becomes this weird bubble of people just acting on what everything "thinks" the content is about without ever having looked at the content. I get that is easier, but intellectually you are doing yourse…
That is the biggest problem on this website, people want to feel smart by 'debunking' things they don't really understand. It leads to a lot of contrarian views with poor signal to noise ratio, especially when the topic is slightly outside of average user experience (midwit programming)
Re: OpenAI O3-Mini
#930Earlier quoted context omitted.
Explain this to me please: we don't have any consensus definition of _mathematics_ that can be applied as a yes/no test. Does that mean we don't know how to do mathematics, or that we don't know whether something, is, or, more importantly, isn't mathematics? For example, if I throw a bunch of sticks in the air and look at their patterns to divine the future- can I call that "mathematics" just because nobody has a "co…
Yeah sure there’s lots of research on reasoning. The papers I’ve seen that make claims about it are usually pretty precise about what it means in the context of that work and that specific claim, at least in the hard sciences listed.