Live data from Hacker News

OpenAI o3 and o4-mini

openai.com

301–310 of 527 posts

Re: OpenAI o3 and o4-mini

#301

Earlier quoted context omitted.

I was a major contributor of Flake. What in particular is so idiotic in your opinion?

I use flakes a lot and I think both flakes and the Nix language are beyond comprehension. Try searching duckduckgo or google for “what is nix flakes” or “nix flake schema” and take an honest read at the results. Insanely complicated and confusing answers, multiple different seemingly-canonical sources of information. Then go look at some flakes for common projects; the almost necessary usage of things like flake-comp…

I apologize. It was my Haskell life period.

Re: OpenAI o3 and o4-mini

#303

Doesn't achieving AGI mean the beginning of the end of humanity's current economic model? I'm not sure I understand the presumption by many that achieving AGI is just another step in some company's offering.

No you see because everyone will become agi engineers actually that makes sense and is going to happen

Re: OpenAI o3 and o4-mini

#304

So at this point OpenAI has 6 reasoning models, 4 flagship chat models, and 7 cost optimized models. So that's 17 models in total and that's not even counting their older models and more specialized ones. Compare this with Anthropic that has 7 models in total and 2 main ones that they promote. This is just getting to be a bit much, seems like they are trying to cover for the fact that they haven't actually done much.…

The old Chinese strategy of having 7343 different phone models with almost the same specs to confuse the customer better

not only that. filling search lists on eBay with your products is old sellers' tactics. Try to search for used Dell workstation or server and you will see pages and pages from the same seller.

Re: OpenAI o3 and o4-mini

#306

Earlier quoted context omitted.

Sonnet and Gemini saw fairly substantial perf increases recenly

Love Sonnet but 3.7 is not obviously an improvement over 3.5 in my real world usage. Gemini 2.5 pro is great, has replaced most others for me (Grok I use for things that require realtime answers)

It does a lot better on philosophy questions.

Re: OpenAI o3 and o4-mini

#307
post #298

Ok, I’m a bit underwhelmed. I’ve asked it a fairly technical question, about a very niche topic (Final Fantasy VII reverse engineering): https://chatgpt.com/share/68001766-92c8-8004-908f-fb185b7549... With right knowledge and web searches one can answer this question in a matter of minutes at most. The model fumbled around modding forums and other sites and did manage to find some good information but then started to…

It can imitate its creator. We reached AGI.

Re: OpenAI o3 and o4-mini

#308

Earlier quoted context omitted.

I use flakes a lot and I think both flakes and the Nix language are beyond comprehension. Try searching duckduckgo or google for “what is nix flakes” or “nix flake schema” and take an honest read at the results. Insanely complicated and confusing answers, multiple different seemingly-canonical sources of information. Then go look at some flakes for common projects; the almost necessary usage of things like flake-comp…

I apologize. It was my Haskell life period.

I forgive you as I hope you forgive me. Flakes are certainly much better than Nix without them, and they’ve saved me much more time than they’ve cost me.

Re: OpenAI o3 and o4-mini

#309

Earlier quoted context omitted.

Do you have a 200k context window? I don't. Most humans can only keep 6 or 7 things in short term memory. Beyond those 6 or 7 you are pulling data from your latent space, or replacing of the short term slots with new content.

But context windows for LLMs include all the “long term memory” things you’re excluding from humans

Long term memory in an LLM is its weights.

Re: OpenAI o3 and o4-mini

#310
post #120

Earlier quoted context omitted.

Im old enough to remember the mystery and hype before o*/o1/strawberry that was supposed to be essentially AGI. We had serious news outlets write about senior people at OpenAI quitting because o1 was SkyNet Now we're up to o4, AGI is still not even in near site (depending on your definition, I know). And OpenAI is up to about 5000 employees. I'd think even before AGI a new model would be able to cover for at least 45…

Yeah, I don't know exactly what at an AGI model will look like, but I think it would have more than 200k context window.

I'm not quite AGI, but I work quite adequately with a much, much smaller memory. Maybe AGI just needs to know how to use other computers and work with storage a bit better.
Post reply on HN