Earlier quoted context omitted.
Looks more like a direct competitor to Aider.
Where do I find out more about Aider ?
OpenAI o3 and o4-mini
201–210 of 527 posts
Re: OpenAI o3 and o4-mini
#202Earlier quoted context omitted.
[flagged]
The closest Elon ever came to anything Hague-worthy is allowing Starlink to be used in Ukrainian attacks on Russian civilian infrastructure. I don't think the Hague would be interested in anything like that. And if his life is worthless, then what would you say about your own? Nonetheless, I commend you on your complete lack of hinges. /s
Re: OpenAI o3 and o4-mini
#203So at this point OpenAI has 6 reasoning models, 4 flagship chat models, and 7 cost optimized models. So that's 17 models in total and that's not even counting their older models and more specialized ones. Compare this with Anthropic that has 7 models in total and 2 main ones that they promote. This is just getting to be a bit much, seems like they are trying to cover for the fact that they haven't actually done much.…
Or perhaps they're trying to make some important customers happy by showing movement on areas the customers care about. Subjectively, customers get locked in by feeling they have the inside track, and these small tweaks prove that. Objectively, the small change might make a real difference to the customer's use case.
Similarly, it's important to force development teams to actually ship, and shipping more frequently reduces risk, so this could reflect internal discipline.
As for media buzz, OpenAI is probably trying to tamp that down; they have plenty of first-mover advantage. More puffery just makes their competitors seem more important, and the risk to their reputation of a flop is a lot larger than the reward of the next increment.
As for "a bit much", before 2023 I was thinking I could meaningfully track progress and trade-off's in selecting tech, but now the cat is not only out of the bag, it's had more litters than I can count. So, yeah - a bit much!
Re: OpenAI o3 and o4-mini
#204Earlier quoted context omitted.
Im old enough to remember the mystery and hype before o*/o1/strawberry that was supposed to be essentially AGI. We had serious news outlets write about senior people at OpenAI quitting because o1 was SkyNet Now we're up to o4, AGI is still not even in near site (depending on your definition, I know). And OpenAI is up to about 5000 employees. I'd think even before AGI a new model would be able to cover for at least 45…
> Im old enough to remember the mystery and hype before o*/o1/strawberry So at least two years old?
Re: OpenAI o3 and o4-mini
#205Earlier quoted context omitted.
Looks more like a direct competitor to Aider.
Where do I find out more about Aider ?
Re: OpenAI o3 and o4-mini
#206Earlier quoted context omitted.
Main advantage over Sonnet is Gemini 2.5 doesn't try to make a bunch of unrelated changes like it's rewriting my project from scratch.
This was incredibly irritating at first, though over time I've learned to appreciate this "extra credit" work. It can be fun to see what Claude thinks I can do better, or should add in addition to whatever feature I just asked for. Especially when it comes to UI work, Claude actually has some pretty cool ideas. If I'm using Claude through Copilot where it's "free" I'll let it do its thing and just roll back to the la…
Too bad Microsoft is widely limiting this -- have you seen their pricing changes?
I also feel like they nerfed their models, or reduced context window again.
Re: OpenAI o3 and o4-mini
#207Earlier quoted context omitted.
As someone who doesn't use anything OpenAI (for all the reasons), I have to agree with the GP. It's all baffling. Why is there an o3-mini and an o4-mini? Why on earth are there so many models? Once you get to this point you're putting the paradox of choice on the user - I used to use a particular brand toothpaste for years until it got to the point where I'd be in the supermarket looking at a wall of toothpaste all b…
(I work at OpenAI.) In ChatGPT, o4-mini is replacing o3-mini. It's a straight 1-to-1 upgrade. In the API, o4-mini is a new model option. We continue to support o3-mini so that anyone who built a product atop o3-mini can continue to get stable behavior. By offering both, developers can test both and switch when they like. The alternative would be to risk breaking production apps whenever we launch a new model and shut…
I have not seen any sort of "If you're using X.122, upgrade to X.123, before 202X. If you're using X.120, upgrade to anything before April 2026, because the model will no longer be available on that date." ... Like all operating systems and hardware manufacturers have been doing for decades.
Side note, it's amusing that stable behavior is only available on a particular model with a sufficiently low temperature setting. As near-AGI shouldn't these models be smart enough to maintain consistency or improvement from version to version?
Re: OpenAI o3 and o4-mini
#208Re: OpenAI o3 and o4-mini
#209Still a knowledge cutoff of August 2023. That is a significant bottleneck to devs using it for AI stuff.
Re: OpenAI o3 and o4-mini
#210As a consumer, it is so exhausting keeping up with what model I should or can be using for the task I want to accomplish.
Gemini 2.5 Pro for every single task was the meta until this release. Will have to reassess now.