Live data from Hacker News

OpenAI declares 'code red' as Google catches up in AI race

theverge.com

371–380 of 960 posts

Re: OpenAI declares 'code red' as Google catches up in AI race

#371

Heard all the news how Gemini 3 is passing everyone on benchmarks, so quickly tested and still find it a far cry from ChatGPT in real world use when testing questions on both platforms. But importantly the ChatGPT app experience at least for iPhone/Mac users is drastically superior vs Google which feels very Google still. So Gemini would have to be drastically better answer wise than ChatGPT to lure users from a bett…

they're deep into a redesign of the gemini app, idk when it will be released or if its going to be good, but at least they agree with you and are putting significant resources into fixing it.

I did notice a bug on the iPhone, even with app background refresh, if the phone shuts off the screen, a prompt that was processing stalls out.

Re: OpenAI declares 'code red' as Google catches up in AI race

#372

Earlier quoted context omitted.

Every so often I try out a GPT model for coding again, and manage to get tricked by the very sparse conversation style into thinking it's great for a couple of days (when it says nothing and then finishes producing code with a 'I did x, y and z' with no stupid 'you're absolutely' right sucking up and it works, it feels very good). But I always realize it's just smoke and mirrors - the actual quality of the code and t…

Can you give some concrete example of programming problem task GPT fails to solve? Interested, because I’ve been getting pretty good results with different tasks using the Codex.

Completely failed for me running the code it changed in a docker container i keep running. Claude did it flawlessly. It absolutely rocks at code reviews but ir‘s terrible in comparison generating code

Re: OpenAI declares 'code red' as Google catches up in AI race

#373

I've seen a rumor going around that OpenAI hasn't had a successful pre-training run since mid 2024. This seemed insane to me but if you give ChatGPT 5.1 a query about current events and instruct it not to use the internet it will tell you its knowledge cutoff is June 2024. Not sure if maybe that's just the smaller model or what. But I don't think it's a good sign to get that from any frontier model today, that's 18 m…

Don’t forget SemiAnalysis’s founder Dylan Patel is supposedly roommates with Anthropics RL tech lead Sholto..

Re: OpenAI declares 'code red' as Google catches up in AI race

#374

Earlier quoted context omitted.

I am not a lawyer, but it is possible he can say whatever he wants without consequences to public because OAI is not a public company.

Kind of, but there are limits. The investors still have LPs who aren’t going to be happy if things get messy. Things can still get really ugly even for a private company.

Most of the credit being throwing around isn't coming from traditional banking companies, mostly private credit being utilized.

Private credit isn't really unregulated.

If you're interested in learning more I believe Matt Stoller has written a few articles about the private credit markets.

Re: OpenAI declares 'code red' as Google catches up in AI race

#375
post #327

Earlier quoted context omitted.

Could say that about any AI company that isn’t at the top as well

You can say it about the AI companies, but Google or Microsoft are far from AI companies.

That's a good point. Google was sleeping on AI and wasn't able to come up with a product before OpenAI and they only scrambled to come out with something when OpenAi became all the rage. Big companies are hard to budge and move in a new direction.

Re: OpenAI declares 'code red' as Google catches up in AI race

#376

The way I've experienced "Code Red" is mostly as a euphemism for "on-going company-wide lack of focus" and a band-aid for mid-level management having absolutely no clue how to meaningfully make progress, upper management panicking, and ultimately putting engineers and ICs on the spot to bear the brunt of that organizational mess. Interestingly enough, apart from Google, I've never seen an organization take the actual…

"Code Red" if implemented correctly should provide a single priority for the company. Engineers will be moved to the most important project(s).

There should already be a single priority for a company...

Why is the bar so low for the billionaire magnate fuck ups? Might as well implement workplace democracy and be done with it, it can't be any worse for the company and at least the workers understand what needs to be done.

Re: OpenAI declares 'code red' as Google catches up in AI race

#377

ChatGPT seems like a huge distraction for OpenAI if their goal is transformative AI IMO: the largest value creation from AGI won’t come from building a better shopping or travel assistant. The real pot of gold is in workflow / labor automation but obviously they can’t admit that openly.

That boat sailed a long time ago

Re: OpenAI declares 'code red' as Google catches up in AI race

#378

I've seen a rumor going around that OpenAI hasn't had a successful pre-training run since mid 2024. This seemed insane to me but if you give ChatGPT 5.1 a query about current events and instruct it not to use the internet it will tell you its knowledge cutoff is June 2024. Not sure if maybe that's just the smaller model or what. But I don't think it's a good sign to get that from any frontier model today, that's 18 m…

It has no idea what it's own knowledge cutoff is.

Re: OpenAI declares 'code red' as Google catches up in AI race

#379

I've seen a rumor going around that OpenAI hasn't had a successful pre-training run since mid 2024. This seemed insane to me but if you give ChatGPT 5.1 a query about current events and instruct it not to use the internet it will tell you its knowledge cutoff is June 2024. Not sure if maybe that's just the smaller model or what. But I don't think it's a good sign to get that from any frontier model today, that's 18 m…

Every so often I try out a GPT model for coding again, and manage to get tricked by the very sparse conversation style into thinking it's great for a couple of days (when it says nothing and then finishes producing code with a 'I did x, y and z' with no stupid 'you're absolutely' right sucking up and it works, it feels very good). But I always realize it's just smoke and mirrors - the actual quality of the code and t…

On the contrary, I cannot use the top Gemini and Claude models because their outputs are so out place and hard to integrate with my code bases. The GPT 5 models integrate with my code base's existing patterns seamlessly.
Post reply on HN