Live data from Hacker News

OpenAI declares 'code red' as Google catches up in AI race

theverge.com

161–170 of 960 posts

Re: OpenAI declares 'code red' as Google catches up in AI race

#161
post #145

Earlier quoted context omitted.

That's really fascinating. Every real world use case I've tried on Gemini (especially math-related) absolutely slaughtered the performance of ChatGPT in speed and quality, not even close. As an Android user, the Gemini app is also far superior, since the ChatGPT app still doesn't properly display math equations, among plenty of other bugs.

What do you mean? It renders LaTex fine on Android.

Some LaTeX, but not all, especially for larger equations. I will admit it has gotten a lot better in recent updates, since it seemed thoroughly broken for quite a while in its early days.

Re: OpenAI declares 'code red' as Google catches up in AI race

#162
post #109

Heard all the news how Gemini 3 is passing everyone on benchmarks, so quickly tested and still find it a far cry from ChatGPT in real world use when testing questions on both platforms. But importantly the ChatGPT app experience at least for iPhone/Mac users is drastically superior vs Google which feels very Google still. So Gemini would have to be drastically better answer wise than ChatGPT to lure users from a bett…

I couldn't even get ChatGPT to let me download code it claimed to program for me. It kept saying the files were ready but refused to let me access or download anything. It was the most basic use case and it totally bombed. I gave up on ChatGPT right then and there. It's amazing how different people have wildly varying experiences with the same product.

Did you wait a while before downloading? The links it provides for temporary projects have a surprisingly brief window where you can download them. I've had similar experience when even waiting 1 minute to download the file.

Re: OpenAI declares 'code red' as Google catches up in AI race

#163

The real code red here is less that Google just one-upped OpenAI but that they demonstrated there’s no moat to be had here. Absent a major breakthrough all the major providers are just going to keep leapfrogging each other in the most expensive race to the bottom of all time. Good for tech, but a horrible business and financial picture for these companies.

Especially if we're approaching a plateau, in a couple years there could be a dozen equally capable systems. It'll be interesting to see what the differentiators turn out to be.

Re: OpenAI declares 'code red' as Google catches up in AI race

#164

Heard all the news how Gemini 3 is passing everyone on benchmarks, so quickly tested and still find it a far cry from ChatGPT in real world use when testing questions on both platforms. But importantly the ChatGPT app experience at least for iPhone/Mac users is drastically superior vs Google which feels very Google still. So Gemini would have to be drastically better answer wise than ChatGPT to lure users from a bett…

Gemini comes with the 1.99 Google One plan. So I use that

Re: OpenAI declares 'code red' as Google catches up in AI race

#165

Earlier quoted context omitted.

I think it’s entirely possible that AI actually has plateaued, or has reached a point where a jump in intelligence comes at the cost of reliability.

I suspect it's reached the point where the distinguishing quality of one model over the others is only observable by true experts -- and only in their respective fields. We are exhausting the well of frontier questions that can be programmatically asked and the answers checked.

Absolutely this. Strong disagree that progress is plateauing, merely that gains are harder for the general public to perceive and typically come from more advanced means than simply scaling. Math performance in particular is improving at an uncomfortably rapid pace.

Re: OpenAI declares 'code red' as Google catches up in AI race

#166

"We’re currently experiencing issues" https://status.openai.com/

That looks pretty... amateurish. I can't imagine selling customer a service that doesn't even hit the third nine

"issues" don't mean "down"...

Re: OpenAI declares 'code red' as Google catches up in AI race

#167

Heard all the news how Gemini 3 is passing everyone on benchmarks, so quickly tested and still find it a far cry from ChatGPT in real world use when testing questions on both platforms. But importantly the ChatGPT app experience at least for iPhone/Mac users is drastically superior vs Google which feels very Google still. So Gemini would have to be drastically better answer wise than ChatGPT to lure users from a bett…

That's really fascinating. Every real world use case I've tried on Gemini (especially math-related) absolutely slaughtered the performance of ChatGPT in speed and quality, not even close. As an Android user, the Gemini app is also far superior, since the ChatGPT app still doesn't properly display math equations, among plenty of other bugs.

Try doing some more casual requests.

When I asked both ChatGPT 5.1 Extended Thinking and Gemini 3 Pro Preview High for best daily casual socks both responses were okay and had a lot of the same options, but while the ChatGPT response included pictures, specs scraped from the product pages and working links, the Gemini response had no links. After asking for links, Gemini gave me ONLY dead links.

That is a recurring experience, Gemini seems to be supremely lazy to its own detriment quite often.

A minute ago I asked for best CR2032 deal for Aqara sensors in Norway, and Gemini recommended the long discontinued IKEA option, because it didn't bother to check for updated information. ChatGPT on the other hand actually checked prices and stock status for all the options it gave me.

Re: OpenAI declares 'code red' as Google catches up in AI race

#168
post #146

WSJ: Altman said OpenAI would be pushing back work on other initiatives, such as advertising, AI agents for health and shopping, and a personal assistant called Pulse. These plus working with Jony Ive on hardware, makes it sound like they took their eyes off the ball.

I don't think this is about Google. This is about advertising being the make or break moment for OpenAI.

The problem with ChatGPT advertising is that it's truly a "bet the farm" situation, unlike any of their projects in the past:

- If it works and prints money like it should, then OpenAI is on a path to become the next Mag 7 company. All the money they raised makes sense.

- If it fails to earn the expected revenue numbers, the ceiling has been penciled in. Sam Altman can't sell the jet pack / meal pill future anymore. Reality becomes cold and stark, as their most significant product has actual revenue numbers attached to it. This is what matters to the accountants, which is the lens through which OpenAI will be evaluated with from this point forward. If it isn't delivering revenue, then they raised way too much money - to an obscene degree. They won't be able to sell the wild far future vision anymore, and will be deleteriously held back by how much they've over-sold themselves.

The other problems that have been creeping up:

- This is the big bet. There is no AGI anymore.

- There is no moat on anything. Google is nipping at their heels. The Chinese are spinning up open source models left and right.

- Nothing at OpenAI is making enough money relative to the costs.

- Selling "AI" to corporate and expecting them to make use of it hasn't been working. Those contracts won't last forever. When they expire, businesses won't renew them.

My guess is that they've now conducted small scale limited tests of advertising and aren't seeing the engagement numbers they need. It's truly a nightmare scenario outcome for them, if so.

They're declaring "code red" loudly and publicly to distract the public from this and to bide more time. Maybe even to raise some additional capital (yikes).

They're saying other things are more important than "working on advertising" right now. And they made sure to mention "advertising" lots so we know "advertising" is on hold. Which is supposedly the new golden goose.

Why drop work on a money printer? What could be more important? Unless the money printer turned out to be a dud.

Didn't we kind of already know advertising would fail on a product like this? Didn't Amazon try to sell via Alexa and have that totally flop? I'm not sure why ChatGPT would be any different from that experience. It's not a "URL bar" type experience like Google has. They don't own every ingress to the web like Google, and they don't own a infinite scroll FOMO feed of fashion like Meta. The ad oppo here is like Quora or Stack Overflow - probably not great.

I have never once asked ChatGPT for shopping ideas. But Google stands in my search for products all the time. Not so much as a "product recommendation engine", but usually just a bridge troll collecting its toll.

Re: OpenAI declares 'code red' as Google catches up in AI race

#169

Heard all the news how Gemini 3 is passing everyone on benchmarks, so quickly tested and still find it a far cry from ChatGPT in real world use when testing questions on both platforms. But importantly the ChatGPT app experience at least for iPhone/Mac users is drastically superior vs Google which feels very Google still. So Gemini would have to be drastically better answer wise than ChatGPT to lure users from a bett…

That's really fascinating. Every real world use case I've tried on Gemini (especially math-related) absolutely slaughtered the performance of ChatGPT in speed and quality, not even close. As an Android user, the Gemini app is also far superior, since the ChatGPT app still doesn't properly display math equations, among plenty of other bugs.

It's generally anecdotal and vibes when people make claims that some AI is better than another for things they do. There are too many variables and not enough eval for any of it to hold water imo. Personal preferences, experience, brand loyalty, and bias at play too

it's contemporary vim vs emacs at this point

Re: OpenAI declares 'code red' as Google catches up in AI race

#170

Earlier quoted context omitted.

chatgpt making targeted "recommendations" (read ads) is a nightmare. especially if it's subtle and not disclosed.

It'll be hard to separate them out from the block of prose. It's not like Google results where you can highlight the sponsored ones.

I mean google does everything possible to blur that line while still trying to say that it is telling you it is an ad.
Post reply on HN