Live data from Hacker News

Google is winning on every AI front

thealgorithmicbridge.com

91–100 of 851 posts

Re: Google is winning on every AI front

#91

Google is winning on every front except... marketing (Google has a chatbot?), trust (who knew the founding fathers were so diverse?), safety (where's the 2.5 Pro model card?), market share (fully one in ten internet users on the planet are weekly ChatGPT users), and, well, vibes (who's rooting for big G, exactly?). But I will admit, Gemini Pro 2.5 is a legit good model. So, hats off for that.

Google is also terribly paranoid of the LLM saying anything controversial. If you want a summary of some hot topic article you might not have the time to read, Gemini will straight up refuse to answer. ChatGPT and Grok don't mind at all.

I noticed the same in Gemini. It would refuse to answer mundane questions that none but the most 'enlightened' could find an offensive twist to.

This makes it rather unusable as a catch all goto resource, sadly. People are curious by nature. Refusing to answer their questions doesn't squash that, it leads them to potentially less trustworthy sources.

Re: Google is winning on every AI front

#92
post #37

Earlier quoted context omitted.

Your issue is because: 1- the cursor agent doesn’t work with gemini. Some times the diff edit even doesn’t work. 2- Cursor does semantic search to lower the token they sent to models. The big advantage for Gemini is the context window, use it with aider, clien or roo code.

> clien or roo code What's the difference between Cline and Roo Code now? Originally Roo was a fork of Cline that added a couple of extra settings. But now it seems like an entirely different app, with it's own website even. https://roocode.com/

I hope this will help:

https://www.reddit.com/r/RooCode/comments/1jn372q/roocode_vs...

Re: Google is winning on every AI front

#94
post #25

In my experience Claude 3.7 is far superior for coding than Gemini 2.5. I tried it in Cursor and I wanted it to work, as a recent ex-Googler. I repeatedly found it inferior. I think it’s still behind Claud 3.5 for coding. It would decide arbitrarily not to finish tasks and suggest that I do them. It made simple errors and failed to catch them.

Same here. I've seen some articles and LLM benchmarks that Gemini 2.5 Pro is better than Claude 3.7 on coding, but base on my recent experience of solving code problems with two products, Claude still gave me better answer, Gemini response are more detail and well structured, but less accurate.

Re: Google is winning on every AI front

#95
post #86
post #78

Earlier quoted context omitted.

> October in Sydney These sound like fairly dated anecdotes. I don't doubt them at all - I had similar horror stories. I disabled Gemini on my phone in order to keep the old assistant for a long time, but it really has gotten a lot better in the last few months.

March 11th was when the alarm one was From, these are just ones that I have screenshot because they were so bad I shared with a friend. https://imgur.com/a/nj4newx Edit: I just asked it for the weather this week and it only showed today. Like this is Amateur hour stuff, Siri 1.0 stuff. https://imgur.com/a/81mz98Y

I can replicate your weather one! I think it's taking "this week" extremely literally and the week ends on Saturday. Asking for "this weekend" gives Saturday and Sunday. Asking for the next few days gives 3 days out, etc .

Definitely not addressing the spirit of the request.

Re: Google is winning on every AI front

#96
post #6

Gemini 2.5 pro is as powerful as everybody says. I still also use Claude Sonnet 3.7 only because the Gemini web UI has issues... (Imagine creating the best AI and then not allowing to attach Python or C files if not renamed .txt) but the way the model is better than anyone else is a "that's another league" experience. They have the biggest search engine and YouTube to leverage the power of the AI they are developing.…

Will there be a winner at all? Perhaps it's going to be like cars where there are dozens of world class manufacturers, or like Linux, where there's just one thing, but its free and impossible to monetize directly.

Re: Google is winning on every AI front

#97

Google is winning on every front except... marketing (Google has a chatbot?), trust (who knew the founding fathers were so diverse?), safety (where's the 2.5 Pro model card?), market share (fully one in ten internet users on the planet are weekly ChatGPT users), and, well, vibes (who's rooting for big G, exactly?). But I will admit, Gemini Pro 2.5 is a legit good model. So, hats off for that.

Didn't GCP manage to lose from this position of strength? I'm not sure even if they're the third biggest

Re: Google is winning on every AI front

#98

As an Ex-OpenAI employee I agree with this. Most of the top ML talent at OpenAI already have left to either do their own thing or join other startups. A few are still there but I doubt if they'll be around in a year. The main successful product from OpenAI is the ChatGPT app, but there's a limit on how much you can charge people for subscription fees. I think soon people expect this service to be provided for free an…

> Google can't as easily burn money

I was actually surprised at Google's willingness to offer Gemini 2.5 Pro via AI Studio for free; having this was a significant contributor to my decision to cancel my OpenAI subscription.

Re: Google is winning on every AI front

#99
> Gemini 2.5 Pro in Deep Research mode is twice as good as OpenAI’s Deep Research

That matches my impression. For the past month or two, I have been running informal side-by-side tests of the Deep Research products from OpenAI, Perplexity, and Google. OpenAI was clearly winning—more complete and incisive, and no hallucinated sources that I noticed.

That changed a few days ago, when Google switched their Deep Research over to Gemini 2.5 Pro Experimental. While OpenAI’s and Perplexity’s reports are still pretty good, Google’s usually seem deeper, more complete, and more incisive.

My prompting technique, by the way, is to first explain to a regular model the problem I’m interested in and ask it to write a full prompt that can be given to a reasoning LLM that can search the web. I check the suggested prompt, make a change or two, and then feed it to the Deep Research models.

One thing I’ve been playing with is asking for reports that discuss and connect three disparate topics. Below are the reports that the three Deep Research models gave me just now on surrealism, Freudian dream theory, and AI image prompt engineering. Deciding which is best is left as an exercise to the reader.

OpenAI:

https://chatgpt.com/share/67fa21eb-18a4-8011-9a97-9f8b051ad3...

Google:

https://docs.google.com/document/d/10mF_qThVcoJ5ouPMW-xKg7Cy...

Perplexity:

https://www.perplexity.ai/search/subject-analytical-report-i...

Re: Google is winning on every AI front

#100
> Add to the above that Gemini 2.5, compared to models of its category, is fast and cheap—I mean, they're giving away free access!

A large player with massive existing streams giving away a product in a new market to undercut new entrants? Looks an awful lot like abuse of monopoly position...

Post reply on HN