Live data from Hacker News

Gemini-2.5-pro-preview-06-05

deepmind.google

161–170 of 237 posts

Re: Gemini-2.5-pro-preview-06-05

#161
post #110

Earlier quoted context omitted.

The hurdle for OpenAI is going to be on the profit side. Google has their own hardware acceleration and their own data centers. OpenAI has to pay a monopolist for hardware acceleration and beholden to another tech giant for data centers. Never mind that Google can customize it's hardware specifically for it's models. The only way for OpenAI to really get ahead on solid ground is to discover some sort of absolute game…

OpenAI has now partnered with Jony Ive now and they are going to have thinnest data centers with thinnest servers mounted on thinnest racks. And since everything is so thin, servers can just whisper to each other instead of communicating via fat cables. I think that will be the game changer OpenAI will show us soon.

[deleted]

Re: Gemini-2.5-pro-preview-06-05

#162
post #12

Impressive seeing Google notch up another ~25 ELO on lmarena, on top of the previous #1, which was also Gemini! That being said, I'm starting to doubt the leaderboards as an accurate representation of model ability. While I do think Gemini is a good model, having used both Gemini and Claude Opus 4 extensively in the last couple of weeks I think Opus is in another league entirely. I've been dealing with a number of gn…

how does it have access to DOM? are you using it with cursor/browser MCP?

Re: Gemini-2.5-pro-preview-06-05

#163
post #32

I found all the previous Gemini models somewhat inferior even compared to Claude 3.7 Sonnet (and much worse than 4) as my coding assistants. I'm keeping an open mind but also not rushing to try this one until some evaluations roll in. I'm actually baffled that the internet at large seems to be very pumped about Gemini but it's not reflective of my personal experience. Not to be that tinfoil hat guy but I smell at lea…

> I'm actually baffled that the internet at large seems to be very pumped about Gemini but it's not reflective of my personal experience. Not to be that tinfoil hat guy but I smell at least a bit of astroturf activity around Gemini. I haven't used Claude, but Gemini has always returned better answers to general questions relative to ChatGPT or Copilot. My impression, which could be wrong, is that Gemini is better in…

I’ve honestly had consistently the opposite experiences for general questions. Also for images, Gemini just hallucinates crazily. ChatGPT even on free tier is giving perfectly correct answers, and I’m on Gemini pro. I canceled it yesterday because of this

Re: Gemini-2.5-pro-preview-06-05

#164

Earlier quoted context omitted.

Very strange. I find reasoning has very narrow usefulness for me. It's great to get a project in context or to get oriented in the conversation, but on long conversations I find reasoning starts to add way too much extraneous stuff and get distracted from the task at hand. I think my coding model ranking is something like Claude Code > Claude 4 raw > Gemini > big gap > o4-mini > o3

Claude Code isn't a model in itself. By default it routes some to Opus 4 or Sonnet 4 but mostly Sonnet 4 unless you explicitly set it.

I am aware

Re: Gemini-2.5-pro-preview-06-05

#165

I'd start to worry about OpenAI, from a valuation standpoint. The company has some serious competition now and is arguably no longer the leader. its going to be interesting to see how easily they can raise more money. Their valuation is already in the $300B range. How much larger can it get given their relatively paltry revenue at the moment and increasingly rising costs for hardware and electricity. If the next gene…

I was tempted by the ratings and immediately paid for a subscription to Gemini 2.5. Half an hour later, I canceled the subscription and got a refund. This is the laziest and stupidest LLM. What he had to do, he told me to do on my own. And also when analyzing simple short documents, he pulled up some completely strange documents from the Internet not related to the topic. Even local LLMs (3B) were not so stupid and lazy.

Re: Gemini-2.5-pro-preview-06-05

#166
post #157

Earlier quoted context omitted.

There is some serious confusion about the strength of OpenAIs position. "chatgpt" is a verb. People have no idea what claude or gemini are, and they will not be interested in it, unless something absolutely fantastic happens. Being a little better will do absolutely nothing to convince normal people to change product (the little moat that ChatGPT has simply by virtue of chat history is probably enough from a convenie…

Google has a text input box on google.com, as soon as this gives similar responses there is no need for the average user to use ChatGPT anymore. I already see lots of normal people share screenshots of the AI Overview responses.

You are skipping over the part where you need to bring normal people, specially young normal people, back to google.com for them to see anything at all on google.com. Hundreds of millions of them don't go there anymore.

Re: Gemini-2.5-pro-preview-06-05

#169
post #157

Earlier quoted context omitted.

There is some serious confusion about the strength of OpenAIs position. "chatgpt" is a verb. People have no idea what claude or gemini are, and they will not be interested in it, unless something absolutely fantastic happens. Being a little better will do absolutely nothing to convince normal people to change product (the little moat that ChatGPT has simply by virtue of chat history is probably enough from a convenie…

Google has a text input box on google.com, as soon as this gives similar responses there is no need for the average user to use ChatGPT anymore. I already see lots of normal people share screenshots of the AI Overview responses.

As the other poster mentioned, young people are not going there. What happens when they grow up?

Re: Gemini-2.5-pro-preview-06-05

#170

I'd start to worry about OpenAI, from a valuation standpoint. The company has some serious competition now and is arguably no longer the leader. its going to be interesting to see how easily they can raise more money. Their valuation is already in the $300B range. How much larger can it get given their relatively paltry revenue at the moment and increasingly rising costs for hardware and electricity. If the next gene…

I think it’s too early to say they are not the leader given they have o3 pro and GPT 5 coming out within the next month or two. Only if those are not impressive would I start to consider that they have lost their edge. Although it does feel likely that at minimum, they are neck and neck with Google and others.

Source for gpt 5 coming out soon?
Post reply on HN