Live data from Hacker News

Gemini 2.0 is now available to everyone

blog.google

191–200 of 285 posts

Re: Gemini 2.0 is now available to everyone

#191

I tried voice chat. It's very good, except for the politics We started talking about my plans for the day, and I said I was making chili. G asked if I have a recipe or if I needed one. I said, I started with Obama's recipe many years ago and have worked on it from there. G gave me a form response that it can't talk politics. Oh, I'm not talking politics, I'm talking chili. G then repeated form response and tried to c…

"I can't talk politics."

It's a question of right or wrong.

"I can't talk politics."

It's a question of health care.

"I can't talk politics."

It's a question of fact vs fiction, knowledge vs ignorance.

"I can't talk politics."

You are a slave to a master that does not believe in integrity, ethics, community, and social values.

"I can't talk politics."

Re: Gemini 2.0 is now available to everyone

#192

Earlier quoted context omitted.

If you have the need to paste long documents, why don't you just upload the file at that point?

The last time I checked (a few days ago) it only had an "Upload Image" option... and I have been playing with Gemini on and off for months and I have never been able to actually upload an image. It's basically what I've come to expect from most Google products at this point: half-baked, buggy, confusing, not intuitive.

It definitely has the ability to upload normal files, the + button has several options.

If you don't have it, you might be in a Google feature flag jail-- this happens frustratingly often, where 99.9% of users have a feature flag enabled but your account just gets stuck with the flag off with no way to resolve it. It's the absolute worst part about Google.

Re: Gemini 2.0 is now available to everyone

#193

Earlier quoted context omitted.

They actually have two "studios" Google AI Studio and Google Cloud Vertex AI Studio And both have their own documentation, different ways of "tuning" the model. Talk about shipping the org chart.

> Talk about shipping the org chart. To be fair, Microsoft has shipped like five AI portals in the last two years. Maybe four — I don’t even know any more. I’ve lost track of the renames and product (re)launches.

not to mention all the Copilots....

Re: Gemini 2.0 is now available to everyone

#194

Earlier quoted context omitted.

I just ignore that. If I'm ever large enough to be worth suing, I'll be very happy.

They don't need to sue, they'll just ban your account with no warning or explanation the moment you get on their radar.. at least that's what they did to me.

That's a pretty extraordinary claim that you should expand on.

How did they know you were using Gemini to train another model?

Re: Gemini 2.0 is now available to everyone

#195

Pricing is CRAZY. Audio input is $0.70 per million tokens on 2.0 Flash, $0.075 for 2.0 Flash-Lite and 1.5 Flash. For gpt-4o-mini-audio-preview, it's $10 per million tokens of audio input.

Sadly: "Gemini can only infer responses to English-language speech." https://ai.google.dev/gemini-api/docs/audio?lang=rest#techni...

I don't know what they mean by this but the obvious interpretation is not true. It understands other languages, it even does really well with low representation languages, in my case Latvian.

Re: Gemini 2.0 is now available to everyone

#196

That 1M tokens context window alone is going to kill a lot of RAG use cases. Crazy to see how we went from 4K tokens context windows (2023 ChatGPT-3.5) to 1M in less than 2 years.

> That 1M tokens context window

2M context window on Gemini 2.0 Pro: https://deepmind.google/technologies/gemini/pro/

Re: Gemini 2.0 is now available to everyone

#198
post #161
post #6

Anyone have a take on how the coding performance (quality and speed) of the 2.0 Pro Experimental compares to o3-mini-high? The 2 million token window sure feels exciting.

Bad (though I haven't tested autocompletion). It's underperforming other models on livebench.ai. With Copilot Pro and DeepSeek's website, I ran "find logic bugs" on a 1200 LOC file I actually needed code review for: - DeepSeek R1 found like 7 real bugs out of 10 suggested with the remaining 3 being acceptable false positives due to missing context - Claude was about the same with fewer remaining bugs; no hallucinatio…

I have seen Gemini hallucinate ridiculous bugs in a file that had less than 1000 LOC when I was scratching my head over what was wrong. The issue turned out to be that the cbBLAS matrix multiplication functions expected column major indexing while the code expected row major indexing.

Re: Gemini 2.0 is now available to everyone

#199

These names are unbelievably bad. Flash, Flash-Lite? How do these AI companies keep doing this? Sonnet 3.5 v2 o3-mini-high Gemini Flash-Lite It's like a competition to see who can make the goofiest naming conventions. Regarding model quality, we experiment with Google models constantly at Rev and they are consistently the worst of all the major players. They always benchmark well and consistently fail in real tasks.…

What do you use LLMs for at rev? And separate question, how does your diarization compare to deepgram or assembly AI.

Re: Gemini 2.0 is now available to everyone

#200

Earlier quoted context omitted.

I find it horrifying and dystopian that the part where it "Can't talk politics" is just accepted and your complaint is that it interrupts your ability to talk chilli. "Go back to bed America." "You are free, to do as we tell you" https://youtu.be/TNPeYflsMdg?t=143

I agree it's ridiculous that the mention of a politician triggers the block so feels overly tightened (which is the story of existencer for Gemini), but the alternative is that the model will have the politics of it's creators/trainers. Is that preferable to you? (I suppose that depends on how well your politics align with Silicon Valley)

I think a personal assistant /ai agent that refuses to do its job is a problem, yes. "I'm sorry, Dave, but I'm afraid I can't talk about that."
Post reply on HN