Live data from Hacker News

Gemini 2.0: our new AI model for the agentic era

blog.google

281–290 of 512 posts

Re: Gemini 2.0: our new AI model for the agentic era

#281

Earlier quoted context omitted.

That's not really any better, since "all of our products" already includes the subset that has at least 2B users. "I brought all my shoes, including all my red shoes."

They're pointing out that seven of their products have more than 2 billion users. "I brought all my shoes, including the pairs that cost over $10,000" is saying something about what shoes you brought, more than "all of them".

Why are they bragging about something completely unrelated in the middle of a sentence about the impact of a piece of technology?

-Hey, are you done packing?

-Yes, I decided I'll bring all my shoes, including the ones that cost over $10,000.

What, they just couldn't help themselves?

Re: Gemini 2.0: our new AI model for the agentic era

#282
post #13

Earlier quoted context omitted.

Are these benchmarks still meaningful?

I've started keeping an eye out for original brainteasers, just for that reason. GCHQ's Christmas puzzle just came out [1], and o1-pro got 6 out of 7 of them right. It took about 20 minutes in total. I wasn't going to bother trying those because I was pretty sure it wouldn't get any of them, but decided to give it an easy one (#4) and was impressed at the CoT. Meanwhile, Google's newest 2.0 Flash model went 0 for 7.…

Why are you comparing flash vs o1-pro, wouldn't a more fair comparison be flash vs mini?

Re: Gemini 2.0: our new AI model for the agentic era

#283

Earlier quoted context omitted.

They're pointing out that seven of their products have more than 2 billion users. "I brought all my shoes, including the pairs that cost over $10,000" is saying something about what shoes you brought, more than "all of them".

Why are they bragging about something completely unrelated in the middle of a sentence about the impact of a piece of technology? -Hey, are you done packing? -Yes, I decided I'll bring all my shoes, including the ones that cost over $10,000. What, they just couldn't help themselves?

The fact that they're using Gemini with even their most important products shows that they trust it.

Re: Gemini 2.0: our new AI model for the agentic era

#284

Earlier quoted context omitted.

Why are they bragging about something completely unrelated in the middle of a sentence about the impact of a piece of technology? -Hey, are you done packing? -Yes, I decided I'll bring all my shoes, including the ones that cost over $10,000. What, they just couldn't help themselves?

The fact that they're using Gemini with even their most important products shows that they trust it.

Again, that's covered by "all our products". Why do we need to be reminded that Google has a lot of users? Someone oblivious to that isn't going to care about this press release.

Re: Gemini 2.0: our new AI model for the agentic era

#285

Earlier quoted context omitted.

I've started keeping an eye out for original brainteasers, just for that reason. GCHQ's Christmas puzzle just came out [1], and o1-pro got 6 out of 7 of them right. It took about 20 minutes in total. I wasn't going to bother trying those because I was pretty sure it wouldn't get any of them, but decided to give it an easy one (#4) and was impressed at the CoT. Meanwhile, Google's newest 2.0 Flash model went 0 for 7.…

Why are you comparing flash vs o1-pro, wouldn't a more fair comparison be flash vs mini?

I just ask o1-mini the first two questions and it got it wrong.

Re: Gemini 2.0: our new AI model for the agentic era

#286
post #278
post #215

Earlier quoted context omitted.

That won't help. Their TOS and policies are vague enough that they can terminate all accounts you own (under "Use of multiple accounts for abuse" for instance).

To be fair, I believe this is reserved for things like fighting fraud.

It has been used a few times by people who had a Google Play app banned, that sometimes the personal account would get banned as well.

https://www.xda-developers.com/google-developer-account-ban-...

Re: Gemini 2.0: our new AI model for the agentic era

#287
post #113

I released a new llm-gemini plugin with support for the Gemini 2.0 Flash model, here's how to use that in the terminal: llm install -U llm-gemini llm -m gemini-2.0-flash-exp 'prompt goes here' LLM installation: https://llm.datasette.io/en/stable/setup.html Worth noting that the Gemini models have the ability to write and then execute Python code. I tried that like this: llm -m gemini-2.0-flash-exp -o code_execution 1…

Question: Have you tried using this for video? Alternately, if I wanted to pipe a bunch of screencaps into it and get one grand response, how would I do that? e.g. "Does the user perform a thumbs up gesture in any of these stills?" [edit: also, do you know the vision pricing? I couldn't find it easily]

Previous Gemini models worked really well for video, and this one can even handle steaming video: https://simonwillison.net/2024/Dec/11/gemini-2/#the-streamin...

Re: Gemini 2.0: our new AI model for the agentic era

#288
post #16

Am I alone in thinking the word “agentic” is dumb as shit? Most of these things seem to just be a system prompt and a tool that get invoked as part of a pipeline. They’re hardly “agents”. They’re modules.

>“agentic” is dumb as shit?

It'll create endless consulting opportunities for projects that never go anywhere and add nothing of value unless you value rich consultants.

Re: Gemini 2.0: our new AI model for the agentic era

#289

Earlier quoted context omitted.

I've started keeping an eye out for original brainteasers, just for that reason. GCHQ's Christmas puzzle just came out [1], and o1-pro got 6 out of 7 of them right. It took about 20 minutes in total. I wasn't going to bother trying those because I was pretty sure it wouldn't get any of them, but decided to give it an easy one (#4) and was impressed at the CoT. Meanwhile, Google's newest 2.0 Flash model went 0 for 7.…

Why are you comparing flash vs o1-pro, wouldn't a more fair comparison be flash vs mini?

[deleted]

Re: Gemini 2.0: our new AI model for the agentic era

#290

Earlier quoted context omitted.

Yet, google continues to show it'll deprecate it's APIs, Services, and Functionality at the detriment of your own business. I'm not sure enterprises will trust Google's LLM over the alternatives. Too many have been burned throughout the years, including GCP customers. The fact GCP needs to have this page, and these lists are not 100% comprehensive is telling enough. https://cloud.google.com/compute/docs/deprecations…

GCP grew 35% last quarter , just saying ...

"just saying" things that are false.

Google Cloud grew 35% year over year, when comparing the 3 months ending September 30th 2024 with 2023.

https://abc.xyz/assets/94/93/52071fba4229a93331939f9bc31c/go... page 12

Post reply on HN