Live data from Hacker News

Gemini 2.0: our new AI model for the agentic era

blog.google

131–140 of 512 posts

Re: Gemini 2.0: our new AI model for the agentic era

#131
No mention of Perplexity yet in the comments but it's obvious to me that they're targeting Perplexity Pro directly with their new Deep Research feature (https://blog.google/products/gemini/google-gemini-deep-resea...). I still wonder why Perplexity is worth $7 billion when the 800-pound gorilla is pounding on their door (albeit slowly).

Re: Gemini 2.0: our new AI model for the agentic era

#132
Unfortunately the 10rpm quota for this experimental model isn't enough to run an actual Agentic experience on.

That's my main issue with google there's several models we want to try with our agent but quota is limited and we have to jump through hoops to see if we can get it raised.

Re: Gemini 2.0: our new AI model for the agentic era

#133
post #72

Earlier quoted context omitted.

i don’t think they need to win the on device market. we need to separate inference and training - the real winners are those who have the training compute. you can always have other companies help with inference

> i don’t think they need to win the on device market. The second Apple comes out with strong on-device AI - and it very much looks like they will - Google will have to respond on Android. They can't just sit and pray that e.g. Samsung makes a competitive chip for this purpose.

But given inference time compute, to give a strong reply reasonably fast, you'll need a lot of compute, very rarely used.

Economically this fits the cloud much better.

Re: Gemini 2.0: our new AI model for the agentic era

#134
post #53

Earlier quoted context omitted.

They’ll care though when they have to pay for it, or when they’re in an area with poor reception.

Poor reception is rapidly becoming a non-issue for most of the developed world. I can’t think of the last time I had poor reception (in America) and wasn’t on an airplane. As the global human population increasingly urbanizes, it’ll become increasingly easy to blanket it with cell towers. Poor(er) regions of the world will increase reception more slowly, but they’re also more likely to have devices that don’t support…

Many major cities have significant dead spots for coverage. It’s not just for developing areas.

Flash is free for api use at a low rate limit. Gemini as a whole is not free to Android users (free right now with subscription costs beyond a time period for advanced features) and isn’t free to Google without some monetary incentive. Hence why I also originally ask about private cloud compute alternatives with Google.

Re: Gemini 2.0: our new AI model for the agentic era

#135
post #94

Earlier quoted context omitted.

And you can’t run the cloud model at all if you can’t talk to the cloud.

Yes, but I can't imagine situations where I "have" to run a model when I don't have internet at that time. My life would be more affected with the rest of the internet than having to run a small stupid model locally. At the very least until the hallucination is completely solved, as I need internet to verify the models.

You’re assuming the model is purely for generation though. Several of the Gemini features are lookup of things across data available to it. A lot of that data can be local to device.

That is currently Apple’s path with Apple Intelligence for example.

Re: Gemini 2.0: our new AI model for the agentic era

#136
Their offering is just so... bad. Even the new model. All the data in the world, yet they trail behind.

They have all of these extensions that they use to prop up the results in the web UI.

I was asking for a list of related YouTube videos - the UI returns them.

Ask the API the same prompt, it returns a bunch of made up YouTube titles and descriptions.

How could I ever rely on this product?

Re: Gemini 2.0: our new AI model for the agentic era

#137
post #17

The Gemini 2 models support native audio and image generation but the latter won't be generally available till January. Really excited for that as well as 4o's image generation (whenever that comes out). Steerability has lagged behind aesthetics in image generation for a while now and it's be great to see a big advance in that. Also a whole lot of computer vision tasks (via LLMs) could be unlocked with this. Think In…

These are not computer vision tasks…

Maybe some of these tasks are arguably not aligned with the traditional applications of CV, but Segmentation and Edge detection are definitely computer vision in every definition I've come across - before and after NNs took over.

Re: Gemini 2.0: our new AI model for the agentic era

#138

Big companies can be slow to pivot, and Google has been famously bad at getting people aligned and driving in one direction. But, once they do get moving in the right direction the can achieve things that smaller companies can't. Google has an insane amount of talent in this space, and seems to be getting the right results from that now. Remains to be seen how well they will be able to productize and market, but hard…

Yet, google continues to show it'll deprecate it's APIs, Services, and Functionality at the detriment of your own business. I'm not sure enterprises will trust Google's LLM over the alternatives. Too many have been burned throughout the years, including GCP customers.

The fact GCP needs to have this page, and these lists are not 100% comprehensive is telling enough. https://cloud.google.com/compute/docs/deprecations https://cloud.google.com/chronicle/docs/deprecations https://developers.google.com/maps/deprecations

Steve Yegge rightfully called this out, and yet no change has been made. https://medium.com/@steve.yegge/dear-google-cloud-your-depre...

Post reply on HN