Live data from Hacker News

Nano Banana Pro

blog.google

361–370 of 718 posts

Re: Nano Banana Pro

#361

There's some really impressive things about this (the speed, the lack of typical AI image gen artifacts) but it also seems less creative than other models I've tried? "mountain dew themed pokemon" is the first search prompt I always try with new image models and Nano Banna Pro just gave me a green pikachu. Other models do a much better job of creating something new.

IMHO I'd rather them focus on strong literal prompt adherence so that more detailed prompts produce more accurate results.

That way you can stick your choice of any number of LLM preprocessors in front of a generic prompt like "mountain dew themed pokemon" and push the responsibility of creating a more detailed prompt upstream.

https://imgur.com/a/s5zfxS5

Note: I'm not particularly impressed with either of the results - this is more a demonstration.

Re: Nano Banana Pro

#362

I've had nano banana pro for a few weeks now, and it's the most impressive AI model I've ever seen The inline verification of images following the prompt is awesome, and you can do some _amazing_ stuff with it. It's probably not as fun anymore though (in the early access program, it doesn't have censoring!)

Genuinely believe that images are 99.5% solved now and unless you’re extremely keen eyed, you won’t be able to tell AI images from real images now

Re: Nano Banana Pro

#363

Does anyone know if this is predicting the entire image at once, or if it's breaking it into constituent steps i.e. "draw text in this font at this location" and then composing it from those "tools"? It would be really interesting if they've solved the garbled text problem within the constraint of predicting the entire image at once.

I’m pretty sure, but no expert on the matter, that correct text rendering was solved by feeding in bitmaps of rasterized fonts as supplemental context to the image generation models.

Re: Nano Banana Pro

#365

Earlier quoted context omitted.

Super important for Google as a search engine so they can filter out and downrank AI generated results. However I expect there are many models out there which don’t do this, that everyone could use instead. So in the end a “feature” like this makes me less likely to use their model because I don’t know how Google will end up treating my blog post if I decide to include an AI generated or AI edited image.

It’s required by EU regulations. Any public generator that doesn’t do it, is in violation of that unless it’s entirely inaccessible from the EU… But of course there’s no way to enforce it on local generation.

The EU didn't define any specific method of watermarking nor does it need to be tamper resistant. Even if they had specified it though, it's easy to remove watermarks like SynthID.

Re: Nano Banana Pro

#366

Earlier quoted context omitted.

> It's still a computer, and we still shouldn't trust everything we see. The fundamentals haven't changed. I think that by now it should be crystal clear to everyone that it matters a lot the sheer scale a new technology permits for $nefarious_intent. Knives (under a certain size) are not regulated. Guns are regulated in most countries. Atomic bombs are definitely regulated. They can all kill people if used badly, th…

I think we're overreacting. Digital fakes will proliferate, and we'll freak out bc it's new. But after a certain amount of time, we'll just get used to it and realize that the world goes on, and whatever major adverse effects actually aren't that difficult to deal with. Which is not the case with nuclear proliferation or things like that. The story of human history is newer generations freaking about progress and nov…

I think the long term effect will be that photos and videos no longer have any evidentiary value legally or socially, absent a trusted chain of custody.

Re: Nano Banana Pro

#367

Google has been stomping around like Godzilla this week, and this is the first time I decided to link my card to their AI studio. I had seen people saying that they gave up and went to another platform because it was "impossible to pay". I thought this was strange, but after trying to get a working API key for the past half hour, I see what they mean. Everything is set up, I see a message that says "You're using Paid…

Oh my, you should have tried to integrate with Google Prism. That was a madness! Nano Banana was just a little tricky to set up in comparison!

Re: Nano Banana Pro

#369

Earlier quoted context omitted.

>This is a GREAT example of the (not so) subtle mistakes AI will make in image generation, or code creation, or your future knee surgery. The mistake is in the prompting (not enough information). The AI did the best it could "What's the biggest known planet" "Jupiter" "NO I MEANT IN THE UNIVERSE!"

It doesn't affect your point but technically since the IAU are insane, exoplanets aren't technically planets and Jupiter is the largest planet in the universe.

I suppose it was too much to hope that chatbots could be trained to avoid pointless pedantry.

Re: Nano Banana Pro

#370
I really hope Google reads these HN posts. They've had some big "product" wins but the pricing, packaging, and user system is a severe blocker to growth. If developers can't or won't figure it out -- how the heck are consumers?
Post reply on HN