Live data from Hacker News

Nano Banana Pro

blog.google

251–260 of 718 posts

Re: Nano Banana Pro

#251
post #100

Earlier quoted context omitted.

Random examples: 1) I have a tricep tendon injury and ChatGPT wants me to check my tricep reflex. I have no idea where on the elbow you're supposed to tap to trigger the reflex. 2) I'm measuring my body fat using skin fold calipers. Show me were the measurement sites are. 3) I'm going hiking. Remind me how to identify poison ivy and dangerous snakes. 4) What would I look like with a buzz cut?

First three are interesting - all question / knowledge based where the answer is a picture. Hadn't really considered this.

The answer is a picture that almost certainly already exists.

Why would you want a program that just makes one up instead?

Re: Nano Banana Pro

#252

The interesting tidbit here is SynthID. While a good first step, it doesn't solve the problem of AI generated content NOT having any kind of watermark. So we can prove that something WITH the ID is AI generated but we can't prove that something without one ISN'T AI generated. Like it would be nice if all photo and video generated by the big players would have some kind of standardized identifier on them - but now you…

Reminder that even in the hypothetical world where every AI image is digitally watermarked, and all cameras have a TPM that writes a hash of every photo to the blockchain, there’s nothing to stop you from pointing that perfectly-verified camera at a screen showing your perfectly-watermarked AI image and taking a picture. Image verification has never been easy. People have been airbrushed out of and pasted into photos…

Competent digital watermarks usually survive the 'analog hole'. Screen-cam resistant watermarks have been in use since at least 2020, and if memory serves, back to 2010 when I first starting reading about them, but I don't recall what it was called back then.

Re: Nano Banana Pro

#253

You can try it out for free on LMArena [0]: New Chat -> Battle dropdown -> Direct Chat -> Click on Generate Image in the chat box -> Click dropdown from hunyuan-image-3.0 -> gemini-3-pro-image-preview (nano-banana-pro). I've only managed to get a few prompts to go through, if it takes longer than 30 seconds it seems to just time out. Image quality seems to vary wildly; the first image I tried looked really good but t…

When I do that, I get two (very similar but not identical) responses side-by-side in one image (I guess as if the model is battling itself?). Is that normal for lmarena?

https://imgur.com/a/h0ncCFN

Re: Nano Banana Pro

#254

Earlier quoted context omitted.

Neat use-case, though the sword literally telescopically inverts itself at the beginning of the scene like a light saber where you would have expected it to be drawn from its scabbard. I'd be interested to see how Wan 2.2 First/Last frame handles those images though...

yeah sadly veo 3.1 has not caught up to the image generation capabilities. May be we need to work on how to make video generation more physically consistent. but the image generation results from banana pro are great.

another interesting use case with synth https://chat.vlm.run/c/1c726fab-04ef-47cc-923d-cb3b005d6262. made a puppet from a image of a model and made the puppet dance.

Re: Nano Banana Pro

#255
post #193

This thing's ability to produce entire infographics from a short prompt is really impressive, especially since it can run extra Google searches first. I tried this prompt: Infographic explaining how the Datasette open source project works Here's the result: https://simonwillison.net/2025/Nov/20/nano-banana-pro/#creat...

I’ve been really excited for you infographic generation. Previous models from Google and openAI had very low detail/resolution for these things.

I’ve found in general that the first generation may not be accurate but a few rolls of the dice and you should have enough to pick a style and format that works, which you can iterate on.

Re: Nano Banana Pro

#256
Does anyone know if this is predicting the entire image at once, or if it's breaking it into constituent steps i.e. "draw text in this font at this location" and then composing it from those "tools"? It would be really interesting if they've solved the garbled text problem within the constraint of predicting the entire image at once.

Re: Nano Banana Pro

#257
I was just playing with the non-pro version of this and it seems to add both a Gemini and Disney watermark. Presumably this was because I referenced beauty and the beast.

Anyone know if this is an hallucination or if they have some kind of deal with content owners to add branding?

Re: Nano Banana Pro

#258

I tried the same prompt as one of the examples ( https://i.imgur.com/iQTPJzz.png ), in the two ways they say you can run it, via Google Gemini and Google AI Studio (I suppose they're different somehow?). The prompt was "Create an infographic that shows hot to make elaichi chai" and Google Gemini created a infographic ( https://i.imgur.com/aXlRzTR.png ), but it was all different from what the example showed. Google AI…

If it were illegal to intentionally mislead people, many magicians would be out of a job :)

Re: Nano Banana Pro

#259
post #58

I've had nano banana pro for a few weeks now, and it's the most impressive AI model I've ever seen The inline verification of images following the prompt is awesome, and you can do some _amazing_ stuff with it. It's probably not as fun anymore though (in the early access program, it doesn't have censoring!)

LLMs might be a dead end, but we're going to have amazing images, video, and 3D. To me the AI revolution is making visual media (and music) catch up with the text-based revolution we've had since the dawn of computing. Computers accelerated typing and text almost immediately, but we've had really crude tools for images, video, and 3D despite graphics and image processing algorithms. AI really pushes the envelope here…

How can LLMs be a dead end? The last improvement in LLMs came out this week.

Re: Nano Banana Pro

#260

I've tried to repaint the exterior of my house. More than 20 times with very detailed prompts. I even tried to optimize it with Claude. No matter what, every time it added one, two or three extra windows to the same wall.

[deleted]
Post reply on HN