Live data from Hacker News

Nano Banana Pro

blog.google

381–390 of 718 posts

Re: Nano Banana Pro

#381

The interesting tidbit here is SynthID. While a good first step, it doesn't solve the problem of AI generated content NOT having any kind of watermark. So we can prove that something WITH the ID is AI generated but we can't prove that something without one ISN'T AI generated. Like it would be nice if all photo and video generated by the big players would have some kind of standardized identifier on them - but now you…

Some days it feels like I'm the only hacker left who doesn't want government mandated watermarking in creative tools. Were politicians 20 years ago as overreative they'd have demanded Photoshop leave a trace on anything it edited. The amount of moral panic is off the charts. It's still a computer, and we still shouldn't trust everything we see. The fundamentals haven't changed.

HN is full of authoritarian bootlickers who can't imagine that people can exist without a paternalistic force to keep them from doing bad things.

Re: Nano Banana Pro

#382

Wow! I was able to combine Nano Banana Pro and Veo 3.1 video generation in a single chat and it produced great results. https://chat.vlm.run/c/38b99710-560c-4967-839b-4578a4146956 . Really cool model

I see many recent accounts posting vlm.run links and if this is what I suspect it is, that's normally not allowed here.

If you have concerns about spam, the right thing to do is to email the mods at hn@ycombinator.com with examples.

Re: Nano Banana Pro

#383

I really hope Google reads these HN posts. They've had some big "product" wins but the pricing, packaging, and user system is a severe blocker to growth. If developers can't or won't figure it out -- how the heck are consumers?

And both their consumer apps are slow. You can replicate this yourself. Go to AI Studio, paste in 80K tokens of text, then type something on your keyboard, and see what happens. The Gemini web app is even worse somehow. A horrifically slow and buggy app. Not new problems either, barely any improvement on this over more than 1 year.

Re: Nano Banana Pro

#385
post #301

Earlier quoted context omitted.

No, this is squarely on the AI. A human would know what you mean without specific instructions.

Seems like you're making a judgment based on your own experience, but as another commenter pointed out, it was wrong. There are plenty of us out there who would confirm, because people are too flawed to trust. Humans double/triple check, especially under higher stakes conditions (surgery). Heck, humans are so flawed, they'll put the things in the wrong eye socket even knowing full well exactly where they should go -…

Intelligence in my book includes error correction. Questioning possible mistakes is part of wisdom.

So the understanding that AI and HI are different entities altogether with only a subset of communication protocols between them will become more and more obvious, like some comments here are already implicitly telling.

Re: Nano Banana Pro

#386

I...worked on the detailed Nano Banana prompt engineering analysis for months ( https://news.ycombinator.com/item?id=45917875 )...and...Google just...Google released a new version. Nano Banana Pro should work with my gemimg package ( https://github.com/minimaxir/gemimg ) without pushing a new version by passing: g = GemImg(model="gemini-3-pro-image-preview") I'll add the new output resolutions and other features ASAP…

>> - Put a strawberry in the left eye socket. >>- Put a blackberry in the right eye socket. >> All five of the edits are implemented correctly This is a GREAT example of the (not so) subtle mistakes AI will make in image generation, or code creation, or your future knee surgery. The model placed the specified items in the eye sockets based on the viewers left/right; when we talk relative in this scenario we usually (…

There's a classic well-illustrated book, _How to Keep Your Volkswagen Alive_, which spends a whole illustrated page at the beginning building up a reference frame for working on the vehicle. Up is sky, down is ground, front is always vehicle's front, left is always vehicle's left.

Sounds a bit silly to write it out, but the diagram did a great job removing ambiguity when you expect someone to be laying on the ground in a tight place looking backwards, upside down.

Also feels important to note that in the theatre, there is stage-right and stage-left, jargon to disambiguate even though the jargon expects you to know the meaning to understand it.

Re: Nano Banana Pro

#387
post #193

This thing's ability to produce entire infographics from a short prompt is really impressive, especially since it can run extra Google searches first. I tried this prompt: Infographic explaining how the Datasette open source project works Here's the result: https://simonwillison.net/2025/Nov/20/nano-banana-pro/#creat...

It didn’t do so well at finding middle C on a piano keyboard: https://gemini.google.com/share/c9af8de05628 I did manage to get one image of a piano keyboard where the black keys were correct, but not consistently.

Fooled me because it was locally correct!

Re: Nano Banana Pro

#388

Google has been stomping around like Godzilla this week, and this is the first time I decided to link my card to their AI studio. I had seen people saying that they gave up and went to another platform because it was "impossible to pay". I thought this was strange, but after trying to get a working API key for the past half hour, I see what they mean. Everything is set up, I see a message that says "You're using Paid…

It's amazing that the "hard problems" are turning out to be "not creating a completely broken user experience".

Is that going to need AGI? Or maybe it will always be out of reach of our silicon overlords and require human input.

Re: Nano Banana Pro

#389
I don't understand the excitement around generating and/or watching AI-produced videos. To me it's probably the single most uninteresting and boring thing related to AI that I can think of. What is the appeal?

Re: Nano Banana Pro

#390
post #85

Earlier quoted context omitted.

Your use case doesn't even make sense. What customers are clamoring for that feature? I doubt any paying customer in the market for (that product) cares. If the law cares, the law has tools to inquire. All of this is trivially easy to circumvent ceremony. Google is doing this to deflect litigation and to preserve their brand in the face of negative press. They'll do this (1) as long as they're the market leader, (2)…

> What customers are clamoring for that feature? If the law cares, the law has tools to inquire. How can they distinguish from real people exploited to AI models autogenerating everything? I mean right now this is possible, largely because a lot of the AI videos have shortcomings. But imagine in 5 years from now on ...

> How can they distinguish from real people exploited to AI models autogenerating everything?

Watermarking by compliant models doesn't help this much because (1) models without watermarking exist and can continue to be developed (especially if absence of a watermark is treated as a sign of authenticity), so you cannot rely on AI fakery being watermarked, and (2) AI models can be used for video-to-video generation without changing much of the source, so you can't rely on something accurately watermarked as "AI-generated" not being based in actual exploitation.

Now, if the watermarking includes provenance information, and you require certain types of content to be watermarked not just as AI using a known watermarking system, but by a registered AI provider with regulated input data safety guardrails and/or retention requirements, and be traceable to a registered user, and...

Well, then it does something when it is present, largely by creating a new content gatekeepiing cartel.

Post reply on HN