The interesting tidbit here is SynthID. While a good first step, it doesn't solve the problem of AI generated content NOT having any kind of watermark. So we can prove that something WITH the ID is AI generated but we can't prove that something without one ISN'T AI generated. Like it would be nice if all photo and video generated by the big players would have some kind of standardized identifier on them - but now you…
Some days it feels like I'm the only hacker left who doesn't want government mandated watermarking in creative tools. Were politicians 20 years ago as overreative they'd have demanded Photoshop leave a trace on anything it edited. The amount of moral panic is off the charts. It's still a computer, and we still shouldn't trust everything we see. The fundamentals haven't changed.
Nano Banana Pro
381–390 of 718 posts
Re: Nano Banana Pro
#382Wow! I was able to combine Nano Banana Pro and Veo 3.1 video generation in a single chat and it produced great results. https://chat.vlm.run/c/38b99710-560c-4967-839b-4578a4146956 . Really cool model
I see many recent accounts posting vlm.run links and if this is what I suspect it is, that's normally not allowed here.
Re: Nano Banana Pro
#383I really hope Google reads these HN posts. They've had some big "product" wins but the pricing, packaging, and user system is a severe blocker to growth. If developers can't or won't figure it out -- how the heck are consumers?
Re: Nano Banana Pro
#384Earlier quoted context omitted.
That is patently false.
So, uh... do you know of an implementation that has both those properties? I'd be quite interested in that.
Re: Nano Banana Pro
#385Earlier quoted context omitted.
No, this is squarely on the AI. A human would know what you mean without specific instructions.
Seems like you're making a judgment based on your own experience, but as another commenter pointed out, it was wrong. There are plenty of us out there who would confirm, because people are too flawed to trust. Humans double/triple check, especially under higher stakes conditions (surgery). Heck, humans are so flawed, they'll put the things in the wrong eye socket even knowing full well exactly where they should go -…
So the understanding that AI and HI are different entities altogether with only a subset of communication protocols between them will become more and more obvious, like some comments here are already implicitly telling.
Re: Nano Banana Pro
#386I...worked on the detailed Nano Banana prompt engineering analysis for months ( https://news.ycombinator.com/item?id=45917875 )...and...Google just...Google released a new version. Nano Banana Pro should work with my gemimg package ( https://github.com/minimaxir/gemimg ) without pushing a new version by passing: g = GemImg(model="gemini-3-pro-image-preview") I'll add the new output resolutions and other features ASAP…
>> - Put a strawberry in the left eye socket. >>- Put a blackberry in the right eye socket. >> All five of the edits are implemented correctly This is a GREAT example of the (not so) subtle mistakes AI will make in image generation, or code creation, or your future knee surgery. The model placed the specified items in the eye sockets based on the viewers left/right; when we talk relative in this scenario we usually (…
Sounds a bit silly to write it out, but the diagram did a great job removing ambiguity when you expect someone to be laying on the ground in a tight place looking backwards, upside down.
Also feels important to note that in the theatre, there is stage-right and stage-left, jargon to disambiguate even though the jargon expects you to know the meaning to understand it.
Re: Nano Banana Pro
#387This thing's ability to produce entire infographics from a short prompt is really impressive, especially since it can run extra Google searches first. I tried this prompt: Infographic explaining how the Datasette open source project works Here's the result: https://simonwillison.net/2025/Nov/20/nano-banana-pro/#creat...
It didn’t do so well at finding middle C on a piano keyboard: https://gemini.google.com/share/c9af8de05628 I did manage to get one image of a piano keyboard where the black keys were correct, but not consistently.
Re: Nano Banana Pro
#388Google has been stomping around like Godzilla this week, and this is the first time I decided to link my card to their AI studio. I had seen people saying that they gave up and went to another platform because it was "impossible to pay". I thought this was strange, but after trying to get a working API key for the past half hour, I see what they mean. Everything is set up, I see a message that says "You're using Paid…
Is that going to need AGI? Or maybe it will always be out of reach of our silicon overlords and require human input.
Re: Nano Banana Pro
#389Re: Nano Banana Pro
#390Earlier quoted context omitted.
Your use case doesn't even make sense. What customers are clamoring for that feature? I doubt any paying customer in the market for (that product) cares. If the law cares, the law has tools to inquire. All of this is trivially easy to circumvent ceremony. Google is doing this to deflect litigation and to preserve their brand in the face of negative press. They'll do this (1) as long as they're the market leader, (2)…
> What customers are clamoring for that feature? If the law cares, the law has tools to inquire. How can they distinguish from real people exploited to AI models autogenerating everything? I mean right now this is possible, largely because a lot of the AI videos have shortcomings. But imagine in 5 years from now on ...
Watermarking by compliant models doesn't help this much because (1) models without watermarking exist and can continue to be developed (especially if absence of a watermark is treated as a sign of authenticity), so you cannot rely on AI fakery being watermarked, and (2) AI models can be used for video-to-video generation without changing much of the source, so you can't rely on something accurately watermarked as "AI-generated" not being based in actual exploitation.
Now, if the watermarking includes provenance information, and you require certain types of content to be watermarked not just as AI using a known watermarking system, but by a registered AI provider with regulated input data safety guardrails and/or retention requirements, and be traceable to a registered user, and...
Well, then it does something when it is present, largely by creating a new content gatekeepiing cartel.