Earlier quoted context omitted.
I don't know if that's so much a mistake as it is ambiguity though? To me, using the viewer's perspective in this case seems totally reasonable. Does it still use the viewer's perspective if the prompt specifies "Put a strawberry in the _patient's left eye_"? If it does, then you're onto something. Otherwise I completely disagree with this.
“Eye on the left” is different from “the left eye”. First can be ambiguous, second really isn’t.
Nano Banana Pro
331–340 of 718 posts
Re: Nano Banana Pro
#332This thing's ability to produce entire infographics from a short prompt is really impressive, especially since it can run extra Google searches first. I tried this prompt: Infographic explaining how the Datasette open source project works Here's the result: https://simonwillison.net/2025/Nov/20/nano-banana-pro/#creat...
"An infographic explaining how player.html works (from the player.html project on Github). https://github.com/pseudosavant/player.html"
And then it made one formatted for social: "Change it to be an infographic formatted to fit on Instagram as a 1:1 square image."
Re: Nano Banana Pro
#333Earlier quoted context omitted.
this is pretty cool! have you found success with image editing in nano banana - i mean photoshop-like stuff. from your article i seem to wonder if nano banana is good for editing versus generating new images.
That IS the use-case for Nano Banana (as opposed to pure generative like Imagen4). In my benchmarks, Nano-Banana scores a 7 out of 12. Seedream4 managed to outpace it, but Seedream can also introduce slight tone mapping variations. NB is the gold standard for highly localized edits. Comparisons of Seedream4, NanoBanana, gpt-image-1, etc. https://genai-showdown.specr.net/image-editing
https://static.simonwillison.net/static/2025/brown-mms-remov...
Re: Nano Banana Pro
#334Maybe I'm an obscure case, but I'm just not sure what I'd use an image generation model for. For people that use them (regularly or not), what do you use them for?
Recently I also started using image generation models to explore ideas for what changes to make in my paintings. Although generally I don't like the suggestions it makes, sometimes it provides me with creative ideas of techniques that are worth experimenting with.
One way to approach thinking about it is that it's good for exploring permutations in an idea-space.
Re: Nano Banana Pro
#335Earlier quoted context omitted.
>This is a GREAT example of the (not so) subtle mistakes AI will make in image generation, or code creation, or your future knee surgery. The mistake is in the prompting (not enough information). The AI did the best it could "What's the biggest known planet" "Jupiter" "NO I MEANT IN THE UNIVERSE!"
No, this is squarely on the AI. A human would know what you mean without specific instructions.
Re: Nano Banana Pro
#336Earlier quoted context omitted.
No, this is squarely on the AI. A human would know what you mean without specific instructions.
If the instructions were actually specific, e.g. Put a blackberry in its right eye socket , then yes, most humans would know what that meant. But the instructions were not that specific: in the right eye socket
Re: Nano Banana Pro
#337Earlier quoted context omitted.
No, this is squarely on the AI. A human would know what you mean without specific instructions.
Seems like you're making a judgment based on your own experience, but as another commenter pointed out, it was wrong. There are plenty of us out there who would confirm, because people are too flawed to trust. Humans double/triple check, especially under higher stakes conditions (surgery). Heck, humans are so flawed, they'll put the things in the wrong eye socket even knowing full well exactly where they should go -…
Re: Nano Banana Pro
#338Google has been stomping around like Godzilla this week, and this is the first time I decided to link my card to their AI studio. I had seen people saying that they gave up and went to another platform because it was "impossible to pay". I thought this was strange, but after trying to get a working API key for the past half hour, I see what they mean. Everything is set up, I see a message that says "You're using Paid…
Re: Nano Banana Pro
#339Something I find weird about AI image generation models is that even though they no longer produce weird "artifacts" that give away that the fact that it was AI generated, you can still recognize that it's AI due to stylistic choices. Not all examples they gave were like this. The example they gave of the word "Typography" would have fooled me as human-made. The infographics stood out though. I would have immediately…
It's a bit odd to say, but another big clue identifying something as AI-generated is that it simply looks "too good" for what it is being used for. If I see a little info graphic demonstrating something relatively mundane, and it has nice 3D rendered characters or graphical elements, at this point it's basically guaranteed to be AI, because you just sort of intuitively know when something would've justified the human…
Of course now a lot of them have learned the lesson and it's much harder to tell.
[0]: I know, I know...
Re: Nano Banana Pro
#340Earlier quoted context omitted.
Some days it feels like I'm the only hacker left who doesn't want government mandated watermarking in creative tools. Were politicians 20 years ago as overreative they'd have demanded Photoshop leave a trace on anything it edited. The amount of moral panic is off the charts. It's still a computer, and we still shouldn't trust everything we see. The fundamentals haven't changed.
Easy to say until it impacts you in a bad way: https://www.nbcnews.com/tech/tech-news/ai-generated-evidence... > “My wife and I have been together for over 30 years, and she has my voice everywhere,” Schlegel said. “She could easily clone my voice on free or inexpensive software to create a threatening message that sounds like it’s from me and walk into any courthouse around the country with that recording.” > “The j…