I don't understand the excitement around generating and/or watching AI-produced videos. To me it's probably the single most uninteresting and boring thing related to AI that I can think of. What is the appeal?
Nano Banana Pro
401–410 of 718 posts
Re: Nano Banana Pro
#402Something I find weird about AI image generation models is that even though they no longer produce weird "artifacts" that give away that the fact that it was AI generated, you can still recognize that it's AI due to stylistic choices. Not all examples they gave were like this. The example they gave of the word "Typography" would have fooled me as human-made. The infographics stood out though. I would have immediately…
We are just very sharp when it comes to seeing small differences in images. I'm reminded of when the air force decided to create a pilot seat that worked for everyone. They took the average body dimensions of all their recruits and designed a seat to fit the average. It turned out, the seat fit none of their recruits. [1] I think AI image generation is a lot like this. When you train on all images, you get to this we…
But it is undeniable that AI images do have an “average” feel to them. What causes this? What is the space over which AI is taking an average to produce its output? One possible answer is that a finite model size means that the model can only explore image space with a limited resolution, and as models get bigger/better they can average over a smaller and smaller portion of this space, but it is always limited.
But that raises the question of why models don't just naturally land on a point in image space. Is this just a limitation of training, which punishes big failures more strongly than it rewards perfection? Or is there something else at play here that's preventing models from landing directly on a “real” image?
Re: Nano Banana Pro
#403Google has been stomping around like Godzilla this week, and this is the first time I decided to link my card to their AI studio. I had seen people saying that they gave up and went to another platform because it was "impossible to pay". I thought this was strange, but after trying to get a working API key for the past half hour, I see what they mean. Everything is set up, I see a message that says "You're using Paid…
100% this. I am using the pro/max plans on both claude and openai. Would love to experiment with gemini but paying is next to impossible. Why do i need the risk of a full blown gcp project just to test gemini. No thx.
Re: Nano Banana Pro
#404Earlier quoted context omitted.
>> - Put a strawberry in the left eye socket. >>- Put a blackberry in the right eye socket. >> All five of the edits are implemented correctly This is a GREAT example of the (not so) subtle mistakes AI will make in image generation, or code creation, or your future knee surgery. The model placed the specified items in the eye sockets based on the viewers left/right; when we talk relative in this scenario we usually (…
>This is a GREAT example of the (not so) subtle mistakes AI will make in image generation, or code creation, or your future knee surgery. The mistake is in the prompting (not enough information). The AI did the best it could "What's the biggest known planet" "Jupiter" "NO I MEANT IN THE UNIVERSE!"
Re: Nano Banana Pro
#405Earlier quoted context omitted.
It doesn't affect your point but technically since the IAU are insane, exoplanets aren't technically planets and Jupiter is the largest planet in the universe.
I suppose it was too much to hope that chatbots could be trained to avoid pointless pedantry.
Re: Nano Banana Pro
#406Google has been stomping around like Godzilla this week, and this is the first time I decided to link my card to their AI studio. I had seen people saying that they gave up and went to another platform because it was "impossible to pay". I thought this was strange, but after trying to get a working API key for the past half hour, I see what they mean. Everything is set up, I see a message that says "You're using Paid…
Re: Nano Banana Pro
#407Alright results are in! I've re-run all my editing based adherence related prompts through Nano Banana Pro. NB Pro managed to successfully pass SHRDLU, the M&M Van Halen test (as verified independently by Simon), and the Scorpio street test - all of which the original NB failed. Model results 1. Nano Banana Pro: 10 / 12 2. Seedream4: 9 / 12 3. Nano Banana: 7 / 12 4. Qwen Image Edit: 6 / 12 https://genai-showdown.spec…
Re: Nano Banana Pro
#408Google has been stomping around like Godzilla this week, and this is the first time I decided to link my card to their AI studio. I had seen people saying that they gave up and went to another platform because it was "impossible to pay". I thought this was strange, but after trying to get a working API key for the past half hour, I see what they mean. Everything is set up, I see a message that says "You're using Paid…
You can use it also in Gemini.
Tell me the model it's using. It's as if Google is trying to unburden me with the knowledge of what model does what but it's just making things more confusing.
Oh, and setting up AI Studio is a mess. First I have to create a project. Then an API key. Then I have to link the API key to the project. Then I have to link the project to the chat session... Come on, Google.
Re: Nano Banana Pro
#409Earlier quoted context omitted.
How do you disagree with having a right and a left hand?
GP is using right as in “correct”, not directionality.
If you are facing a wall-plate with two power sockets on it side by side and you are telling someone to plug something in, which one would be "the right socket", and which would be "the left socket"?
If above the wall-plate is a photo of a person and you are someone to draw a tattoo on the photo, which is "the right arm" and which is "the left arm"?
Same wording, different expectation.
Re: Nano Banana Pro
#410Alright results are in! I've re-run all my editing based adherence related prompts through Nano Banana Pro. NB Pro managed to successfully pass SHRDLU, the M&M Van Halen test (as verified independently by Simon), and the Scorpio street test - all of which the original NB failed. Model results 1. Nano Banana Pro: 10 / 12 2. Seedream4: 9 / 12 3. Nano Banana: 7 / 12 4. Qwen Image Edit: 6 / 12 https://genai-showdown.spec…
thanks, I love your website. Are you planning to do NB Pro for the text-to-image benchmark too?
I'll try to have the generative comparisons for NB Pro up later this afternoon once I catch my breath.