Like most image generators, it didn’t pass the piano keyboard test. (Black keys are wrong.) https://aistudio.google.com/app/prompts?state=%7B%22ids%22:%...
The selling point of this model really seems to be it's consistency between generations rather than it's raw generating ability. for instance: https://aistudio.google.com/app/prompts/1gTG-D92MyzSKaKUeBu2...
Gemini 2.5 Flash Image
291–300 of 504 posts
Re: Gemini 2.5 Flash Image
#292Re: Gemini 2.5 Flash Image
#293Earlier quoted context omitted.
Agreed. I find myself alternating between Qwen Image Edit 20B, Kontext, and now Flash 2.5 depending on the situation and style. And of course, Flash isn't open-weights, so if you need more control / less censorship then you're SOL.
Has there been a sufficient indication to conclude these weights will not (now or ever) be released?
I don't think we can really answer the question if Flash will ever be released.
Re: Gemini 2.5 Flash Image
#294Re: Gemini 2.5 Flash Image
#295Earlier quoted context omitted.
The hype is about image editing, not pure text-to-image. Upload an input image, say what you want changed, get the output. That's the idea. Much better preservation of characters and objects.
Can it edit the photo at the original resolution? Most of my photos these days are 48MP and I don't want to lose a ton of resolution just to edit them.
Re: Gemini 2.5 Flash Image
#296Earlier quoted context omitted.
[flagged]
I don't think he's the gullible one, check their bio ;)
Re: Gemini 2.5 Flash Image
#297Re: Gemini 2.5 Flash Image
#298Re: Gemini 2.5 Flash Image
#299Earlier quoted context omitted.
[flagged]
This presumes that you're okay with giving the real Elon your wallet but not a fake Elon, but why?
Re: Gemini 2.5 Flash Image
#300I've updated the GenAI Image comparison site (which focuses heavily on strict text-to-image prompt adherence) to reflect the new Google Gemini 2.5 Flash model (aka nano-banana). https://genai-showdown.specr.net This model gets 8 of the 12 prompts correct and easily comes within striking distance of the best-in-class models Imagen and gpt-image-1 and is a significant upgrade over the old Gemini Flash 2.0 model. The re…
> Though fair warning, gpt-image-1 is borderline useless as an "editor" since it almost always changes the whole image instead of doing localized inpainting-style edits like Kontext, Qwen, or Nano-Banana. Came into this thread looking for this post. It's a great way to compare prompt adherence across models. Have you considered adding editing capabilities in a similar way given the recent trend of inpainting-style pr…
I've done some experimentation with Qwen and Kontext and been pretty impressed, but it would be nice to see some side by sides now that we have essentially three models that are capable of highly localized in-painting without affecting the rest of the image.