Live data from Hacker News

Gemini 2.5 Flash Image

developers.googleblog.com

351–360 of 504 posts

Re: Gemini 2.5 Flash Image

#353

I digitised our family photos but a lot of them were damaged (shifted colours, spills, fingerprints on film, spots) that are difficult to correct for so many images. I've been waiting for image gen to catch up enough to be able to repair them all in bulk without changing details, especially faces. This looks very good at restoring images without altering details or adding them where they are missing, so it might fina…

I've been waiting for image gen to catch up enough to be able to repair them all in bulk without changing details, especially faces. I've been waiting for that, too. But I'm also not interesting in feeding my entire extended family's visual history into Google for it to monetize. It's wrong for me to violate their privacy that way, and also creepy to me. Am I correct to worry that any pictures I send into this system…

You're looking for Flux Kontext, a model you can run yourself offline on a high end consumer GPU. Performance and accuracy are okay, not groundbreaking, but probably enough for many needs.

Re: Gemini 2.5 Flash Image

#354

Unfortunately, it suffers from the same safetyism than other many releases. Half of the prompts get rejected. How can you have character consistency if the model is forbidden from editing any human. And most of my photo editing involves humans, so basically this is just a useless product. I get that Google doesn't want to be responsible for deep fake advances, but that seems inevitable, so this is just slightly delay…

I was using Veo two days ago when video generations were free. I removed all words that sounded even remotely bad, but it still refused. Eventually gave up but now I'm thinking it's because I tried to generate myself

Re: Gemini 2.5 Flash Image

#355

I've updated the GenAI Image comparison site (which focuses heavily on strict text-to-image prompt adherence) to reflect the new Google Gemini 2.5 Flash model (aka nano-banana). https://genai-showdown.specr.net This model gets 8 of the 12 prompts correct and easily comes within striking distance of the best-in-class models Imagen and gpt-image-1 and is a significant upgrade over the old Gemini Flash 2.0 model. The re…

I really like your site.

Do you know of any similar sites that that compares how well the various models can adhere to a style guide? Perhaps you could add this?

I.e. pride the model with a collection of drawings in a single style, then follow prompts and generate images in the same style?

For example if you wanted to illustrate a book, and have all the illustrations look like they were from the same artists.

Re: Gemini 2.5 Flash Image

#356

   All images created or edited with Gemini 2.5 Flash Image will include an invisible SynthID digital watermark, so they can be identified as AI-generated or edited.
Obviously I understand what is the purpose and the good intention, but I think sad to see that we are not not anymore responsible adults but big corps deciding for us what we can and what we cannot do. Snitching on your back.

Re: Gemini 2.5 Flash Image

#357

Unfortunately, it suffers from the same safetyism than other many releases. Half of the prompts get rejected. How can you have character consistency if the model is forbidden from editing any human. And most of my photo editing involves humans, so basically this is just a useless product. I get that Google doesn't want to be responsible for deep fake advances, but that seems inevitable, so this is just slightly delay…

I have an old photo of my girlfriend with her cousin when they were young, wearing Christmas dresses in front of the tree, not long before they were separated to other sides of the world for decades now. The photo is itself low quality on top of the photo itself being physically beat up.

So far no model is willing to clean it up :/

Re: Gemini 2.5 Flash Image

#358

All images created or edited with Gemini 2.5 Flash Image will include an invisible SynthID digital watermark, so they can be identified as AI-generated or edited. Obviously I understand what is the purpose and the good intention, but I think sad to see that we are not not anymore responsible adults but big corps deciding for us what we can and what we cannot do. Snitching on your back.

dont worry, you can just screenshot the image to get rid of the watermark

Re: Gemini 2.5 Flash Image

#359
There is one thing Gemini 2.5 Flash Image can do that no other edit model can do: incorporate multiple images simultaneously without shenanigans due to its multimodality, e.g. for Flux Kontext, if you want to "put the person in the first image into the second image", you have to concatenate them pre-VAE which can be unwieldly, but this model doesn't have that issue. You can even incorporate more than two images, but that may cause too much chaos.

In quick testing, prompt adherence does appear to be much better for massive prompt and the syntatic sugar does appear to be more effective. And there are other tricks not covered which I suspect may allow more control, but I'm still testing.

Given that generations are at the same price as its competitors, this model will shake things up.

Re: Gemini 2.5 Flash Image

#360

All images created or edited with Gemini 2.5 Flash Image will include an invisible SynthID digital watermark, so they can be identified as AI-generated or edited. Obviously I understand what is the purpose and the good intention, but I think sad to see that we are not not anymore responsible adults but big corps deciding for us what we can and what we cannot do. Snitching on your back.

I don't see the problem as it's not like you're forced to use their image generation model.
Post reply on HN