Live data from Hacker News

Gemini 2.5 Flash Image

developers.googleblog.com

21–30 of 504 posts

Re: Gemini 2.5 Flash Image

#22

What is the difference between Gemini Flash Image models and the Imagen models?

Imagen is a diffusion text to image model. You write some text that describes your image, you get an image out and that's it.

Flash Image is an image (and text) predicting large language model. In a similar fashion to how trained LLMs can manipulate/morph text, this can do that for images as well. Things like style transfer, character consistency etc.

You can communicate with it in a way you can't for imagen, and it has a better overall world understanding.

Re: Gemini 2.5 Flash Image

#23
post #2

Anyone know how it handles '1920s nazi officer'? They stopped doing humans for a while but now I see they're back so I wonder how they're handling the criticism they got from that

What is a "1920s nazi officer" what do they look like?

brown uniform, red armband with swastika was the usual SA look in the 1920s.

Re: Gemini 2.5 Flash Image

#25
post #8

Earlier quoted context omitted.

when giving more context it replied: """ Unfortunately, I can't generate images of people. My purpose is to be helpful and harmless, and creating realistic images of humans can be misused in ways that are harmful. This is a safety policy that helps prevent the generation of deepfakes, non-consensual imagery, and other problematic content. If you'd like to try a different image prompt, I can help you create images of…

What a weird rejection. You have to scroll pretty far in the article to see an example output that doesn't have a realistic depiction of a person.

The rejection message doesn’t seem to be accurate. I tried “happy person” as a prompt in AI Studio and it generated a happy human without any complaints.

It’s possible that they relaxed the safety filtering to allow humans but forgot to update the error message.

Re: Gemini 2.5 Flash Image

#26
I have a certain use case for such image generators. Feed them an entire news article I fetch from bbc and ask it to create an image to accompany the article. Thus far only midjourney managed to understand context. And now this, which is even more impressive. We live in interesting times.

Re: Gemini 2.5 Flash Image

#29
post #25

Earlier quoted context omitted.

What a weird rejection. You have to scroll pretty far in the article to see an example output that doesn't have a realistic depiction of a person.

The rejection message doesn’t seem to be accurate. I tried “happy person” as a prompt in AI Studio and it generated a happy human without any complaints. It’s possible that they relaxed the safety filtering to allow humans but forgot to update the error message.

[deleted]
Post reply on HN