Live data from Hacker News

Gemini 2.5 Flash Image

developers.googleblog.com

121–130 of 504 posts

Re: Gemini 2.5 Flash Image

#121
post #60

This is the gpt 4 moment for image editing models. Nano banana aka gemini 2.5 flash is insanely good. It made a 171 elo point jump in lmarena! Just search nano banana on Twitter to see the crazy results. An example. https://x.com/D_studioproject/status/1958019251178267111

> This is the gpt 4 moment for image editing models. No it's not. We've had rich editing capabilities since gpt-image-1, this is just faster and looks better than the (endearingly? called) "piss filter". Flux Kontext, SeedEdit, and Qwen Edit are all also image editing models that are robustly capable. Qwen Edit especially. Flux Kontext and Qwen are also possible to fine tune and run locally. Qwen (and its video gen s…

In other words, this is the gpt 4 moment for image editing models.

Gpt4 isn't "fundamentally different" from gpt3.5. It's just better. That's the exact point the parent commenter was trying to make.

Re: Gemini 2.5 Flash Image

#122

Half the time I ask Gemini to generate some image it claims it doesn't have the capability. And in general I've felt it's so hard to actually use the features Google announce? Like, a third of them is in one product, some in another which I can't use, and no idea what or where I should pay to get access. So confusing.

Yeah, in fact the website says "Try it in Gemini" and I'm not sure if I'm already trying it or not - if I choose Gemini 2.5 Flash in the regular Gemini UI, I'm using this?

It’s going to be a messy rollout as usual. The web app (gemini.google.com) shows “Images with Imagen” for me under tools for 2.5 flash but I just tried a few image edits and remixes in the iOS app and it looks like it’s been updated to this model.

Re: Gemini 2.5 Flash Image

#125

Earlier quoted context omitted.

> This is the gpt 4 moment for image editing models. No it's not. We've had rich editing capabilities since gpt-image-1, this is just faster and looks better than the (endearingly? called) "piss filter". Flux Kontext, SeedEdit, and Qwen Edit are all also image editing models that are robustly capable. Qwen Edit especially. Flux Kontext and Qwen are also possible to fine tune and run locally. Qwen (and its video gen s…

In other words, this is the gpt 4 moment for image editing models. Gpt4 isn't "fundamentally different" from gpt3.5. It's just better. That's the exact point the parent commenter was trying to make.

did you see the generated pic demis posted on X? it looks like slop from 2 years ago. https://x.com/demishassabis/status/1960355658059891018

Re: Gemini 2.5 Flash Image

#126
post #85

I digitised our family photos but a lot of them were damaged (shifted colours, spills, fingerprints on film, spots) that are difficult to correct for so many images. I've been waiting for image gen to catch up enough to be able to repair them all in bulk without changing details, especially faces. This looks very good at restoring images without altering details or adding them where they are missing, so it might fina…

Do you happen to know some software to repair/improve video files? I'm in the process of digitalizing a couple of Video 2000 and VHS casettes of childhood memories of my mom who start suffering from dementia. I have a pretty streamlined setup for digitalizing the videos but I'd like to improve the quality a bit.

I didn't do any videos, just pictures, but considering how little I found for pictures I doubt you'll find much

Re: Gemini 2.5 Flash Image

#127
post #118
post #108

Earlier quoted context omitted.

Hmm, I think the hype is mainly for image editing, not generating. Although note I haven't used it! How are you testing it?

I tested it with two prompts: // In this one, Gemini doesn't understand what "cinematic" is "A cinematic underwater shot of a turtle gracefully swimming in crystal-clear water [...]" // In this one, the reflection in the water in the background has different buildings "A modern city where raindrops fall upward into the clouds instead of down, pedestrians calmly walking [...]" Midjourney created both perfectly.

As others have said, this is an image editing model.

Editing models do not excel at aesthetic, but they can take your Midjourney image, adjust the composition, and make it perfect.

These types of models are the Adobe killer.

Re: Gemini 2.5 Flash Image

#128

Earlier quoted context omitted.

It seems like every combination of "nano banana" is registered as a domain with their own unique UI for image generation... are these all middle actors playing credit arbitrage using a popular model name?

I'd assume they are just fake, take your money and use a different model under the hood. Because they already existed before the public release. I doubt that their backend rolled the dice on LMArena until nano-banana popped up. And that was the only way to use it until today.

Agreed, I didn't mean to imply that they were even attempting to run the actual nano banana, even through LMarena.

There is a whole spectrum of potential sketchiness to explore with these, since I see a few "sign in with Google" buttons that remind me of phishing landing pages.

Post reply on HN