Live data from Hacker News

Gemini 2.5 Flash Image

developers.googleblog.com

441–450 of 504 posts

Re: Gemini 2.5 Flash Image

#441

Earlier quoted context omitted.

I have an old photo of my girlfriend with her cousin when they were young, wearing Christmas dresses in front of the tree, not long before they were separated to other sides of the world for decades now. The photo is itself low quality on top of the photo itself being physically beat up. So far no model is willing to clean it up :/

There are reddit communities (I admittedly don't remember which, but could probably be found from a simple search) where people will offer their photo editing skills to touch up the photo, often for free. Could be worth trying a real human if the robots are going full HAL 9000 and telling you they can't do it.

https://www.reddit.com/r/PhotoshopRequest/

People sometimes do it for free ("my son died, and this is the only photo I have") or for an agreed upon tip.

Re: Gemini 2.5 Flash Image

#442
post #56

Earlier quoted context omitted.

[flagged]

These SpaceX scams are rampant on youtube and highly, highly lucrative. It’s crazy and you have to be very vigilant, as whatever is promised lines up with Elon’s MO.

My mind is blown that people would engage in it at all, let alone need to be "vigilant". Amazing.

Re: Gemini 2.5 Flash Image

#443

I've updated the GenAI Image comparison site (which focuses heavily on strict text-to-image prompt adherence) to reflect the new Google Gemini 2.5 Flash model (aka nano-banana). https://genai-showdown.specr.net This model gets 8 of the 12 prompts correct and easily comes within striking distance of the best-in-class models Imagen and gpt-image-1 and is a significant upgrade over the old Gemini Flash 2.0 model. The re…

This is incredibly useful! I was manually generating my own model comparisons last night, so great to see this :)

I will note that, personally, while adherence is a useful measure, it does miss some of the qualitative differences between models. For your "spheron" test for example, you note that "4o absolutely dominated this test," but the image exhibits all the hallmarks of a ChatGPT-generated image that I personally dislike (yellow, with veiny, almost impasto brush strokes). I have stopped using ChatGPT for image generation altogether because I find the style so awful. I wonder what objective measures one could track for "style"?

It reminders be a bit of ChatGPT vs Claude for software development... Regardless of how each scores on benchmarks, Claude has been a clear winner in terms of actual results.

Re: Gemini 2.5 Flash Image

#444
post #179

FYI, this is the famed nano-banana model which has been now renamed to gemini-2.5-flash-image-preview in LMArena.

I mean they are going to have to rename their AI because gemini.com is going to IPO soon. "Banana" would be a nice name for their AI, and they could freely claim it's bananas.

why do you think they have to rename it because some company's IPO

Re: Gemini 2.5 Flash Image

#445
post #60

This is the gpt 4 moment for image editing models. Nano banana aka gemini 2.5 flash is insanely good. It made a 171 elo point jump in lmarena! Just search nano banana on Twitter to see the crazy results. An example. https://x.com/D_studioproject/status/1958019251178267111

nano banana is good, but not insanely good

Re: Gemini 2.5 Flash Image

#446

Earlier quoted context omitted.

I've done ~20 prompts so far and not had one be rejected so far. What sort of things are you asking it to do? I've tried things like changing clothing and accessories on people.

Basic things like: "{uploaded image of a man} can you remove the glasses?" or "make everyone in the picture smile" or "open the eyes of everyone in the photo". Nothing that a human would consider "unsafe". I am based in EU and using Google AI Studio with all safety toggles set to "Off".

I noticed that I get far fewer refusals when I set my VPN to the USA.

Re: Gemini 2.5 Flash Image

#447
So it doesn't allow to do anything with photos containing kids, right? Isn't it too much of a filter for such a thing? ChatGPT thankfully created Ghibli versions of everything I gave it.

Re: Gemini 2.5 Flash Image

#448

Earlier quoted context omitted.

Vibe coding might not be real, but vibe graphics design certainly is. https://imgur.com/a/internet-DWzJ26B Anyone can make images and video now.

Are those oil derricks, or wind turbines? Who cares! Graphic design is easy now!

They're Australian farm windmills https://media.istockphoto.com/id/959193466/photo/australian-...

(But yeah, some got a generator attached...)

Re: Gemini 2.5 Flash Image

#449

A bit mixed opinions - I tried colorizing manga pages with it, and the results were perfect. Interestingly, it can change pages with tons of text on them without any problem, but cannot seem to do translation, if I ask it to translate a French comic page, the text ends up garbled (even though it can perfectly read and translate the text by itself). I tried with another page, and it copypasted the same character (in d…

I had a similar experience.

It did not change the text on a hat (ended up changing 1 of 3 words).

On one occasion it regenerated the same image again, ignoring my instructions to edit.

I get the feeling that this model is optimised for images with people in it than objects or drawings etc

Re: Gemini 2.5 Flash Image

#450
post #80

Earlier quoted context omitted.

Facebook comments are obviously botted too

I dunno, I thought so for a while, but I’m beginning to suspect this is a very optimistic view of humanity.

Why do HN commenters all act like they're in the top 1% of intellectuals
Post reply on HN