Live data from Hacker News

GPT Image 1.5

openai.com

161–170 of 272 posts

Re: GPT Image 1.5

#161

Okay results are in for GenAI Showdown with the new gpt-image 1.5 model for the editing portions of the site! https://genai-showdown.specr.net/image-editing Conclusions - OpenAI has always had some of the strongest prompt understanding alongside the weakest image fidelity. This update goes some way towards addressing this weakness. - It's leagues better at making localized edits without altering the entire image's ae…

This showdown benchmark was and still is great, but an enormous grain of salt should be added to any model that was released after the showdown benchmark itself. Maybe everyone has a different dose of skepticism. Personally I'm not even looking at results for models that were released after the benchmark, for all this tells us, they might as well be one-trick ponies that only do well in the benchmark. It might be too…

I think training image models to pass these very specific tests correctly will be very difficult for any of these companies. How would they even do that?

Re: GPT Image 1.5

#162
post #126

I get the tech implementation is amazing, I wonder if it takes away from genuineness of events, like the Astronaut photo, I get it's just a joke/funny too but it's like a photo of you in a supercar vs. actually buying one. Or fake AI companions vs. real people. Beauty filters/skinny filters vs. actually being healthy.

the next generation of humans growing up will not even care whether media is real or not any more. The saturation of AI content and FUD around real content is going to blur the lines to the extent that there's no point even caring about it. And it's an intractable problem.

hopefully this leads to greater importance of seeing things with your own wetware.

Re: GPT Image 1.5

#163
post #112

Was it ever explained or understood why ChatGPT Images always has (had?) that yellow cast?

Not always, it started at a very specific point. Studio Ghibli craze + reinforcement learning on the likes.

The Studio Ghibli craze started with the initial release of images in ChatGPT, and the yellow filter has always existed even at that time. They did not make changes to the model as a result of RL (until pontentially today, with a new model)

Re: GPT Image 1.5

#164

Earlier quoted context omitted.

> the only model that legitimately passed the Giraffe prompt. 10 years ago I would have considered that sentence satire. Now it allegedly means something. Somehow it feels like we’re moving backwards.

> Somehow it feels like we’re moving backwards. I don't understand why everyone isn't in awe of this. This is legitimately magical technology. We've had 60+ years of being able to express our ideas with keyboards. Steve Jobs' "bicycle of the mind". But in all this time we've had a really tough time of visually expressing ourselves. Only highly trained people can use Blender, Photoshop, Illustrator, etc. whereas almos…

You basically described magic mushrooms, where the description came from you while high on magic mushrooms.

It’s just a tool. It’s not a world-changing tech. It’s a tool.

Re: GPT Image 1.5

#165

Okay results are in for GenAI Showdown with the new gpt-image 1.5 model for the editing portions of the site! https://genai-showdown.specr.net/image-editing Conclusions - OpenAI has always had some of the strongest prompt understanding alongside the weakest image fidelity. This update goes some way towards addressing this weakness. - It's leagues better at making localized edits without altering the entire image's ae…

Z-image was released recently and that's what /r/StableDiffusion all talks about these days. Consider adding that too. It is very good quality for its size (Requires only 6 or 8 gigs of ram).

I've actually done a bit of preliminary testing with ZiT. I'm holding off on adding it to the official GenAI site until the base and edit models have been released since the Turbo model is pretty heavily distilled.

https://mordenstar.com/other/z-image-turbo

Re: GPT Image 1.5

#166

Okay results are in for GenAI Showdown with the new gpt-image 1.5 model for the editing portions of the site! https://genai-showdown.specr.net/image-editing Conclusions - OpenAI has always had some of the strongest prompt understanding alongside the weakest image fidelity. This update goes some way towards addressing this weakness. - It's leagues better at making localized edits without altering the entire image's ae…

So when you say "X attempts" what does that mean? You just start a new chat with the same exact prompt and hope for a different result?

Re: GPT Image 1.5

#167

Earlier quoted context omitted.

> the only model that legitimately passed the Giraffe prompt. 10 years ago I would have considered that sentence satire. Now it allegedly means something. Somehow it feels like we’re moving backwards.

> Somehow it feels like we’re moving backwards. I don't understand why everyone isn't in awe of this. This is legitimately magical technology. We've had 60+ years of being able to express our ideas with keyboards. Steve Jobs' "bicycle of the mind". But in all this time we've had a really tough time of visually expressing ourselves. Only highly trained people can use Blender, Photoshop, Illustrator, etc. whereas almos…

I'm struggling to see the benefits. All I see people using this for is generating slop for work presentations, and misleading people on social media. Misleading might be understating it too. It's being used to create straight up propaganda and destruction of the sense of reality.

Re: GPT Image 1.5

#168

Okay results are in for GenAI Showdown with the new gpt-image 1.5 model for the editing portions of the site! https://genai-showdown.specr.net/image-editing Conclusions - OpenAI has always had some of the strongest prompt understanding alongside the weakest image fidelity. This update goes some way towards addressing this weakness. - It's leagues better at making localized edits without altering the entire image's ae…

So when you say "X attempts" what does that mean? You just start a new chat with the same exact prompt and hope for a different result?

All images are generated using independent, separate API calls. See the FAQ at the bottom under “Why is the number of attempts seemingly arbitrary?” and “How are the prompts written?” for more detail, but to quickly summarize:

In addition to giving models multiple attempts to generate an image, we also write several variations of each prompt. This helps prevent models from getting stuck on particular keywords or phrases, which can happen depending on their training data. For example, while “hippity hop” is a relatively common name for the ball-riding toy, it’s also known as a “space hopper.” In some cases, we may even elaborate and provide the model with a dictionary-style definition of more esoteric terms.

This is why providing an “X Attempts” metric is so important. It serves as a rough measure of how “steerable” a given model is - or put another way how much we had to fight with the model in order for it to consistently follow the prompt’s directives.

Re: GPT Image 1.5

#170

Okay results are in for GenAI Showdown with the new gpt-image 1.5 model for the editing portions of the site! https://genai-showdown.specr.net/image-editing Conclusions - OpenAI has always had some of the strongest prompt understanding alongside the weakest image fidelity. This update goes some way towards addressing this weakness. - It's leagues better at making localized edits without altering the entire image's ae…

This leaderboard feels incredibly accurate given my own experience.
Post reply on HN