Okay results are in for GenAI Showdown with the new gpt-image 1.5 model for the editing portions of the site! https://genai-showdown.specr.net/image-editing Conclusions - OpenAI has always had some of the strongest prompt understanding alongside the weakest image fidelity. This update goes some way towards addressing this weakness. - It's leagues better at making localized edits without altering the entire image's ae…
This showdown benchmark was and still is great, but an enormous grain of salt should be added to any model that was released after the showdown benchmark itself. Maybe everyone has a different dose of skepticism. Personally I'm not even looking at results for models that were released after the benchmark, for all this tells us, they might as well be one-trick ponies that only do well in the benchmark. It might be too…
GPT Image 1.5
161–170 of 272 posts
Re: GPT Image 1.5
#162I get the tech implementation is amazing, I wonder if it takes away from genuineness of events, like the Astronaut photo, I get it's just a joke/funny too but it's like a photo of you in a supercar vs. actually buying one. Or fake AI companions vs. real people. Beauty filters/skinny filters vs. actually being healthy.
hopefully this leads to greater importance of seeing things with your own wetware.
Re: GPT Image 1.5
#163Was it ever explained or understood why ChatGPT Images always has (had?) that yellow cast?
Not always, it started at a very specific point. Studio Ghibli craze + reinforcement learning on the likes.
Re: GPT Image 1.5
#164Earlier quoted context omitted.
> the only model that legitimately passed the Giraffe prompt. 10 years ago I would have considered that sentence satire. Now it allegedly means something. Somehow it feels like we’re moving backwards.
> Somehow it feels like we’re moving backwards. I don't understand why everyone isn't in awe of this. This is legitimately magical technology. We've had 60+ years of being able to express our ideas with keyboards. Steve Jobs' "bicycle of the mind". But in all this time we've had a really tough time of visually expressing ourselves. Only highly trained people can use Blender, Photoshop, Illustrator, etc. whereas almos…
It’s just a tool. It’s not a world-changing tech. It’s a tool.
Re: GPT Image 1.5
#165Okay results are in for GenAI Showdown with the new gpt-image 1.5 model for the editing portions of the site! https://genai-showdown.specr.net/image-editing Conclusions - OpenAI has always had some of the strongest prompt understanding alongside the weakest image fidelity. This update goes some way towards addressing this weakness. - It's leagues better at making localized edits without altering the entire image's ae…
Z-image was released recently and that's what /r/StableDiffusion all talks about these days. Consider adding that too. It is very good quality for its size (Requires only 6 or 8 gigs of ram).
Re: GPT Image 1.5
#166Okay results are in for GenAI Showdown with the new gpt-image 1.5 model for the editing portions of the site! https://genai-showdown.specr.net/image-editing Conclusions - OpenAI has always had some of the strongest prompt understanding alongside the weakest image fidelity. This update goes some way towards addressing this weakness. - It's leagues better at making localized edits without altering the entire image's ae…
Re: GPT Image 1.5
#167Earlier quoted context omitted.
> the only model that legitimately passed the Giraffe prompt. 10 years ago I would have considered that sentence satire. Now it allegedly means something. Somehow it feels like we’re moving backwards.
> Somehow it feels like we’re moving backwards. I don't understand why everyone isn't in awe of this. This is legitimately magical technology. We've had 60+ years of being able to express our ideas with keyboards. Steve Jobs' "bicycle of the mind". But in all this time we've had a really tough time of visually expressing ourselves. Only highly trained people can use Blender, Photoshop, Illustrator, etc. whereas almos…
Re: GPT Image 1.5
#168Okay results are in for GenAI Showdown with the new gpt-image 1.5 model for the editing portions of the site! https://genai-showdown.specr.net/image-editing Conclusions - OpenAI has always had some of the strongest prompt understanding alongside the weakest image fidelity. This update goes some way towards addressing this weakness. - It's leagues better at making localized edits without altering the entire image's ae…
So when you say "X attempts" what does that mean? You just start a new chat with the same exact prompt and hope for a different result?
In addition to giving models multiple attempts to generate an image, we also write several variations of each prompt. This helps prevent models from getting stuck on particular keywords or phrases, which can happen depending on their training data. For example, while “hippity hop” is a relatively common name for the ball-riding toy, it’s also known as a “space hopper.” In some cases, we may even elaborate and provide the model with a dictionary-style definition of more esoteric terms.
This is why providing an “X Attempts” metric is so important. It serves as a rough measure of how “steerable” a given model is - or put another way how much we had to fight with the model in order for it to consistently follow the prompt’s directives.
Re: GPT Image 1.5
#169Re: GPT Image 1.5
#170Okay results are in for GenAI Showdown with the new gpt-image 1.5 model for the editing portions of the site! https://genai-showdown.specr.net/image-editing Conclusions - OpenAI has always had some of the strongest prompt understanding alongside the weakest image fidelity. This update goes some way towards addressing this weakness. - It's leagues better at making localized edits without altering the entire image's ae…