Live data from Hacker News

GenAI Image Editing Showdown

genai-showdown.specr.net

21–30 of 55 posts

Re: GenAI Image Editing Showdown

#21

There isn’t a date in the article, but I know I had read this months ago. And sure enough, wayback has the text-to-image page from April. But the image editing page linked at the top is more recent, and was added sometime in September. (And was presumably the intended link) I hadn’t read that page yet. Odd there is no dates, at first glance one might think the pages were made at the same time.

> There isn’t a date in the article

SEO guys convinced everyone that articles without dates do better on search engines. I hope both sides of their pillow is hot.

Re: GenAI Image Editing Showdown

#22

I'd assume that behind the scenes the models generate several passes and only show the user the best one, that would be smart, as to to make it seem their model is better than others Is also pretty obvious that the models have some built in prompt system rules that makes the final output a certain style. They seem very consistent It also looks like 40 has the temperature turned way down, to ensure max adherence, whil…

You can run some image models locally if you want to prove to yourself how well they can do with just a single generation from a prompt with no extra steps.

I've done this enough to suspect that most hosted image models don't increase their running costs to try and get better results through additional passes without letting the user know what they are doing.

Many of the LLM-driven models do implement a form of prompt rewriting though (since effectively prompting image models is really hard) - some notes on how DALL-E 3 did that here: https://simonwillison.net/2023/Oct/26/add-a-walrus/

Re: GenAI Image Editing Showdown

#23
Gpt4o shows the huge annoyance of the company/model being a moral judge of your requests and refusing quite often for anything negative.

It's like 1964 but corporate enforced. Now there are tasks that you are not allowed to do despite being legal.

In the same way, using gpt5 is now very unbearable to me as it almost always starts all responses of a conversation by things like: "Great question", "good observation worthy of an expert", "you totally right", "you are right to ask the question"...

Re: GenAI Image Editing Showdown

#24

Gpt4o shows the huge annoyance of the company/model being a moral judge of your requests and refusing quite often for anything negative. It's like 1964 but corporate enforced. Now there are tasks that you are not allowed to do despite being legal. In the same way, using gpt5 is now very unbearable to me as it almost always starts all responses of a conversation by things like: "Great question", "good observation wort…

Try some of the Chinese models. Much less restrictive. With some obvious exceptions.

Re: GenAI Image Editing Showdown

#25

Gpt4o shows the huge annoyance of the company/model being a moral judge of your requests and refusing quite often for anything negative. It's like 1964 but corporate enforced. Now there are tasks that you are not allowed to do despite being legal. In the same way, using gpt5 is now very unbearable to me as it almost always starts all responses of a conversation by things like: "Great question", "good observation wort…

People gave Altman shit for enabling NSFW in ChatGPT, but I see that as a step in the right direction. The right direction being: the one that leads to less corporate censorship.

>In the same way, using gpt5 is now very unbearable to me as it almost always starts all responses of a conversation by things like: "Great question"

User preference data is toxic. Doing RLHF on it gives LLM sycophancy brainrot. And by now, all major LLMs have it.

At least it's not 4o levels of bad - hope they learned that fucking lesson.

Re: GenAI Image Editing Showdown

#26

The "editing" showdown is very good. Introduced me to the Seedream model which i didn't know about until now. I don't fully understand the iterative methodology tho - they allow multiple attempts, which are judged by another multimodal llm? Won't they have limited accuracy in itself?

"LLMs judged by LLMs" is the industry standard. Can't put a human judge in a box and have him evaluate and rate a set of 7600 responses on demand.

Now, are LLM judges flawed? Obviously. But they are more shelf stable than humans, so it's easier to compare different results. And as long as you use an LLM judge as a performance thermometer and not a direct optimization target, you aren't going to be facing too many issues from that.

If you are using an LLM judge as a direct optimization target though? You'll see some funny things happen. Like GPT-5 prose. Which isn't even the weirdest it gets.

Re: GenAI Image Editing Showdown

#27
Is there any AI image generator/editor that is good at creating graphics with transparent background? Nano Banana and some others output a white grey checkered background (fake transparency).

Re: GenAI Image Editing Showdown

#29

Gpt4o shows the huge annoyance of the company/model being a moral judge of your requests and refusing quite often for anything negative. It's like 1964 but corporate enforced. Now there are tasks that you are not allowed to do despite being legal. In the same way, using gpt5 is now very unbearable to me as it almost always starts all responses of a conversation by things like: "Great question", "good observation wort…

People gave Altman shit for enabling NSFW in ChatGPT, but I see that as a step in the right direction. The right direction being: the one that leads to less corporate censorship. >In the same way, using gpt5 is now very unbearable to me as it almost always starts all responses of a conversation by things like: "Great question" User preference data is toxic. Doing RLHF on it gives LLM sycophancy brainrot. And by now,…

I have seen a few normally progressive types act quite conservative puritan over the NSFW ChatGPT thing. It seems there are quite a lot of people consider things to be uniformly good or bad and their opinion of the whole colours their opinion of the parts.

OpenAI are in a difficult position when it comes to global standards. It's probably easier to see from outside of the United States, because the degree to which the historical puritanism has influenced everything is remarkable. I remember the release of the Watchmen film and being amazed at how pervasive the preoccupation with a penis was in the media coverage.

Post reply on HN