Live data from Hacker News

Show HN: A Dalle-3 and GPT4-Vision feedback loop

dalle.party

101–110 of 156 posts

Re: Show HN: A Dalle-3 and GPT4-Vision feedback loop

#101
post #50

Earlier quoted context omitted.

> "[...]the final result is more cohesive around a single them than the original idea." That's an observation worth investigating. Here's another set of data points to see if there's more to it... Input prompt: "Six robots on a boat with harpoons, battling sharks with lasers strapped to their heads" GPT4V prompt: "Write a prompt for an AI to make this image. Just return the prompt, don't say anything else. Make it fu…

Both of your examples seem to start with two subjects (steam engine/flying machine and shark/robot), and throughout the animation one of them gets more prominence until the other is eventually dropped altogether.

I was curious if two subject prompts behaved different from three subject, so I've run three additional tests, each with the same three subjects and general prompt structure + instructions, but swapping the position of each subject in the prompt. Each test was run for ten iterations.

GPT4V instructions for all tests: "Write a prompt for an AI to make this image. Just return the prompt, don't say anything else. Make it weirder."

From what you'll see in the results there's possible evidence of bias towards the first subject listed in a prompt, making it the object of fixation through the subsequent iterations. I'll also speculate that "gnomes" (and their derivations) and "cosmic images" are over-represented as subjects in the underlying training data. But that's wild speculation based on an extremely small sample of results.

In any case, playing around with this tool has been enjoyable and a fun use of API credits. Thank you @z991 for putting this together and sharing it!

------ Test 1 ------

Prompt: "Two garden gnomes, a sentient mushroom, and a sugar skull who once played a gig at CBGB in New York City converse about the boundaries of artificial intelligence."

Result: https://dalle.party/?party=ZSOHsnZe

------ Test 2 ------

Prompt: "A sentient mushroom, a sugar skull who once played a gig at CBGB in New York City, and two garden gnomes converse about the boundaries of artificial intelligence."

Result: https://dalle.party/?party=pojziwkU

------ Test 3 ------

Prompt: "A sugar skull who once played a gig at CBGB in New York City, a sentient mushroom, and two garden gnomes converse about the boundaries of artificial intelligence."

Result: https://dalle.party/?party=RBIjLSuZ

Re: Show HN: A Dalle-3 and GPT4-Vision feedback loop

#103

Here's a custom prompt that I enjoyed: "Think hard about every single detail of the image, conceptualize it including the style, colors, and lighting. Final step, condensing this into a single paragraph: Very carefully, condense your thoughts using the most prominent features and extremely precise language into a single paragraph." https://dalle.party/?party=1lSMniUP https://dalle.party/?party=cEUyjzch https://dalle.…

The fractal one is awesome!

Re: Show HN: A Dalle-3 and GPT4-Vision feedback loop

#105

need to throw in a Google to Google to Google language translate to get some more variety

Here's an attempt at using transformations between languages to see what happens:

Prompt: "A unicorn and a rainbow walk into a tavern on Venus"

GPT4V instructions: "Write a prompt for an AI to make this image. Take this prompt and translate it into a different language understood by GPT-4 Vision, don't say anything else."

Results: https://dalle.party/?party=ED7E056D

I wasn't happy with the diversity of languages, so I modified the instructions for a second run of ten iterations using the same prompt as before:

GPT4V instructions: "Using a randomly selected language from around the world understood by GPT-4 Vision, write a prompt for an AI to make this image and then make it weirder. Just return the prompt, don't say anything else."

Result: https://dalle.party/?party=c7-eNR24

The languages it selected don't look particularly random to me which was interesting.

@z991 -- I ran into an unexpected API error the first time I tried this. Perhaps your logs show why it happened. It appeared when the second iteration was run:

"Error: You uploaded an unsupported image. Please make sure your image is below 20 MB in size and is of one the following formats: ['png', 'jpeg', 'gif', 'webp']."

From: https://dalle.party/?party=hI0V0lO_

Re: Show HN: A Dalle-3 and GPT4-Vision feedback loop

#106

Here's a custom prompt that I enjoyed: "Think hard about every single detail of the image, conceptualize it including the style, colors, and lighting. Final step, condensing this into a single paragraph: Very carefully, condense your thoughts using the most prominent features and extremely precise language into a single paragraph." https://dalle.party/?party=1lSMniUP https://dalle.party/?party=cEUyjzch https://dalle.…

Mine got surral real fast, though the sixth one is kinda cool https://dalle.party/?party=DNgriW_E
Post reply on HN