Live data from Hacker News

Show HN: A Dalle-3 and GPT4-Vision feedback loop

dalle.party

71–80 of 156 posts

Re: Show HN: A Dalle-3 and GPT4-Vision feedback loop

#71
post #70

My results are disappoitingly noisy but I love the concept https://dalle.party/?party=bxrPClVg https://dalle.party/?party=mmBxT8G- https://dalle.party/?party=kxra0OKY (the last prompt got a content warning) https://dalle.party/?party=Q8VYXU0_

You have a custom prompt enabled (probably from viewing another one and pressing "start over") that is asking for opposites which will increase the noise a lot.

Oh wow, I completely missed that, thanks!

Re: Show HN: A Dalle-3 and GPT4-Vision feedback loop

#72

The "create text version of image" prompt matters a ton. I tried three, demo here: default https://dalle.party/?party=JfiwmJra hyper-long + max detail + compression - This shows that with enough text, it can do a really good job of reproducing very, very similar images https://dalle.party/?party=QtEqq4Mu hyper-long + max detail + compression + telling it to cut all that down to 12 words - This seems okay. I might be…

I like your prompt! Some results:

https://dalle.party/?party=Vwuu9ipd

https://dalle.party/?party=Pc3g4Har

My intuition says that the "poetry" part skews the images in a bit of a kitchy direction.

Re: Show HN: A Dalle-3 and GPT4-Vision feedback loop

#73
post #23
post #2

Also, descent into Corgi insanity: https://dalle.party/?party=oxXJE9J4

So do I understand correctly that the corgi was purely made up from GPT-4's interpretation of the picture?

It was created by uploading the previous picture to GPT-4 to generate a prompt by using the vision API and using this prompt to create the new prompt:

"Write a prompt for an AI to make this image. Just return the prompt, don't say anything else. Replace everything with corgi."

Then it takes that new prompt and feeds it to Dall-E to generate a new image. And then it repeats.

Re: Show HN: A Dalle-3 and GPT4-Vision feedback loop

#74
post #70

My results are disappoitingly noisy but I love the concept https://dalle.party/?party=bxrPClVg https://dalle.party/?party=mmBxT8G- https://dalle.party/?party=kxra0OKY (the last prompt got a content warning) https://dalle.party/?party=Q8VYXU0_

You have a custom prompt enabled (probably from viewing another one and pressing "start over") that is asking for opposites which will increase the noise a lot.

Clicking start over selects the default prompt but it seems like you are right.

Starting over by removing the permalink parameter gives me much more consistent results! An exampe from before: https://dalle.party/?party=Sk8srl2F

I wonder what the default prompt is. There still seems to be a heavy bias towards futuristic cityscapes, deserts, and moonlight. It might just be the model bit it's a bit cheesy if you ask me!

Re: Show HN: A Dalle-3 and GPT4-Vision feedback loop

#77
post #22

Cool idea! I made one with the starting prompt "an artificial intelligence painting a picture of itself": https://dalle.party/?party=wszvbrOx It consistently shows a robot painting on a canvas. The first 4 are paintings of robots, the next 3 are galaxies, and the final 2 are landscapes.

I tried something similar! Interestingly, picture 2 was what I wanted. After that... weirdness ensued https://dalle.party/?party=C2w7zuwe

Re: Show HN: A Dalle-3 and GPT4-Vision feedback loop

#79

This is hilarious, thanks for sharing At the same time, it perfectly illustrates my main issue with these AI art tools: they very often generate pictures that are interesting to look at while very rarely generating exactly what you want them to. I imagine a study in which participants are asked to create N images of their choosing and rate them from 0-10 on how satisfied they are with the results. One try per image o…

That's not an AI issue. A few sentences can't exactly capture the contents of a drawing - regardless of "intelligence".

Yeah, try commissioning art with a single paragraph prompt and getting exactly what you want without iteration.

Re: Show HN: A Dalle-3 and GPT4-Vision feedback loop

#80
post #50

Earlier quoted context omitted.

I find it somewhat fascinating that in both examples, the final result is more cohesive around a single them than the original idea.

> "[...]the final result is more cohesive around a single them than the original idea." That's an observation worth investigating. Here's another set of data points to see if there's more to it... Input prompt: "Six robots on a boat with harpoons, battling sharks with lasers strapped to their heads" GPT4V prompt: "Write a prompt for an AI to make this image. Just return the prompt, don't say anything else. Make it fu…

Both of your examples seem to start with two subjects (steam engine/flying machine and shark/robot), and throughout the animation one of them gets more prominence until the other is eventually dropped altogether.
Post reply on HN