Live data from Hacker News

Show HN: Stable Diffusion image gallery of 300 AI generated pictures

github.com

1–7 of 7 posts

Re: Show HN: Stable Diffusion image gallery of 300 AI generated pictures

#2
I feel like 'AI Whisperer' will be a job title within the next few months if it isn't already. Having played around with Stable Diffusion (on a low powered machine) I've found that choosing the right terms and aesthetics in your prompt are huge influences between good and weird outcomes.

Re: Show HN: Stable Diffusion image gallery of 300 AI generated pictures

#3
Assuming you have seen:

https://lexica.art/

There’s also this prompt guide; which is for Dall-e, but a lot of general advice too:

http://dallery.gallery/the-dalle-2-prompt-book/

- PDF is here: http://dallery.gallery/wp-content/uploads/2022/07/The-DALL·E...

Re: Show HN: Stable Diffusion image gallery of 300 AI generated pictures

#4

Assuming you have seen: https://lexica.art/ There’s also this prompt guide; which is for Dall-e, but a lot of general advice too: http://dallery.gallery/the-dalle-2-prompt-book/ - PDF is here: http://dallery.gallery/wp-content/uploads/2022/07/The-DALL·E...

Thank you! I haven't seen those links. Very much appreciated.

Re: Show HN: Stable Diffusion image gallery of 300 AI generated pictures

#5
post #2

I feel like 'AI Whisperer' will be a job title within the next few months if it isn't already. Having played around with Stable Diffusion (on a low powered machine) I've found that choosing the right terms and aesthetics in your prompt are huge influences between good and weird outcomes.

You're right on. I noticed that Stable Diffusion does well with some kinds of prompts, such as "in the style of ", or with outdoor scenery, or with mythology and illustration. Yet any images of real people are massively distorted, as if Stable Diffusion doesn't understand what a human face should look like.

Re: Show HN: Stable Diffusion image gallery of 300 AI generated pictures

#6
post #5
post #2

I feel like 'AI Whisperer' will be a job title within the next few months if it isn't already. Having played around with Stable Diffusion (on a low powered machine) I've found that choosing the right terms and aesthetics in your prompt are huge influences between good and weird outcomes.

You're right on. I noticed that Stable Diffusion does well with some kinds of prompts, such as "in the style of ", or with outdoor scenery, or with mythology and illustration. Yet any images of real people are massively distorted, as if Stable Diffusion doesn't understand what a human face should look like.

It seems that a common "pipeline" is to run the outputs of SD through GFPGAN to correct the faces. Although it can't help with the number of limbs or fingers.