Playing with the prompt-only demo at https://huggingface.co/spaces/stabilityai/stable-diffusion I got the impression that many apparently harmless requests corner the model into a very sparse set of examples exhibiting extreme bias and utter nonsense. For example (seed 0, other advanced options at default values): Blue hamsters filling donuts with nails mostly edible donuts and quasi-donuts; some blue, but no hamster…
I think many of your prompts have trouble getting the right representation because you're using a lot of abstract and implied language. For these prompts to work right, you need to be very specific and often add keywords as if you're specifying a mood board. You also need some overlap between what the model has likely been trained on and the expected output. The system works much better in img2img mode. Slapping toge…
Ooooo. Ouch. That's... Kind of a death knell for a good tool in my experience, and suggests the tool in question is a glorified search engine in a sense. With a really confusing query syntax composed of words, and graphical starting states.
If I have to become to get anything done... Why not just learn to draw/hire a creative?