I am completely uninformed in this space. Would someone be kind to explain what the current state of the art in image generation is (how does this compare to Midjourney and others)? How do open source models stack up? Also what are the most common use cases for image generation?
Midjourney may be better for plain prompts, but Stable Diffusion is SOTA because of the tooling and finetuning surrounding it.
For the longest time I thought it was google imaging things and doing some photoshop to make things look like Pixar because it was so bad.