Live data from Hacker News

How a Stable Diffusion prompt changes its output for the style of 1500 artists

gorgeous.adityashankar.xyz

171–180 of 205 posts

Re: How a Stable Diffusion prompt changes its output for the style of 1500 artists

#172
It's really interesting how good some of the results are, but the styles seem to be completely mixed up. Just looking up Claude Monet, and there isn't anything impressionist. Leonardo da Vinci, on the other hand, give results that rather look like British impressionist paintings.

I've seen paintings in the style of Donato Giancola that really looked like his style, but in the examples of this site none of the result do. Maybe there needs to be some adjustments to the prompts?

Re: How a Stable Diffusion prompt changes its output for the style of 1500 artists

#173

Earlier quoted context omitted.

I've added a search function now, I'll change it to be in alphabetical order gimmie a sec

done

I'm not sure what you did, but now the name you select do not match with the style of the painting :)

Re: How a Stable Diffusion prompt changes its output for the style of 1500 artists

#174
post #110

Earlier quoted context omitted.

The positioning seems very consistent, almost to the point where I wonder if that was part of the selection process to demonstrate the differences in style. There are only four per style, where the position of a subject could be a selection factor. Hard to tell if the position similarities are driven by the Stable Diffusion model or by the selection of representative images.

The composition and positioning come from the original seed. If the same seed is used, the same background image noise is applied for the initial image which is transformed into all the styles. Thus the similarities you see would make sense if using also the same seed for the tests.

Yes, if you generate a bunch of images of waves or the sea, or other repeating patterns with the same seed, you can see how all the 'peeks and toughs' of those patterns line up in the same place.

Re: How a Stable Diffusion prompt changes its output for the style of 1500 artists

#175

Earlier quoted context omitted.

done

I'm not sure what you did, but now the name you select do not match with the style of the painting :)

I’m seeing that too. Results don’t match the artist at all. Also, the names at the top are alpha by firstname, and all the Japanese artists are bunched up at the bottom.

Re: How a Stable Diffusion prompt changes its output for the style of 1500 artists

#176

Earlier quoted context omitted.

It's the nature of ML models--nobody is 100% sure what it understands until they try something and get results. It was given a lot of tagged data: 600 million captioned images from LAION-5B. So if you want to know what it might support, you could try any one of the captions from those 600 million images.

But why isn't the list of words from those captions available anywhere (at least as far as I can tell)? There may be 600 million captions, but the number of unique words would probably be 10 or 20 thousand at most, completely feasible to browse or grep.

The complete vocabulary is available here: https://huggingface.co/openai/clip-vit-base-patch32/resolve/...

It's a bit less than 50k words, but that includes space-padded duplicates.

Re: How a Stable Diffusion prompt changes its output for the style of 1500 artists

#177
post #41

Earlier quoted context omitted.

But why isn't the list of words from those captions available anywhere (at least as far as I can tell)? There may be 600 million captions, but the number of unique words would probably be 10 or 20 thousand at most, completely feasible to browse or grep.

I don't think the underlying model is word based, but character based. You could download the caption data for LAION and grep that, but it's not strictly 1:1 with what SD was trained against.

No, it's word based.

The vocabulary is here: https://huggingface.co/openai/clip-vit-base-patch32/resolve/...

It is contextual though, so words in different orders mean different things.

Re: How a Stable Diffusion prompt changes its output for the style of 1500 artists

#179

Earlier quoted context omitted.

People said this about cameras. About digital cameras. About digital photo editing software. The next generation will normalize these tools and find incredible ways to be creative within their new cutting edge medium. The post-art world is here! Just think about how history books will remember this period! The styles that will be borne of necessity, of the need to break down art and find what makes it tick.

Yep, I see this as a start, and very curious to see the ways in which it’ll get used with a human in the loop, and also the ways human artists will be pushed to creat art that’s out of distribution for these models.

Even then, few artist got famous on technical skill alone, surely less than those who got famous primarily for their message irregardless of their skill. And this is before getting into the endless pit of defining what art is.

Besides having an ai doing the legwork is no much different than Veronese giving large swath of paintings to his novices while focusing on the two/three major parts.

Re: How a Stable Diffusion prompt changes its output for the style of 1500 artists

#180
post #115
post #97

Earlier quoted context omitted.

Some people would probably argue that humans are essentially the same.

Whatever else this emergent "creative AI" phenomenon may or may not do, it's definitely touching nerves in people who still believe there's something ineffable and transcendent about the creative experience.

Stable Diffusion is literally copy-pasting existing artwork. It's not creating anything new.

If anything, this sort of "AI" only makes human ineffable creative experiences even more valuable, because without it there would be no training set and no "AI".

Post reply on HN