Live data from Hacker News

Stable Diffusion Public Release

stability.ai

191–200 of 437 posts

Re: Stable Diffusion Public Release

#191
post #185

this one was incredibly easy to trick into producing uncensored nudes. which. personally.. I think is great.. but to each their own (NSFW!!!) this ween does not exist : https://i.ibb.co/D7qJ7HC/23456532.png

They have a content filter classifier at the top wondering how this escaped that.

Re: Stable Diffusion Public Release

#192
Played with it for a bit in DreamStudio so I could control more of the settings. So far everything it generates is "high quality", but the AI seems to lack the creativity and breadth of understanding that DALL-E 2 has. OpenAI's model is better at taking wildly differing concepts and figuring out creative ways to glue them together, even if the end result isn't perfect. Stable Diffusion is very resistant to that, and errs towards the goal of making a high quality image. If it doesn't understand the prompt, it'll pick and choose what parts of the prompt are easiest for it and generate fantastic looking results for those. Which is both good and bad.

For example, I asked it in various ways for a bison dressed as an astronaut. The results varied from just photos of astronauts, to bisons on earth, to bisons on the moon. The bison was always drawn hyper realistically, which is cool, but none of them were dressed as an astronaut. DALLE on the other hand will try all kinds of different ways that a bison might be portrayed as an astronaut. Some realistic, some more imaginative. All of them generally trying to fulfill the prompt. But many results will be crude and imperfect.

I personally find DALLE to be more satisfying to play with right now, because of that creativity. I'm not necessarily looking for the highest quality results. I just want interesting results that follow my prompt. (And no, SD's Scale knob didn't seem to help me). But there's also a place for SD's style if you just want really great looking, but generic stuff.

That said, the current version of SD was explicitly finetuned on an "aesthetically" ranked dataset. So these results aren't really surprising. I'm sure the next generations of SD will start knocking DALLE out of the park in both metrics. And, of course, massive massive props to Stability.ai for releasing this incredible work as open source. Imagine all the tinkering and evolving people are going to do on top of this work. It's going to be incredible.

Re: Stable Diffusion Public Release

#193
I'm getting a 403 on the Colab (while successfully logging in and providing a huggingface token). Is it already disabled? Do you have to pay huggingface to download the model? It's unclear from the Colab and post where the issue is.

Re: Stable Diffusion Public Release

#195

Earlier quoted context omitted.

Human works are needed to create the initial datasets, but an increasing amount of models use generative feedback loops to create more training data. This layer can easily introduce novel styles and concepts without further human input. The time is coming where we will need to, as patrons, reevaluate our relationships with art. I fear art is returning to a patronage model, at least for now, as certainly an industry w…

Why would people want to consume art that says nothing and means nothing? While this technology is fascinating, it produces the visual equivalent of muzak, and will continue to do so in perpetuity without the ability to reason.

Can photography be good art? Is Marcel Duchamp (found object) art? Can good art be discovered almost serendipitously, or can good art only be created by slowly learning and applying a skill?

I think art is mostly about perception and selection, by the viewer. There are others that think art is more about the crafting process by the artist. How do you tell the difference between an artist and a craftsperson?

One way I categorise artists I have met is engineer-type artists versus discovery-type artists: https://news.ycombinator.com/item?id=31981875

Disclaimer: I am engineer.

Re: Stable Diffusion Public Release

#197
post #84

Earlier quoted context omitted.

We do a decent job of banning child pornography. And bringing the two ideas together, is child pornography that is provably created by an AI still illegal?

As far as I'm aware countries fall broadly into two camps. Camp 1, USA for example, is concerned purely with the abuse of children, i.e., anything that depicts or is constructed of pieces of real children is illegal but other things such as drawings, stories, adults role playing, etc is not. Camp 2 outlaws any representation of it whether or not a child was involved. Nowhere will a training set featuring pictures of…

> Nowhere will a training set featuring pictures of naked children be legal.

Appropriately from the recent news stories, but it's easy to imagine at least portions of such pictures being available for medical diagnostic purposes. I've sent pictures of my children to my doctor, so presumably in the future it's easy to imagine sending pictures to an AI to diagnose which would require a suitably fleshed out (pardon the pun) training set.

Re: Stable Diffusion Public Release

#198

DALL-E 2 just got smoked. Anyone with a graphics card isn't going to pay to generate images, or have their prompts blocked because of the overly aggressive anti-abuse filter, or have to put up with the DALL-E 2 "signature" in the corner. It makes me wonder how OpenAI is going to work around this because this makes DALL-E 2 a very uncompetitive proposition. Except, of course, for people without graphics cards, but it'…

I have been considering building a "modern" computer and now wonder exactly what I need to load this puppy up.

People are saying you need a GPU with 6.9GB of RAM for the current model, so in practice at least an 8GB GPU.

Thankfully, GPU prices have finally calmed down and you can get one for a reasonable price. I think any of the RTX 3000 series desktop GPU's should do it, for example.

Re: Stable Diffusion Public Release

#199
post #69

Earlier quoted context omitted.

I just tried it via this link. I'm not sure what I'm looking at here but the results were extremely underwhelming. I've used Dall E 2 and Midjourney extensively so I know what they're capable of. Maybe I'm missing something?

I've only used it via discord, but it's much better than Midjourney and sometimes better than Dall-E. So maybe that site isn't the same thing or you need to work on your prompts.

I haven't seen anything beat midjourney for creating atmosphere.

DALL-E is great at making 'things' and generally good/great at faces.

Re: Stable Diffusion Public Release

#200
post #147

Earlier quoted context omitted.

You'll need a GPU. One with a LOT of RAM, like an RTX 3090, which has 24 GB.

According to this post, it needs 6.9 Gb. So the 3070, 3070-Ti, 3080, etc. can all run it. Sadly, my RTX 2060 is below that limit...

Apparently the model decompresses, and it won't fit very well on the 8gb models... I'm willing to give the max settings a spin on my 3070ti, but I'm not very hopeful.
Post reply on HN