Live data from Hacker News

Stable Diffusion 2.0

stability.ai

241–250 of 519 posts

Re: Stable Diffusion 2.0

#241
post #237

Earlier quoted context omitted.

Not exactly, since the abuse needed to actually happen for the derivative images to be possible to generate.

Is Stable Diffusion only able to generate images of things that have actually happened?

Hmm, that’s a good point. It seems to be able to “transfer knowledge” for lack of a better term, so maybe it wouldn’t need to be in the dataset at all…

Re: Stable Diffusion 2.0

#242
post #27

I am a solo dev working on a creative content creation app to leverage the latest developments in AI. Demoing even the v1 of stable diffusion to the non-technical general users blows them away completely. Now that v2 is here, it’s clear we’re not able to keep pace in developing products to take advantage of it. The general public still is blown away by autosuggest in mobile OS keyboards. Very few really know how far…

> The general public still is blown away by autosuggest in mobile OS keyboards.

And it's still not available in my language on iOS... :( (Norwegian)

Re: Stable Diffusion 2.0

#243

Earlier quoted context omitted.

Porn has driven many tech advances. I predict that models trained on specific porn genres will appear as soon as training a good model is doable for under $5000. They’ll get here much quicker if we get video to that mark first.

> Porn has driven many tech advances. This is an urban myth.

[deleted]

Re: Stable Diffusion 2.0

#244
post #81

Earlier quoted context omitted.

I have... ~11TBs of free disk space and a 1080ti. Obviously nowhere close to being able to crunch all of Wikimedia Commons, but I'm also not trying to beat Stability AI at their own game. I just want to move the arguments people have about art generators beyond "this is unethical copyright laundering" and "the model is taking reference just like a real human".

To put things in perspective, the dataset it's trained on is ~240TB and Stability has over ~4000 Nvidia A100 (which is much faster than a 1080ti). Without those ingredients, you're highly unlikely to get a model that's worth using (it'll produce mostly useless outputs). That argument also makes little sense when you consider that the model is a couple gigabytes itself, it can't memorize 240TB of data, so it "learned"…

Stable Diffusion 1 was trained with 256 A100s running for a little over three weeks. These days that would cost less than a Tesla…

Re: Stable Diffusion 2.0

#245

Is there a good explanation of how to train this from scratch with a custom dataset[0]? I've been looking around the documentation on Huggingface, but all I could find was either how to train unconditional U-Nets[1], or how to use the pretrained Stable Diffusion model to process image prompts (which I already know how to do). Writing a training loop for CLIP manually wound up with me banging against all sorts of stra…

It will be worthwhile to use images from commons. I have found that my photography is used in the stable diffusion data set. What was funny is that they have taken the images from other URLs than my flickr account.

Re: Stable Diffusion 2.0

#246
post #27

I am a solo dev working on a creative content creation app to leverage the latest developments in AI. Demoing even the v1 of stable diffusion to the non-technical general users blows them away completely. Now that v2 is here, it’s clear we’re not able to keep pace in developing products to take advantage of it. The general public still is blown away by autosuggest in mobile OS keyboards. Very few really know how far…

And we are not doing a good job at educating people and preparing them for what's to happen. People are so used to BigTech making decisions for them

Re: Stable Diffusion 2.0

#247

Is there a good explanation of how to train this from scratch with a custom dataset[0]? I've been looking around the documentation on Huggingface, but all I could find was either how to train unconditional U-Nets[1], or how to use the pretrained Stable Diffusion model to process image prompts (which I already know how to do). Writing a training loop for CLIP manually wound up with me banging against all sorts of stra…

> Writing a training loop for CLIP manually wound up with me banging against all sorts of strange roadblocks and missing bits of documentation, and I still don't have it working. There is working training code for openCLIP https://github.com/mlfoundations/open_clip But training multi-modal text-to-image models is still a _very_ new thing, in terms of the software world. Given that, my experience has been that it's ne…

All this horsepower deployed to image generation is interesting but somebody wake me up when there is a stable diffusion for SQL or when on demand generative User Interfaces are spun up on the fly to suit the purpose.

Re: Stable Diffusion 2.0

#248

I just thought about this, so bare in mind that I don't know much of the technical implications of this, but: Couldn't we train a very good model by distributing the dataset along with the computing power using something similar to folding@home?

i think eventually someone will do

Re: Stable Diffusion 2.0

#249

Earlier quoted context omitted.

The main reason why Stable Diffusion is worried about NSFW is that people will use it to generate disgusting amounts of CSAM. If LAION-5B or OpenAI's CLIP have ever seen CSAM - and given how these datasets are literally just scraped off the Internet, they have - then they're technically distributing it. Imagine the "AI is just copying bits of other people's art" argument, except instead of statutory damages of up to…

Is artificially generated CSAM that doesn't actually involve children in its production not an improvement over the status quo?

I believe the status quo is non-realistic drawings (think Lisa Simpson) can be illegal.

I don't think the fact that it's artificially generated has any bearing for some important purposes.

Re: Stable Diffusion 2.0

#250
post #27

I am a solo dev working on a creative content creation app to leverage the latest developments in AI. Demoing even the v1 of stable diffusion to the non-technical general users blows them away completely. Now that v2 is here, it’s clear we’re not able to keep pace in developing products to take advantage of it. The general public still is blown away by autosuggest in mobile OS keyboards. Very few really know how far…

I'm in a kind of same boat. I think indie games are the way to show true potential of SD. Hence, I'm working on http://diffudle.com/ which is a mix of Wheel Of Fortune + Stable Diffusion + Wordle. I Can't figure it out but feels to me like its lacking something.

It confused me that the letter boxes were divided in 7+3, thus I thought it would be two words while the correct answer was a single 10 letter word. Maybe try to avoid wrapping words.
Post reply on HN