Live data from Hacker News

Stable Diffusion 2.0

stability.ai

491–500 of 519 posts

Re: Stable Diffusion 2.0

#491

Earlier quoted context omitted.

What? No. Capitalism is a more specific system for organizing goods and services, wherein the means of production and distribution of those goods and services (buildings, land, machines and other tools, vehicles etc) are privately owned and operated by workers (who are paid a wage) for the profit of the owners. That's only been the norm for a few hundred years, and only in certain places. Also, capitalism is separate…

> That's only been the norm for a few hundred years, and only in certain places. Can you point to a system that worked well before that you'd like to go back to?

Your question assumes the only alternatives involve going back, not forwards. There are still many untried sociopolitical systems.

Re: Stable Diffusion 2.0

#492

Earlier quoted context omitted.

That is a very apropos reference. If you're familiar with Cubism, you know that there's Picasso, and then there's Braque. The one is an art celebrity beyond almost any other, and the other isn't. But they developed Cubism in parallel. There were periods where their work was almost indistinguishable. "Houses at l'Estaque", the trope namer for Cubism thanks to the remarks of a critic, was in fact by Braque. You can gen…

> You can generate infinite recognizable Basquiat from an AI, but is it Basquiat? No, of course not, because Basquiat's style operates within the context of a specific individual human making a point about expectations and the interface between his race and his artistic boldness and audacity as experienced by his wealthy audience. I'm not sure how I feel about this - I agree with the conclusion, but not the reasoning…

On reflection, I'm going to say 'nope'. Because it's Basquiat, I'm pretty sure you couldn't get him to make a model of himself (maybe he would, and call it 'samo'?). I don't think you could pay him to draw a circle with a pencil: I think he'd have been offended and angry. And so that is not 'his work'. It trips over what makes him Basquiat, so doing these things is not Basquiat (though it's very, very Warhol).

Even more than that, you couldn't do Rothko that way: the man would be beyond offended and would not deal with you at all. But by contrast, you ABSOLUTELY are doing a Warhol if you train an AI on him and have it generate infinite works, and furthermore I think he'd be absolutely delighted at the notion, and would love exploring the unexplored conceptual space inside the neural net.

In a sense, an AI Warhol is megaWarhol, an unexplored level of Warholiness that wasn't attainable within his lifetime.

Context and intent matter. All of modern art ended up exploring these questions AS the artform itself, so boiling it down to 'did a specific person make a mark on a thing' won't work here.

Re: Stable Diffusion 2.0

#493
post #61

Is there a good explanation of how to train this from scratch with a custom dataset[0]? I've been looking around the documentation on Huggingface, but all I could find was either how to train unconditional U-Nets[1], or how to use the pretrained Stable Diffusion model to process image prompts (which I already know how to do). Writing a training loop for CLIP manually wound up with me banging against all sorts of stra…

> train this from scratch If you're talking about training from scratch and not fine tuning, that won't be cheap or easy to do. You need thousands upon thousands of dollars of GPU compute [1] and a gigantic data set. I trained something nowhere near the scale of Stable Diffusion on Lambda Labs, and my bill was $14,000. [1] Assuming you rent GPUs hourly, because buying the hardware outright will be prohibitively expen…

Hi, do you have a writeup of that anywhere? Would love to hear (read) more about it

Re: Stable Diffusion 2.0

#494
post #27

I am a solo dev working on a creative content creation app to leverage the latest developments in AI. Demoing even the v1 of stable diffusion to the non-technical general users blows them away completely. Now that v2 is here, it’s clear we’re not able to keep pace in developing products to take advantage of it. The general public still is blown away by autosuggest in mobile OS keyboards. Very few really know how far…

I'm in a kind of same boat. I think indie games are the way to show true potential of SD. Hence, I'm working on http://diffudle.com/ which is a mix of Wheel Of Fortune + Stable Diffusion + Wordle. I Can't figure it out but feels to me like its lacking something.

Bookmarked! I love it, would be good to play past games

Re: Stable Diffusion 2.0

#495
post #282

Earlier quoted context omitted.

I really want to try this. Please add mobile iOS support!

I'm able to use this on my Iphone browser. Can you elaborate if you're facing any difficulties?

It worked eventually for me but the scrolling seems stuck in the beginning or maybe only certain areas are scrollable?

Re: Stable Diffusion 2.0

#496

To put things in perspective, the dataset it's trained on is ~240TB and Stability has over ~4000 Nvidia A100 (which is much faster than a 1080ti). Without those ingredients, you're highly unlikely to get a model that's worth using (it'll produce mostly useless outputs). That argument also makes little sense when you consider that the model is a couple gigabytes itself, it can't memorize 240TB of data, so it "learned"…

>> it can't memorize 240TB of data, so it "learned" learning is a form of memorization but yeah

It compresses a whole image down to 1 byte, 60000:1 ratio. That's how much it is allowed to "memorise" from each input on average. Less than a pixel from a whole image.

Re: Stable Diffusion 2.0

#497
post #325

Earlier quoted context omitted.

Human artists derive their inspiration and styles from a large set of copyrighted works, but they are free to produce new art despite of that. Art would have developed much slower and be much poorer if, for example, Impressionism or Cubism had been entangled in long ownership confrontations in courts. Then there's the fact that humanity has been able to develop and share art and literary works for thousands of years…

> It would be interesting to see if this technology can erode the copyright concept a bit Copyright law (especially in US) only ever changes in the direction that suits corporations. So - no. What I expect instead is artists being sued by a big tech company for copyright violations because that big tech company used the artist Public Domain image for training their copyrighted AI and as a result it created a copyrigh…

>Copyright law (especially in US) only ever changes in the direction that suits corporations. So - no.

Just objectively false.

Re: Stable Diffusion 2.0

#498

Earlier quoted context omitted.

They reportedly did so to stop people from generating CSAM [0]. [0] https://old.reddit.com/r/StableDiffusion/comments/y9ga5s/sta...

I regularly skimmed 4Chan’s /b/ to get a frame of reference for fringe internet culture. But I’ve had to stop because the CSAM they generate by the hundreds per hour is just freakishly and horrifyingly high fidelity. There’s a lot of important social questions to ask about the future of pornography, but I’m sure not going to be the one to touch that with a thousand foot pole.

This comment is so far off it might as well be an outright lie. There hasn't been CSAM on /b/ for years. The 4chan you speak of hasn't existed in a decade.

Re: Stable Diffusion 2.0

#499

Earlier quoted context omitted.

> Specifically, Wikimedia Commons images in the PD-Art-100 category, because the images will be public domain in the US and the labels CC-BY-SA. Doesn't the "BY" part of the license mean you have to provide attribution along with your models' output[0]? I feel you'll have the equivalent of Github Copilot problem: it might be prohibitive to correctly attribute each output, and listing the entire dataset in attribution…

If I was generating image labels I absolutely would need to worry about that. However, since we're only generating images alone, we don't need to worry about bits of the labels getting into the output images. The attribution requirement would absolutely apply to the model weights themselves, and if I ever get this thing to train at all I plan to have a script that extracts attribution data from the Wikimedia Commons…

>If I was generating image labels I absolutely would need to worry about that. However, since we're only generating images alone, we don't need to worry about bits of the labels getting into the output images.

Just to be correct, SD generates labels on images sometimes, so, we need to worry ;)

Re: Stable Diffusion 2.0

#500

Earlier quoted context omitted.

I'm in a kind of same boat. I think indie games are the way to show true potential of SD. Hence, I'm working on http://diffudle.com/ which is a mix of Wheel Of Fortune + Stable Diffusion + Wordle. I Can't figure it out but feels to me like its lacking something.

Bookmarked! I love it, would be good to play past games

Thanks :)
Post reply on HN