Live data from Hacker News

Stable Diffusion Public Release

stability.ai

31–40 of 437 posts

Re: Stable Diffusion Public Release

#31

DALL-E 2 just got smoked. Anyone with a graphics card isn't going to pay to generate images, or have their prompts blocked because of the overly aggressive anti-abuse filter, or have to put up with the DALL-E 2 "signature" in the corner. It makes me wonder how OpenAI is going to work around this because this makes DALL-E 2 a very uncompetitive proposition. Except, of course, for people without graphics cards, but it'…

This, however, is unconditionally good for the end users. I expect OpenAI to lower their prices significantly quite soon.

Re: Stable Diffusion Public Release

#33
The most interesting part, to me, of a release like this is the amount of "please don't abuse this technology" pleading. No licence will ever stop people from doing things that the licence says they can't. There will always be someone who digs into the internals and makes a version that does not respect your hopes and dreams. It's going to be bad.

As I see it, within a couple years this tech will be so widespread and ubiquitous that you can fully expect your asshole friends to grab a dozen photos of you from Facebook and then make a hyperrealistic pornographic image of you with a gorilla[0]. Pandora's box is open, and you cannot put the technology back once it's out there.

You can't simply pass laws against it because every country has its own laws and people in other places will just do whatever it is you can't do here.

And it's only going to get better/worse. Video will follow soon enough, as the tech improves. Politics will be influenced. You can't trust anything you see anymore (if you even could before, since Photoshop became readily available).

Why bother asking people not to? I guess if it helps you sleep at night that you tried, I guess?

[0]A gorilla if you're lucky, to be honest.

Re: Stable Diffusion Public Release

#34
post #30
post #2

This is one of the most important moments in all of art history. Millions of people just got unconditional access to the state-of-the-art in AI text-to-image for absolutely free less the cost of hardware. I have an Nvidia GPU myself and am thrilled beyond belief with the possibilities that this opens up. Am planning on doing some deep dives into latent-space exploration algorithms and hypernetworks in the coming days…

>This is one of the most important moments in all of art history. I agree, but not for the reasons you imply. It will force real artists to differentiate themselves from AI, since the line is now sufficiently blurred. It's probably the death of an era of digital art as we know it.

Maybe this will signal a return to "real" non-digital artwork and methods...

Re: Stable Diffusion Public Release

#35
post #7

Earlier quoted context omitted.

> particularly interested in training a hypernetwork to translate natural language instructions into latent-space navigation instructions with the end goal of enabling me to give the model natural-language feedback on its generations. What are you doing exactly?

AFAICT: making a navigation/direction model that can translate phrase-based directions into actual map-based directions, with the caveat that the model would be updated primarily by giving it feedback the same way that you would give a person feedback. Sounds only a couple of steps removed from basically needing AGI?

I suspect you’d want to start by trying to translate differences between images into descriptive differences. Maybe you could generate examples by symbolic manipulation to generate pairs of images or maybe nlp can let us find differences between pairs of captions? Large nlp models already feel pretty magical to me and encompass things that we would have said required AGI until recently so it seems possible, though really tough

Re: Stable Diffusion Public Release

#36
post #30
post #2

This is one of the most important moments in all of art history. Millions of people just got unconditional access to the state-of-the-art in AI text-to-image for absolutely free less the cost of hardware. I have an Nvidia GPU myself and am thrilled beyond belief with the possibilities that this opens up. Am planning on doing some deep dives into latent-space exploration algorithms and hypernetworks in the coming days…

>This is one of the most important moments in all of art history. I agree, but not for the reasons you imply. It will force real artists to differentiate themselves from AI, since the line is now sufficiently blurred. It's probably the death of an era of digital art as we know it.

[deleted]

Re: Stable Diffusion Public Release

#39
While neat, and no doubt impressive, it still utterly fails on prompts that should be completely reasonable to any sane human being/artist.

Take something like "A cat dancing atop a cow, with utters that are made out of ar-15s that shoot lazer-beam confetti". A vivid description should be aroused in your head, and no doubt, I could imagine an artist have a lot of fun creating such a description... Alas, what the model spits out is pure unusable garbage.

Re: Stable Diffusion Public Release

#40
Really interesting. I wonder if at some point it would be possible to optimize a network for size and speed by focusing on a specific genre, like impressionist or only pixel art. I like that I can get an image in any style I want, but that has to increase the workload substantially.
Post reply on HN