Live data from Hacker News

Stable Diffusion Public Release

stability.ai

11–20 of 437 posts

Re: Stable Diffusion Public Release

#15
post #2

This is one of the most important moments in all of art history. Millions of people just got unconditional access to the state-of-the-art in AI text-to-image for absolutely free less the cost of hardware. I have an Nvidia GPU myself and am thrilled beyond belief with the possibilities that this opens up. Am planning on doing some deep dives into latent-space exploration algorithms and hypernetworks in the coming days…

> state-of-the-art in AI image-to-text

I think you meant text-to-image!

Re: Stable Diffusion Public Release

#17
post #7
post #2

This is one of the most important moments in all of art history. Millions of people just got unconditional access to the state-of-the-art in AI text-to-image for absolutely free less the cost of hardware. I have an Nvidia GPU myself and am thrilled beyond belief with the possibilities that this opens up. Am planning on doing some deep dives into latent-space exploration algorithms and hypernetworks in the coming days…

> particularly interested in training a hypernetwork to translate natural language instructions into latent-space navigation instructions with the end goal of enabling me to give the model natural-language feedback on its generations. What are you doing exactly?

Imagine every conceivable image is laid out on the ground, images which are similar to each other are closer together. You’re looking at an image of a face. Some nearby images might be happier, sadder, with different hair or eye colours, every possibility in every combination all around it. There are a lot of images, so it is hard to know where to look if you want something specific, even if it is nearby. They’re going to write software to point you in the right direction, by describing what you want in text.

Here’s an example of this sort of manipulation: https://arxiv.org/abs/2102.01187

Re: Stable Diffusion Public Release

#18

I've been looking forward to this. The license however strikes me as too aspirational, and it may be hard to enforce legally: > You agree not to use the Model or Derivatives of the Model: > - In any way that violates any applicable national, federal, state, local or international law or regulation; > - For the purpose of exploiting, harming or attempting to exploit or harm minors in any way; > - To generate or dissem…

I dunno. I can imagine any of those points being the subject of a civil suit and for someone to win damages, for e.g. psychological harm. The parts talking about “effect” instead of intent are of questionable enforceability - how can I agree not to cause an unanticipated effect on a third party? I cannot. But having said that, I can be asked to account for effects that a “reasonable person” would anticipate, so there’s that.

These are all things that someone could sue over (especially in California) and so they’re wanting to place the responsibility on the artist and not their tools.

Re: Stable Diffusion Public Release

#19
post #7
post #2

This is one of the most important moments in all of art history. Millions of people just got unconditional access to the state-of-the-art in AI text-to-image for absolutely free less the cost of hardware. I have an Nvidia GPU myself and am thrilled beyond belief with the possibilities that this opens up. Am planning on doing some deep dives into latent-space exploration algorithms and hypernetworks in the coming days…

> particularly interested in training a hypernetwork to translate natural language instructions into latent-space navigation instructions with the end goal of enabling me to give the model natural-language feedback on its generations. What are you doing exactly?

AFAICT: making a navigation/direction model that can translate phrase-based directions into actual map-based directions, with the caveat that the model would be updated primarily by giving it feedback the same way that you would give a person feedback.

Sounds only a couple of steps removed from basically needing AGI?

Re: Stable Diffusion Public Release

#20

I've been looking forward to this. The license however strikes me as too aspirational, and it may be hard to enforce legally: > You agree not to use the Model or Derivatives of the Model: > - In any way that violates any applicable national, federal, state, local or international law or regulation; > - For the purpose of exploiting, harming or attempting to exploit or harm minors in any way; > - To generate or dissem…

[deleted]
Post reply on HN