Live data from Hacker News

Stable Diffusion is a big deal

simonwillison.net

161–170 of 488 posts

Re: Stable Diffusion is a big deal

#161

Some of the tech and especially the platform they're building is impressive, but in terms of raw image generation quality from results I've seen and my own experience, I don't find it anything close to DALLE-2

Just compare what was proprietary a couple years ago to what is open now. What is publicly available will be better than DALLE-2 in short order.

Re: Stable Diffusion is a big deal

#162
post #121

These new models completely changed my mind about how much impact AI will have in my lifetime. They are the most impressive software achievements in decades and anyone who has a “meh” reaction will absolutely end up looking silly. I’m not optimistic that their impact will be positive though.

It is capable of surprising beauty. It will however be as beautiful as the disaster that is modern Facebook -- and the misinformation will be more convincing and subtler.

Re: Stable Diffusion is a big deal

#164

Some of the tech and especially the platform they're building is impressive, but in terms of raw image generation quality from results I've seen and my own experience, I don't find it anything close to DALLE-2

Really? I find the output to be in general incredibly better than DALL-E.

If you zoom in on a DALL-E generated image the illusion falls appart and you can see it's a vortex of pixel that fuzzily creates a shape.

So far my experience with Stable Diffusion is that, no matter how basic the prompt, the ouput will be sharper and more coherent.

Re: Stable Diffusion is a big deal

#165
post #46

After using SD heavily for a week, I half agree with this. It is incredibly disruptive, and it's wild how much it accelerates the creative process. I'll give you that. But two things I've noticed: First, artists will still have a massive advantage over non-artists with this tool. A photographer who intimately knows the different lenses and cameras and industry terms will get to a representation of their idea much fas…

I think this can be seen a bit like the invention of an "index fund" for art. Active investors are still needed to generate the market signals that an ETF can aggregate, but for the majority of people an ETF is preferable to the cost of hiring an active investor yourself. And similarly SD needs artists to generate the signal that it aggregates, but for the majority of people it might be (or might soon be) preferable…

There is bound to be a smart kid who already turned this idea into a shitcoin (meaning a pump and dumb money grab, not an actual attempt to make a art-index and tokenize it).

  It seems someone indeed did this (the p&d, not the index):  https://www.coingecko.com/en/coins/artonline

Re: Stable Diffusion is a big deal

#166
post #114

Maybe this current explosion in the relevance and visibility of this kind of AI model will finally lead us to rethink how insanely nonsensical our IP systems are. I'm not holding my breath, but there's hope that this sort of thing will (combined with situations like the HBO debacle) clarify the need for massive IP reform in the cultural zeitgeist. The problem here isn't that the model was trained on copyrighted works…

What will happen is AI will be fed IP legal corpus as part of training. It will only show images that can't be linked (in a justice setting) to any particular work. Which will be interesting because it will produce media culturally alien but still appealing and probably addictive. It will literally be the engine of advancement of human culture. I'm not the conservative type, but I'm still slightly concerned as much a…

>What will happen is AI will be fed IP legal corpus

I first read this as a prediction that AIs will be employed to generate future IP legislation, and now I don't know if I'll be able to go back to sleep.

Re: Stable Diffusion is a big deal

#167

Is this going to start changing the world as much as search did? My guess within 10 years AI will have changed everything, it won’t look like AGI but it’ll be good enough to be better than humans at most creative endeavours including things like generating whole films and possibly games with a few carefully arranged text prompts and start images. Quite terrifying. I wonder if it will be possible to train a neural net…

Reading about how it's done it's not so clear those other things are close at hand. Most of these SD things generate very low-res images and then use upscaling to make them high enough resolution that they don't look like crap.

Apparently the computational power required to make them larger grows exponentially (or geometrically?) so until we find new algorithms it might be a long time before the same technique can generate whole films and games and any coherent way.

Of course maybe that breakthrough will be announced tomorrow.

Re: Stable Diffusion is a big deal

#168
post #50

I am a contrarian by nature. Wearing my investor hat, I continue to be unimpressed with AI tools, including the latest image generation enhancements. I do find stable diffusion interesting, but I don't see the disruption. Similarly, I don't see the disruption from Github Copilot considering I run a company full of highly paid and extremely skilled developers and not one of them uses copilot. How often have you wanted…

What affect will it have on kids? Every kid can now just stick figure draw into painting or else ask Siri "draw me a picture of Trogdor holding up a sword" etc... I can imagine that having some kind of large influence on creativity (not sure negative or positive)

Re: Stable Diffusion is a big deal

#169
post #4

Earlier quoted context omitted.

Fair use is a defense. You can still be sued for fair use, you have to go to court and prove to a judge or jury that your work constitutes fair use.

Yes but we have a common law system and there's already tons of precedent that training AI systems is transformative. It's also quite obvious just by looking at the generated images that it's clearly transformative. The images generated are unique and you can't trace the original copyrighted image from what's generated. You really don't need a judge to see that Fair Use covers Stable Diffusion.

What happens if you give an image prompt like "mona lisa", "daffodils van gogh", or similar designed to describe an image the model was trained on. Will it generate that image?

Or for written works, start with a sentance from a copyrighted work, or part of licensed code. Will it start reproducing that work word for word (like code pilot can do with the GPL license)? Getting these to generate copies of GPL'd, company owned, or other code with restrictions can lead to complex issues for the person/company using that code. Or likewise if a story contains significant elements of copyrighted works; worse if the works have trademarked elements.

Re: Stable Diffusion is a big deal

#170
post #97

I've been trying to get it to draw a picture of a man trapped inside of a light bulb. Can anyone think of a prompt that works? It draws all sorts of freaky things featuring men and light bulbs but none with the former inside the latter.

https://supernotes-resources.s3.amazonaws.com/direct-uploads...

what did you have to write to get it to do this?
Post reply on HN