Live data from Hacker News

Dall-E 2

openai.com

401–410 of 511 posts

Re: Dall-E 2

#401
post #50

The correct response here from the artists point of view should be a widespread coming together against their art being used as training data for ML models. With a quickly spread new license on most major art submission sites that explicitly forbids AI algorithms from using their work, artists would effectively starve OpenAI and others from using their own works to put them out of a job.

There has been precedent for such a movement. In 2011, an "art collective" sourced user-submitted artwork without the artists' consent for an installation where visitors were instructed to step all over printouts of the art on the floor. The artists complained that their work was being used inappropriately. A large number of those artists to left for other art websites en masse.[0]

There doesn't seem to be an equivalent movement with AI-generated art, probably because the understanding of how the models are trained from large datasets is not mainstream yet. I would imagine thousands of those same artists/consumers would be up in arms if they had a basic understanding of ML and millions of average people were beginning to feed the models their own keywords.

This I think ties in with the "responsibility" principles that OpenAI outlines. Once the generation technique has been reverse-engineered and can be used without limits, there is no way to uninvent it. It can be made illegal, but humans can always find a way around laws if they want something badly enough. This could have drastic consequences if enough artists believe that the training violates their respect or other intangible humanistic qualities. With technological advancement that can never be put back in the bottle and spreads to occupy the entire consciousness of the Internet, their options for recourse will be far different than being able to tell a single fringe art group siphoning others' content to pack up and leave.

[0] https://en.wikipedia.org/wiki/Pixiv#Chaos_Lounge

Re: Dall-E 2

#402

>We’ve limited the ability for DALL·E 2 to generate ... adult images. I think that using something like this for porn could potentially offer the biggest benefit to society. So much has been said about how this industry exploits young and vulnerable models. Cheap autogenerated images (and in the future videos) would pretty much remove the demand for human models and eliminate the related suffering, no? EDIT: typo

It'd be ironic if we ended up destroying our planet by using so much electricity to train models to generate a maximally optimal version of the type of content that you refer to similar to crypto mining.

Re: Dall-E 2

#403

Earlier quoted context omitted.

How do you propose we talk about what it is doing if not by using the terminology from the human editing process it is replacing? I'm struggling to express things. My issue is that it appears to not be possible to explain what the AI is doing at all. If you could, you'd be able to actually control the output. And talking about how the model is trained is interesting but not an answer. Of course there is a superimposi…

> Of course there is a superimposing step, that just means it adds its layer on top of the photo you provide. That's all it means and that's literally what it is doing, that's all I tried to say, heh. It is not doing this. You are wrong. You are mistaken. You are confused. You do not understand what is happening. (People have tried to tell you this several times, but you're not listening. shrug One more can't hurt.)

I am specifically referring to the flamingo example: "DALL·E 2 can make realistic edits to existing images from a natural language caption."

You provide the background image and a text prompt and it doodles on top of the image you provided as per their demonstration. I wasn't referring to the other examples down the page where it conjures up a brand new image from scratch based on your image input.

It is great that you can tell it to add a flamingo and it fits into the background you provide nicely due to the well tuned style transfer. That part is cool. And it is impressive that sometimes the flamingo it adds is reflected in the water. But sometimes it isn't reflected. And it isn't up to you, it is up to it. And you can't tell it to add a reflection as a discrete step.

Look more carefully. This is more akin to a clipart finder, except if the clipart doesn't exist it uses the most similar thing in its training set to what it guesses you want as a starting point to synthesize new clipart from.

It doesn't add it in like an artist would and you can't control it at all. I don't know how to better express this.

This isn't unimpressive or un-useful but not quite as mind blowing on second glance.

Re: Dall-E 2

#404

Earlier quoted context omitted.

I'm not sure that's "compositing" except in the most abstract sense? But maybe that's the sense in which you mean it. I'd argue that at no point is there a representation of a "teddy bear" and "a background" that map closely to their visual representation - that are combined. (I'm aware I'm being imprecise so give me some leeway here)

This model's predecessor could do image editing with some help: https://arxiv.org/pdf/2112.10741.pdf so it could distinguish individual objects from backgrounds. Other ML models can definitely do that; it's called "panoptic segmentation".

Thank you! Fascinating, I didn't know about panoptic segmentation - that makes things much more interesting.

It really needs to expose the whole pipeline to become truly useful.

Re: Dall-E 2

#405
post #261

Earlier quoted context omitted.

Imagine asking it to generate a picture for "duck wearing a hat on Mars": First, it creates a random 10x10 pixel blurry image and asks a neural net: "Could this be a duck wearing a hat on Mars?" and the neural net replies "No, because all the pictures I've ever seen of Mars have lots of red color in them" so the system tweaks the pixels to make them more red, put some pixels in the center that have a plausible duck c…

Not the case, though in a handwave-y way, same idea - instead of iteratively scaling, you're iteratively denoising. See here, links out to the Cornell NLP PhD describe in even more detail: https://www.jpohhhh.com/articles/inflection-point-ml-art

I was kind of explaining how I picture the process in my head, fully aware that it isn't really possible to do an ELI5 on this stuff, and not really having an understanding of the technical details myself, either.

Re: Dall-E 2

#406
post #353

>We’ve limited the ability for DALL·E 2 to generate ... adult images. I think that using something like this for porn could potentially offer the biggest benefit to society. So much has been said about how this industry exploits young and vulnerable models. Cheap autogenerated images (and in the future videos) would pretty much remove the demand for human models and eliminate the related suffering, no? EDIT: typo

No. If people are exposed to stimuli, they will pursue increasingly stimulating versions of it. I.e., if they see artificial CP, they will often begin to become desensitized (habituated) and pursue real CP or even live children thereafter. Conversely, if people are not exposed to certain stimuli, they will never be able to conceptualize them, and thus will be unable to think about them. Obviously you cannot eliminate…

    If people are exposed to stimuli, they will pursue 
    increasingly stimulating versions of it.
This is not true in any kind of universal way.

If you enjoy car chases in movies, does that mean you're going to require more and more intense chase scenes, and then consume real-life crash footage, and ultimately progress to doing your own daredevil driving stunts in real life?

No, because at some point it's "enough."

Same with... literally anything we enjoy. Did you enjoy your lunch? Did you compulsively feel the need to work up to crazier and crazier lunches?

What about sex? Have you had sex? Do you feel the need to seek out crazier and crazier versions of it?

Re: Dall-E 2

#407
post #369
post #360

Earlier quoted context omitted.

> If people are exposed to stimuli, they will pursue increasingly stimulating versions of it. I.e., if they see artificial CP, they will often begin to become desensitized (habituated) and pursue real CP or even live children thereafter. I have accumulated tens of thousands of headshots in video games but have yet to ever shoot a single real person in the face. More importantly, I have never had the urge to seek out…

The point is more "can you conceive of a headshot before you've ever witnessed one?" And the assertion is, no. I should be explicit -- I am saying the exposure which makes one seek stimulus is merely a catalyst for deeper urges, not a generator of them as such. A certain level of inhibition (e.g. sociopathy) is required but IMO so is a prior conception of the deed. In your example, if someone is predisposed to wantin…

Is this just your own personal theory or opinion? Do you have some proof?

To put it as nicely as possible, this wildly contradicts reality as I have experienced it and observed others experiencing it.

Re: Dall-E 2

#408

Earlier quoted context omitted.

The tail end of programming will be the last thing to be replaced, maybe. I don’t see why CRUD apps get to hide under the umbrella of programming ultra-advanced AI.

Let me know when you can speak English to a computer and have it generate CRUD code that satisfies all engineering and design constraints. The AI will need to be dynamic enough to understand nuance, missing gaps in the requirements spec, have context on the application being built, able to suggest improvements on product design, know how to make changes through the same conversational interface, etc. Accomplishing th…

Is it that hard to do? Just design a solution that uses Alexa voice services to parse the vocal input via NLP and then invoke a lambda function to call a sagemaker or gpt-3 model to generate code. Granted it will take a little while to be perfect but are we really far from it?

Re: Dall-E 2

#409
post #353

>We’ve limited the ability for DALL·E 2 to generate ... adult images. I think that using something like this for porn could potentially offer the biggest benefit to society. So much has been said about how this industry exploits young and vulnerable models. Cheap autogenerated images (and in the future videos) would pretty much remove the demand for human models and eliminate the related suffering, no? EDIT: typo

No. If people are exposed to stimuli, they will pursue increasingly stimulating versions of it. I.e., if they see artificial CP, they will often begin to become desensitized (habituated) and pursue real CP or even live children thereafter. Conversely, if people are not exposed to certain stimuli, they will never be able to conceptualize them, and thus will be unable to think about them. Obviously you cannot eliminate…

Wow I didn't even think of this, that people could use this for something so horrifying. I'm relived that the geniuses behind this seem so smart that they even thought of this too and prohibit using the AI for sexual images.

> Our content policy does not allow users to generate violent, adult, or political content, among other categories. We won’t generate images if our filters identify text prompts and image uploads that may violate our policies. We also have automated and human monitoring systems to guard against misuse.

Re: Dall-E 2

#410
post #338

Earlier quoted context omitted.

Depends whether you think models should be able to generate cp. It's almost impossible to even give an affirmative answer to that question without making yourself a target. And as much as I err on the side of creator freedom, I find myself shying away from saying yes without qualifications. And if you don't allow cp, then by definition you require some censoring. At that point it's just a matter of where you censor,…

I don't think it's necessarily certain villainy for those who fight that fight as long as they are fighting it correctly. There's a huge case to be made that flooding the darknet with AI generated CP reduces the revictimization of those in authentic CP images, and would cut down on the motivating factors to produce authentic CP (for which original production is often a requirement to join CP distribution rings). As w…

Thought-provoking post, thanks.

Especially the part about maybe generating specifically tailored material to "train" folks. Although, while obviously moral instead of immoral like "gay conversion therapy", I wonder if it would be just as ineffective.

    and would cut down on the motivating factors 
    to produce authentic CP (for which original 
    production is often a requirement to join 
    CP distribution rings).
Hmmmmm. Will machine-generated "normal" (i.e., non-CP) porn really eliminate the motivating factors to produce normal porn?

I obviously can't speak for enjoyers of CP. But when watching normal porn, I think part of the thrill for many/most people is knowing that what's happening is real.

Another potential risk is that a flood of publicly available, machine-generated CP might actually help the producers and distributors of real CP by serving as camouflage. Finding and prosecuting the people who make real CP is difficult enough already. Now, imagine if the good guys couldn't even reliably tell what was real and there were 100000x as many fake images as real ones floating around.

Yikes.

Post reply on HN