Live data from Hacker News

Dall-E 2

openai.com

411–420 of 511 posts

Re: Dall-E 2

#411

Is there a geometric model relative to this? EG: "corgi near the fireplace" but the output is a 3d model of the corgi and fireplace with shaders rather than an image.

Wait until you see the same concept combined with NeRF idea. The output won’t be 3d shapes but another model that can generate realistic and geometrically consistent images of a scene viewed from different angles.

Re: Dall-E 2

#412

A few comments by someone who's spent way too much time in the AI-generated space: * I recommend reading the Risks and Limitations section that came with it because it's very through: https://github.com/openai/dalle-2-preview/blob/main/system-c... * Unlike GPT-3, my read of this announcement is that OpenAI does not intend to commercialize it, and that access to the waitlist is indeed more for testing its limits (and…

Regarding cherry-picking, the images of astronauts on horses look stunning, except for their hands. There's something seriously wrong with their hands. Maybe give it another five years, a few more $billion and a few more petabytes/flops and it will be good. Then finally everyone can generate art for their own Magic: the Gathering cards. (That's the end goal, right?)

As I keep telling people: "hands are hard". This is why I went so far as to make a hand-specific dataset ("PALM" https://www.gwern.net/Crops#palm which of course now everyone is going to confuse with 'PaLM'...). Hands are just way too variable to learn easily.

My dataset is a start, but it may benefit from focused training, the way Facebook's new Make-A-Scene https://arxiv.org/abs/2203.13131#facebook (not DALL-E 2 quality but not far from it) has focused losses on faces.

Re: Dall-E 2

#413
post #3

Preventing Harmful Generations We’ve limited the ability for DALL·E 2 to generate violent, hate, or adult images. By removing the most explicit content from the training data, we minimized DALL·E 2’s exposure to these concepts. We also used advanced techniques to prevent photorealistic generations of real individuals’ faces, including those of public figures. "And we've also closed off a huge range of potentially int…

I instinctively want to "flip the sign" on all of the automated controls they put in, just out of the morbid interest to see what comes out. The moment you have a "avoid_harm_to_humans:bool" training parameter, someone's going to set it to -1. Their document about all the measures they took to prevent unethical use is also a document about how to use a re-implementation of their system unethically. They literally hir…

The kind of measures they are taking, like simply deleting wholesale anything problematic, don't really have a '-1'.

But amusingly, exactly that did happen in one of their GPT experiments! https://openai.com/blog/fine-tuning-gpt-2/

Re: Dall-E 2

#414
post #142

Something about this makes me nauseous. Perhaps is the fact that soon the market value for creatives is going to fall to a hair about zero for all but the most famous. We will be all the poorer for it when 95% of images you see are AI generated. There will be niches of course but in a few short years it'll be over for a huge swathe of creative professionals who are already struggling. Some of the images also hit me w…

I paid $1500 for a commissioned painting from an artist I respect and follow as a birthday present for a friend. The painting meant something to me because I worked with the artist to have some input about what kind of a person my friend is, what kind of features I want to see in the painting and how I want it to feel. The artist gave me 5 different sketches and we had tons of back and forth. The process and the act…

You would still work with the model back and forth with editing the prompt and image to figure out what it meant, what kind of person your friend is, what you were looking for (even the things you couldn't verbalize and only knew when you spotted them in a large array of diverse samples, the sort you could never hire a human to do), how you wanted it to feel... And then you would also have $1500 for another gift. Personally, I would prefer the scenario in which I received a unique meaningful painting from my friend, plus $1500.

Re: Dall-E 2

#415

The most interesting item to me is the variations on the garden shop and bathroom sink idea. The realism of these leaks the AI lacking intuition of the requirements. This makes for a number of nonsensical designs that look right at first like: This Sink lacks sensical faucets. https://cdn.openai.com/dall-e-2/demos/variations/modified/ba... This doorway is downright impossible https://cdn.openai.com/dall-e-2/demos/var…

[deleted]

Re: Dall-E 2

#416
perhaps in the not so distant future, we can simply feed a movie script to the program and out comes a feature film.

Re: Dall-E 2

#417
post #369
post #360

Earlier quoted context omitted.

> If people are exposed to stimuli, they will pursue increasingly stimulating versions of it. I.e., if they see artificial CP, they will often begin to become desensitized (habituated) and pursue real CP or even live children thereafter. I have accumulated tens of thousands of headshots in video games but have yet to ever shoot a single real person in the face. More importantly, I have never had the urge to seek out…

The point is more "can you conceive of a headshot before you've ever witnessed one?" And the assertion is, no. I should be explicit -- I am saying the exposure which makes one seek stimulus is merely a catalyst for deeper urges, not a generator of them as such. A certain level of inhibition (e.g. sociopathy) is required but IMO so is a prior conception of the deed. In your example, if someone is predisposed to wantin…

> "can you conceive of a headshot before you've ever witnessed one?"

Am totally blind, have never been able to see, can still conceive of a headshot. So, yes?

Re: Dall-E 2

#418
post #414
post #142

Earlier quoted context omitted.

I paid $1500 for a commissioned painting from an artist I respect and follow as a birthday present for a friend. The painting meant something to me because I worked with the artist to have some input about what kind of a person my friend is, what kind of features I want to see in the painting and how I want it to feel. The artist gave me 5 different sketches and we had tons of back and forth. The process and the act…

You would still work with the model back and forth with editing the prompt and image to figure out what it meant, what kind of person your friend is, what you were looking for (even the things you couldn't verbalize and only knew when you spotted them in a large array of diverse samples, the sort you could never hire a human to do), how you wanted it to feel... And then you would also have $1500 for another gift. Per…

Don't entirely disagree with what you're saying - I believe Dall-E 6 or whatever will get to that level of sophistication. One more thing though - I felt the painting is worth more because the artist toiled for it. It's like a "lofi 10 hour soundtrack" on youtube vs. an album from an acclaimed artist. I listen to each song 100 times over from the latter meanwhile the lofi video just plays in the background. Knowing someone toiled over it and put their heart and soul into the art gives it the value, for me.

Re: Dall-E 2

#419

A few comments by someone who's spent way too much time in the AI-generated space: * I recommend reading the Risks and Limitations section that came with it because it's very through: https://github.com/openai/dalle-2-preview/blob/main/system-c... * Unlike GPT-3, my read of this announcement is that OpenAI does not intend to commercialize it, and that access to the waitlist is indeed more for testing its limits (and…

The Risks and Limitations section is particularly interesting to me. It's like a time capsule of society's current fears about technology. They talk about many ways this tech could be misused, but I don't think they've even scratched the surface.

An example off the top of my head: this could be used as advertising or recruitment for controversial organizations or causes. Would it be wrong for the USA to use this for military recruitment? Israel? Ukraine? Russia?

Another example: this could be used to glorify and reinforce actions which our society does not consider to immoral but other societies - or our own future society - will. It wasn't long ago that the US and Europe did a full 180 on their treatment of homosexuality. Will we eventually change our minds about eating meat, driving cars, etc?

Have they gone too far in a desperate bid to prevent the AI from being capable of harm? Have they not gone far enough? I don't know. If I was that worried about something being misused, I don't think I could ever bring myself to work on it in the first place. But I suppose the onward march of technology is inevitable.

Re: Dall-E 2

#420
post #5

Some freely available models GLID-3: https://colab.research.google.com/drive/1x4p2PokZ3XznBn35Q5B... and a new Latent Diffusion notebook: https://colab.research.google.com/github/multimodalart/laten... have both appeared recently and are getting remarkably close to the original Dall-E (maybe better as I can't test the real thing...) So - this was pretty good timing if OpenAI want to appear to be ahead of the pack. Of…

a cow and a farmer in their field looking at the sky
Post reply on HN