Live data from Hacker News

A prompt engineering guide for DALLE-2

dallery.gallery

31–40 of 65 posts

Re: A prompt engineering guide for DALLE-2

#32

Earlier quoted context omitted.

I've played with Midjourney for awhile and just got my invite to DALL-E last night. One thing I think is really cool about Midjourney is the ability to give it image URLs as part of the prompts. I can't say I've had tremendous success with it, and it still feels a little half-baked, but I wish DALL-E had something along those lines. (Unless it does and I'm missing it). It's much easier to show examples of a particula…

You can upload an image to DALL-E, edit it and add a prompt to it as well.

DALLE2 isn't as flexible as the more open colab notebooks here; you can do "variations" of an image but you can't edit an image except through inpainting, so it's hard to generate "AI art" style images of the kind Midjourney and Diffusion are good at.

It also won't allow uploading images with faces in them.

Re: A prompt engineering guide for DALLE-2

#33

Earlier quoted context omitted.

You can have free access here to dalle at https://replicate.com/nicholascelestin/dalle-mega and dalle mini https://huggingface.co/spaces/dalle-mini/dalle-mini

My propmt of 'penguin smoking a bong' does not disappoint on either, although hugging face more accurately portrayed the act of smoking, while replicate gave me images of penguin shaped bongs

Replicate is a newer version being trained on the same data set so it should theoretically catch up soon, no guarantees of course.

Re: A prompt engineering guide for DALLE-2

#34
post #23
post #19

Dall-e still has a lot of work to be done with face construction. Maybe that’s a feature not a bug.

It's seems to be by far the best of any other drawing AI besides the "this person does not exist" series, but those are quite specialized. You could be right though. It does "digital art" well, but realistic faces poorly, and they slap down lots of restrictions to avoid deepfaking.

Google's internal models (Imagen and Parti) are much better. It looks like DALLE2 is just not big enough to accurately draw faces, which are very detailed things.

"This person doesn't exist" uses StyleGAN which can definitely do faces, but can't do general pictures.

Re: A prompt engineering guide for DALLE-2

#35
post #23

Earlier quoted context omitted.

It's seems to be by far the best of any other drawing AI besides the "this person does not exist" series, but those are quite specialized. You could be right though. It does "digital art" well, but realistic faces poorly, and they slap down lots of restrictions to avoid deepfaking.

Google's internal models (Imagen and Parti) are much better. It looks like DALLE2 is just not big enough to accurately draw faces, which are very detailed things. "This person doesn't exist" uses StyleGAN which can definitely do faces, but can't do general pictures.

Are there samples of faces by the Google models? The websites don't seem to show any. Though their 20B samples are incredibly impressive.

Re: A prompt engineering guide for DALLE-2

#37
post #22
post #19

Dall-e still has a lot of work to be done with face construction. Maybe that’s a feature not a bug.

I think they’re not training on faces on purpose.

You are probably right. Having used it I sometimes get images with white polygons covering the faces of people as if they have been blanked out.

Re: A prompt engineering guide for DALLE-2

#38

Earlier quoted context omitted.

Hang in there — I only got my invitation a couple days ago. They're still rolling out invitations at a steady pace. But, just as a side note, one of the first things they tell you is that they own the full copyright for any images you generate. You definitely have to play around with prompts to get a feel for how it works and to maximize the chance of getting something closer to what you want.

I don't think the provider of an AI image generator service can decide they own the copyrights to it (perhaps they can require you assign the copyrights, though it may not even be copyrightable?), only courts can (and they decided the person setting up cameras for monkeys didn't own the copyrights to the monkey photos)?

DALL-E needs human input to start generating, the monkey pressed the shutter all on its own.
Post reply on HN