Live data from Hacker News

DALL·E now available in beta

openai.com

371–380 of 579 posts

Re: DALL·E now available in beta

#371

I have been having a blast with DALL-E, spending about an hour a day trying out wild combinations and cracking my friends up. I cannot imagine getting bored of it; it's like getting bored with visual stimulus, or art in general. In fact, I've been glad to have a 50/day limit, because it helps me contain my hyperfocus instincts. The information about new pricing is, to me as someone just enjoying making crazy imagines…

> trying out wild combinations and cracking my friends up

Wait until the next edition comes out where it automatically learns the sorts of things that crack you up and starts generating them without any input from you.

Re: DALL·E now available in beta

#372
post #76

I fully expect stock image sites to be swamped by DALL-E generated images that match popular terms (e.g. "business person shaking hands"). Generate the image for $0.15. Sell it for $1.00.

They won't. DALL-E images are mostly not as high quality. The high quality stuff which everyone has been sharing is result of lots of cherry picking.

Even the high quality stuff still can't do human faces right.

Re: DALL·E now available in beta

#373
post #58

I was supposed to be making a video game, but got a bit sidetracked when DALL·E came out and made this website on the side: http://dailywrong.com/ (yes I should get SSL). It's like The Onion, but all the articles are made with GPT-3 and DALL·E. I start with an interesting DALL·E image, then describe it to GPT-3 and ask it for an Onion-like article on the topic. The results are surprisingly good.

Thanks, finally a legit news publication :) This was really funny :) http://dailywrong.com/man-finally-comfortable-just-holding-a...

So the other men in the pictures are the uncomfortable ones?

Re: DALL·E now available in beta

#374
post #255

Earlier quoted context omitted.

The monkey selfie was not derived from millions of existing works, and that is the difference. If an artist has a well-known art style, and this algorithm was trained on it and can copy that style, would the artist have grounds to sue? I don't know.

Well, music is not "pictures" but Marvin Gaye's family got 5 million because Blurred Lines sounds similar enough to a Marvin Gaye song (even though it was not a sample): https://en.wikipedia.org/wiki/Pharrell_Williams_v._Bridgepor...

[deleted]

Re: DALL·E now available in beta

#375

Earlier quoted context omitted.

In the unCLIP/DALL-E 2 paper[0], they train the encoder/decoder with 650M/250M images respectively. The decoder alone has 3.5B parameters, and the combined priors with the encoder/decoder are the in the neighborhood of ~6B parameters. This is large, but small compared to the name-brand "large language models" (GPT3 et. al.) This means the parameters of the trained model fit in something like 7GB (decoder only, half-p…

> This means the parameters of the trained model fit in something like 7GB (decoder only, half-precision floats) to 24GB (full model, full-precision) > you would probably want an enterprise cloud/data-center GPU like an NVIDIA A100, especially if running batches of more than one image. That doesn't seem so bad. looks up price of NVIDIA A100 - $20,000 oh...ok I'll probably just pay for the service then

p4d.24xlarge is only $33/hr! And you get 400 Gbe so it should be quick to load.

Re: DALL·E now available in beta

#376

Earlier quoted context omitted.

They won't. DALL-E images are mostly not as high quality. The high quality stuff which everyone has been sharing is result of lots of cherry picking.

Even the high quality stuff still can't do human faces right.

They avoided using real human faces in the training data.

Re: DALL·E now available in beta

#377

Earlier quoted context omitted.

Honestly I would rather that they not try. I don't understand why a computer tool has to be held to a political standard.

There are legitimate reasons to reduce externalizations of societies innate biases. A mortgage AI that calculates premiums for the public shouldn't bias against people with historically black names, for example. This problem is harder to tackle because it is difficult to expose and resign the "latent space" that results in these biases; it's difficult to massage the ML algo's to identify and remove the pathways that…

[deleted]

Re: DALL·E now available in beta

#378

Earlier quoted context omitted.

I'm already bored of it. When you have everything, you have nothing.

I don't know how to say this without sounding like a jerk, even if I bend over backwards to preface that this isn't my intent: this statement says more about your creativity and curiosity than a ceiling on how entertaining DALL-E can be to someone who could keep multiple instances busy, like grandma playing nine bingo cards at once. Knowing that it will only get better - animation cannot be far behind - makes me feel…

Dall-e has novelty, but no intent, meaning, originality. Yes the author can be creative at generating prompts, but visually I haven’t seen it generate anything that feels artistically interesting. If you want pre-existing concepts in novel combinations then yes it works.

It’s good at “in the style of” but there’s no “in a new style”.

It has a house style too that tends to feel Reddit-like.

Re: DALL·E now available in beta

#379

Earlier quoted context omitted.

> inoffensive tool. Wouldn't that result end up being like "inoffensive art" or "inoffensive comedy"? Bland, boring and Corporate-PC.

Being offensive is only one way to be interesting. There are others, like being clever, or being absurd, or being goofy, or being poignant, or being refreshing. Of the good stuff, offensive humor is only a tiny slice.

offensive to whom is the sticking point when it comes to comedy

it takes a special talent to please everybody

Re: DALL·E now available in beta

#380
post #14

One of the commercial use cases this post mentions is authors who want to add illustrations to children's stories. I wonder if there is a way for DALL-E to generate a character, then persist that character over subsequent runs. Otherwise, it would be pretty difficult to generate illustrations that depict a coherent story. Example ... Image 1 prompt: A character named Boop, a green alien with three arms, climbs out of…

You cannot. But a workaround would be to say something like “generate an alien in three different poses— running, walking, waving” Then use inpainting to only preserve that pose and generate new content around it. It’s definitely not perfect.

You can do better than this. Draw/generate your character.

Then put that at the side of a transparent image, and use as the prompt, "Two identical aliens side by side. One is jumping"

Post reply on HN