Live data from Hacker News

DALL·E now available in beta

openai.com

261–270 of 579 posts

Re: DALL·E now available in beta

#261
post #58

I was supposed to be making a video game, but got a bit sidetracked when DALL·E came out and made this website on the side: http://dailywrong.com/ (yes I should get SSL). It's like The Onion, but all the articles are made with GPT-3 and DALL·E. I start with an interesting DALL·E image, then describe it to GPT-3 and ask it for an Onion-like article on the topic. The results are surprisingly good.

Thanks, finally a legit news publication :)

This was really funny :)

http://dailywrong.com/man-finally-comfortable-just-holding-a...

Re: DALL·E now available in beta

#262

I'm blown away by this: "Starting today, users get full usage rights to commercialize the images they create with DALL·E, including the right to reprint, sell, and merchandise. This includes images they generated during the research preview." I assumed this was going to be the sticking point for wider usage for a long time. They're now saying that you have full rights to sell Dall-E 2 creations?

Previously, OpenAI asserted they owned the generated images, so the new licensing is a shift in that aspect. GPT-3 also has a "you own the content" clause as well. Of course, that clause won't deter a third party from filing a lawsuit against you if you commercialize a generated image too close to something realistic, as the copyrights of AI generated content still hasn't been legally tested.

As far as I can tell they still own the images they just license your use of them commercially.

Re: DALL·E now available in beta

#263

I'm blown away by this: "Starting today, users get full usage rights to commercialize the images they create with DALL·E, including the right to reprint, sell, and merchandise. This includes images they generated during the research preview." I assumed this was going to be the sticking point for wider usage for a long time. They're now saying that you have full rights to sell Dall-E 2 creations?

I think they are reacting to competition. MidJourney is amazing, was easier to get into, gives you commercial rights, and frankly I found more fun to use and even better output in most instances.

The only thing I don’t like about MidJourney is the Discord based interface. I think I can grok why Dave chose this route as it bakes in an active community element and allows users to pick up prompt engineering techniques osmotically… but I’d prefer a clean DALL-E style app and cli / api access.

Re: DALL·E now available in beta

#264

Earlier quoted context omitted.

There are legitimate reasons to reduce externalizations of societies innate biases. A mortgage AI that calculates premiums for the public shouldn't bias against people with historically black names, for example. This problem is harder to tackle because it is difficult to expose and resign the "latent space" that results in these biases; it's difficult to massage the ML algo's to identify and remove the pathways that…

That's great, but by doing so you are also inadvertently favoring, in your example, the people with black names. For example, Chinese people save on average, 50 times more than Americans according to the Fed [1]. That would mean they would generally be overrepresented in loan approvals because they have a better balance sheet. Does that necessarily mean that Americans are discriminated against in the approval process…

I would agree that it is not.

The government, and many people, have moved the definition and goal posts; so that anything that has the end result of a non-proportional uniformity can be labeled and treated as bias.

Ultimately it is a nuanced game; is discriminating against certain clothing or hair-styles racist? Of course. Yet, neither of those are explicitly tied to one's skin color or ethnicity, but are an indirect associative trait because of culture.

In America, we have intentionally muddled the waters of demarcation between culture and race, and are starting to see the cost of that.

Re: DALL·E now available in beta

#265

I'm blown away by this: "Starting today, users get full usage rights to commercialize the images they create with DALL·E, including the right to reprint, sell, and merchandise. This includes images they generated during the research preview." I assumed this was going to be the sticking point for wider usage for a long time. They're now saying that you have full rights to sell Dall-E 2 creations?

Previously, OpenAI asserted they owned the generated images, so the new licensing is a shift in that aspect. GPT-3 also has a "you own the content" clause as well. Of course, that clause won't deter a third party from filing a lawsuit against you if you commercialize a generated image too close to something realistic, as the copyrights of AI generated content still hasn't been legally tested.

they still own the generated content, only grant usage. I have mixed feelings about this confused approach, it won’t last long.

> …you own your Prompts and Uploads, and you agree that OpenAI owns all Generations…

Re: DALL·E now available in beta

#266

Two questions: (1) Any opinions on if removing the watermark is possible? Is doing so against the terms of service? (2) Appears the output is still at 1024x1024 - what are options to upscale the resolution, for example would OpenCV super resolution work?

It is possible, they confirmed on discord you can remove the watermark. Yep... The output is an issue, I'd like to pay if that was an upgrade.

Annoying that if removing the watermark is allowed that it is even inserted. Imagine if Adobe did that.

Here’s more information on super resolution options beyond what Adobe already offers:

(1) List of options current options for super resolutions:

https://upscale.wiki/wiki/Different_Neural_Networks

(2) Older example of one way to benchmark:

https://docs.opencv.org/4.x/dc/d69/tutorial_dnn_superres_ben...

Re: DALL·E now available in beta

#267
post #87

Interesting. I got access couple weeks ago (was on waitlist since the initial announcement) and frankly as much as really want to be excited and like it, DALL-E ended up being a bit underwhelming. IMHO - often results that produced are of low quality (distorted images, or quite wacky representation of the query). Some styles of imagery are certainly a better fit for being generated by DALL-E, but as far as commercial…

I also got access a couple of weeks ago and I can't fathom how anyone could be underwhelmed by it. What were you expecting?

Dalle seems to only have a few "styles" of drawing that it is actually "good" at. It is particularly strong at these styles but disappointingly underwhelming at anything else, and will actively fight you and morph your prompt into one of these styles even when given an inpainting example of exactly what you want.

It's great at photorealistic images like this: https://labs.openai.com/s/0MFuSC1AsZcwaafD3r0nuJTT, but it's intentionally lobotomized to be bad at faces, and often has an uncanny valley feel in general, like this: https://labs.openai.com/s/t1iBu9G6vRqkx5KLBGnIQDrp (never mind that it's also lobotomized to be unable to recognize characters in general). It's basically as close to perfect as an AI can be at generating dogs and cats though, but anything else will be "off" in some meaningful ways.

It has a particular sort of blurry, amateur oil painting digital art style it often tries to use for any colorful drawings, like this: https://labs.openai.com/s/EYsKUFR5GvooTSP5VjDuvii2 or this: https://labs.openai.com/s/xBAJm1J8hjidvnhjEosesMZL . You can see the exact problem in the second one with inpainting: it utterly fails at the "clean" digital art style, or drawing anything with any level of fine detail, or matching any sort of vector art or line art (e.g. anime/manga style) without loads of ugly, distracting visual artifacts. Even Craiyon and DALLE-mini outperform it on this. I've tried over 100 prompts to get stuff like that to generate and have not had a single prompt that is able to generate anything even remotely good in that style yet. It seems almost like it has a "resolution" of detail for non-photographic images, and any detail below a certain resolution just becomes a blobby, grainy brush stroke, e.g. this one: https://labs.openai.com/s/jtvRjiIZRsAU1ukofUvHiFhX , the "fairies" become vague colored blobs here. It can generate some pretty ok art in very specific styles, e.g. classical landscape paintings: https://labs.openai.com/s/6rY7AF7fWPb5wWiSH0rAG0Rm , but for anything other than this generic style it disappoints hard.

The other style it is ok at is garish corporate clip art, which is unremarkable and there's already more than enough clip art out there for the next 1000 years of our collective needs -- it is nevertheless somewhat annoying when it occasionally wastes a prompt generating that crap because you weren't specific that you wanted "good" images of the thing you were asking for.

The more I use DALLE-2 the more I just get depressed at how much wasted potential it has. It's incredibly obvious they trimmed a huge amount of quality data and sources from their databases for "safety" reasons, and this had huge effects on the actual quality of the outputs in all but the most mundane of prompts. I've got a bunch more examples of trying to get it to generate the kind of art I want (cute anime art, is that too much to ask for?) and watching it fail utterly every single time. The saddest part is when you can see it's got some incredible glimpse of inspiration or creative genius, but just doesn't have the ability to actually follow through with it.

Re: DALL·E now available in beta

#268

Earlier quoted context omitted.

> If an artist has a well-known art style, and this algorithm was trained on it and can copy that style, would the artist have grounds to sue? I don't know. While nothing has been commercialized yet on the DALLE2 subreddit, I know that it can do Dave Choe's work remarkably well. I also saw Alex Gray's work to be close, but not really identical either. It wasn't as intricate as his work is. It will be interesting if t…

I'm going to guess there's not going to be much value placed on anything out of DALLE for a long while. Digital art is typically worth much less than physical art and I would say these GAN images are going to worth less than digital art generated by human hand. There will be outliers of course but I would be shocked if there's much of a market in it for at least the present.

When these tools can generate layered tiff/psd images, polygon meshes and automate UV packing; then we’ll be talking.

Re: DALL·E now available in beta

#269

Earlier quoted context omitted.

Imagine you're in South Korea (or any other ethnically homogenous country). Do you want "black" "female" randomly appended to your input?

If I was using this in South Korea, how is showing all white people any better than showing whites, blacks, latinos and asians?

You would presumably input “South Korean CEO”. DALL-E would then unhelpfully add “black” “female” without your knowledge.

Re: DALL·E now available in beta

#270
How do you interface with DALL-E?

For MidJourney I was painfully surprised to find that everything is done through chat messages on a Discord server.

I'm not a paid member, so I have to enter my prompts in public channels. It's extremely easy to lose your own prompts in the rapidly flowing stream of prompts going on. I can kind of see why they did it that way--maybe, if I squint really hard--to try to promote visibility and community interaction, but it's just not happening. It's hard enough to find my own images, say nothing about follow what someone else is doing. This is literally the worst user experience I have ever had with a piece of software.

There are dozens of channels. It's so spammy, doing it through Discord. It's constantly pinging new notifications and I have to go through and manually mute each and every one of the channels. Then they open a few dozen more. Rinse. Repeat.

I understand paid users can have their own channels to generate images, but I really don't see the point in paying for it when, even subtracting the firehose of prompts and images, it's still an objectively shitty interface to have to do everything through Discord chat messages.

Post reply on HN