Live data from Hacker News

DALL·E now available in beta

openai.com

31–40 of 579 posts

Re: DALL·E now available in beta

#31
post #23

> Starting today, users get full usage rights to commercialize the images they create with DALL·E, including the right to reprint, sell, and merchandise. This includes images they generated during the research preview. So DALL·E 2 is going to restart, revive and cause another renaissance of fully automated and mass generated NFTs, full of derivatives and remixing etc to pump up the crypto NFT hype squad? Either way,…

I don't see why there's any credible reason to expect that DALL-E will do anything at all to help those promoting the NFT silliness. Two separate issues.

Re: DALL·E now available in beta

#33
post #27
post #22

Earlier quoted context omitted.

Watermarks are still there and resolution still 1024x1024.

I wonder if they have plans to allow SVG exports in the future. I mean, the file size would probably be ridiculous in a lot of the cases, but for my use case I wouldn't mind it. And sucks about the watermark, maybe they will introduce an option to pay for removing it.

SVG exports would only be meaningful if the model is generating vector images, which are then converted to bitmaps. I highly doubt that's the case, but perhaps someone who has actually looked at the model structure can confirm?

Re: DALL·E now available in beta

#34
post #28

Something I haven’t seen anyone talking about with these huge models: how do future models get trained when more content online is model generated to start with? Presumably you don’t wanna train a model on autogenerated images or text, but you can’t necessarily know which is which.

Makes me think of Ouroboros

https://en.m.wikipedia.org/wiki/Ouroboros

Re: DALL·E now available in beta

#35
> Reducing bias: We implemented a new technique so that DALL·E generates images of people that more accurately reflect the diversity of the world’s population. This technique is applied at the system level when DALL·E is given a prompt about an individual that does not specify race or gender, like “CEO.”

Will it do it "more accurately" as they claim? As in, if 90% of CEOs are male, then the odds of a CEO being male in a picture is 90%? Or less "accurately reflect the diversity of the world’s population" and show what they would like the real world to be like?

Re: DALL·E now available in beta

#36
post #28

Something I haven’t seen anyone talking about with these huge models: how do future models get trained when more content online is model generated to start with? Presumably you don’t wanna train a model on autogenerated images or text, but you can’t necessarily know which is which.

This should be a step in cleaning your data to begin with. If you don't know the providence of your data then you shouldn't be even training with it.

Getting humans to refine your data is the best solution right now and many companies and researches go with this approach.

Re: DALL·E now available in beta

#37

I'm blown away by this: "Starting today, users get full usage rights to commercialize the images they create with DALL·E, including the right to reprint, sell, and merchandise. This includes images they generated during the research preview." I assumed this was going to be the sticking point for wider usage for a long time. They're now saying that you have full rights to sell Dall-E 2 creations?

Previously, OpenAI asserted they owned the generated images, so the new licensing is a shift in that aspect. GPT-3 also has a "you own the content" clause as well. Of course, that clause won't deter a third party from filing a lawsuit against you if you commercialize a generated image too close to something realistic, as the copyrights of AI generated content still hasn't been legally tested.

AFAIK only people can own copyright (the monkey selfie case tested this), and machine-generated outputs don't count as creative work (you can't write an algorithm that generates every permutation of notes and claim you own every song[1]), so DALL-E-generated images are most likely copyright-free. I presume OpenAI only relies on terms of service to dictate what users are allowed to do, but they can't own the images, and neither can their users.

[1]: https://felixreda.eu/2021/07/github-copilot-is-not-infringin...

Re: DALL·E now available in beta

#39
post #27
post #22

Earlier quoted context omitted.

Watermarks are still there and resolution still 1024x1024.

I wonder if they have plans to allow SVG exports in the future. I mean, the file size would probably be ridiculous in a lot of the cases, but for my use case I wouldn't mind it. And sucks about the watermark, maybe they will introduce an option to pay for removing it.

SVG isn't really possible with the model architecture they're using. The diffusion+upscaling step basically outputs 1024x1024 pixels; at no point does the model have a vector representation.

I suppose it's possible that at some point they'll try to make an image -> svg translation model?

Re: DALL·E now available in beta

#40

I'm blown away by this: "Starting today, users get full usage rights to commercialize the images they create with DALL·E, including the right to reprint, sell, and merchandise. This includes images they generated during the research preview." I assumed this was going to be the sticking point for wider usage for a long time. They're now saying that you have full rights to sell Dall-E 2 creations?

And I just used it to create cover art for a book published in Amazon :)

https://twitter.com/nutanc/status/1549798460290764801?s=20&t...

Post reply on HN