Live data from Hacker News

DALL·E now available in beta

openai.com

91–100 of 579 posts

Re: DALL·E now available in beta

#91

> Reducing bias: We implemented a new technique so that DALL·E generates images of people that more accurately reflect the diversity of the world’s population. This technique is applied at the system level when DALL·E is given a prompt about an individual that does not specify race or gender, like “CEO.” Will it do it "more accurately" as they claim? As in, if 90% of CEOs are male, then the odds of a CEO being male i…

It's also odd since you'd think that this would be an issue solved by training with representative images in the first place.

If you used good input you'd expect an appropriate output, I don't know why manual intervention would be necessary unless it's for other purposes than stated. I suspect this is another case where "diversity" simply means "less whites".

Re: DALL·E now available in beta

#92

That's disappointing given up until this point you could have 50 free uses per 24h. I expected it to get monetized eventually, but not so fast and drastically. Well, still had my fun and have to say the creations are so good it's often mind blowing there's an AI behind it.

they're a non-profit so the price is probably still dirt cheap

Re: DALL·E now available in beta

#93
post #23

> Starting today, users get full usage rights to commercialize the images they create with DALL·E, including the right to reprint, sell, and merchandise. This includes images they generated during the research preview. So DALL·E 2 is going to restart, revive and cause another renaissance of fully automated and mass generated NFTs, full of derivatives and remixing etc to pump up the crypto NFT hype squad? Either way,…

I have tried playing around with the beta access to make it generate NFT art with different prompts, but in vail. I think it has not been trained on NFT art (crypto punks and so on).

> I think it has not been trained on NFT art (crypto punks and so on).

How exactly are you defining NFT art?

I mean, it can literately be anything: Dorsey sold a screencap of his 1st tweet, Nadya from Pussy Riot did some creative stuff, and the Ape crap was the bulk of this stuff that got passed around.

I think what can be gleaned from that short-lived non-sense is that value is subjective and that the quality of a valuabe piece of 'art' is equally as hard to define. Much the same with its predecessor: cryptokitties.

Re: DALL·E now available in beta

#94
post #37

Earlier quoted context omitted.

Previously, OpenAI asserted they owned the generated images, so the new licensing is a shift in that aspect. GPT-3 also has a "you own the content" clause as well. Of course, that clause won't deter a third party from filing a lawsuit against you if you commercialize a generated image too close to something realistic, as the copyrights of AI generated content still hasn't been legally tested.

AFAIK only people can own copyright (the monkey selfie case tested this), and machine-generated outputs don't count as creative work (you can't write an algorithm that generates every permutation of notes and claim you own every song[1]), so DALL-E-generated images are most likely copyright-free. I presume OpenAI only relies on terms of service to dictate what users are allowed to do, but they can't own the images, a…

> DALL-E-generated images are most likely copyright-free

The US Copyright Office did make a ruling that might suggest that recently[1], but crucially, in that case, the AI "didn't include an element of human authorship." The board might rule differently about DALL-E because the prompts do provide an opportunity for human creativity.

And there's another important caveat that the felixreda.eu link seems to miss. DALL-E output, whether or not it's protected by copyright, can certainly infringe other copyrights, just like the output of any other mechanical process. In short, Disney can still sue if you distribute DALL-E generated images of Marvel characters.

1: https://www.theverge.com/2022/2/21/22944335/us-copyright-off...

Re: DALL·E now available in beta

#95
post #23

> Starting today, users get full usage rights to commercialize the images they create with DALL·E, including the right to reprint, sell, and merchandise. This includes images they generated during the research preview. So DALL·E 2 is going to restart, revive and cause another renaissance of fully automated and mass generated NFTs, full of derivatives and remixing etc to pump up the crypto NFT hype squad? Either way,…

I have tried playing around with the beta access to make it generate NFT art with different prompts, but in vail. I think it has not been trained on NFT art (crypto punks and so on).

Heads up: I think you meant "in vain" rather than "in vail". However, a similar phrase is "to no avail" which also means that something was not successful.

Re: DALL·E now available in beta

#96
post #37

Earlier quoted context omitted.

AFAIK only people can own copyright (the monkey selfie case tested this), and machine-generated outputs don't count as creative work (you can't write an algorithm that generates every permutation of notes and claim you own every song[1]), so DALL-E-generated images are most likely copyright-free. I presume OpenAI only relies on terms of service to dictate what users are allowed to do, but they can't own the images, a…

The monkey selfie was not derived from millions of existing works, and that is the difference. If an artist has a well-known art style, and this algorithm was trained on it and can copy that style, would the artist have grounds to sue? I don't know.

> If an artist has a well-known art style, and this algorithm was trained on it and can copy that style...

A lawyer could argue that the algorithm is producing a derivative work of the copyrighted input.

Re: DALL·E now available in beta

#97
post #28

Something I haven’t seen anyone talking about with these huge models: how do future models get trained when more content online is model generated to start with? Presumably you don’t wanna train a model on autogenerated images or text, but you can’t necessarily know which is which.

I wonder if human artists can demand that their work not be used for modelling. So as the robots are stuck using older styles for their creations, the humans will keep creating new styles of art.

Re: DALL·E now available in beta

#98
post #37

Earlier quoted context omitted.

AFAIK only people can own copyright (the monkey selfie case tested this), and machine-generated outputs don't count as creative work (you can't write an algorithm that generates every permutation of notes and claim you own every song[1]), so DALL-E-generated images are most likely copyright-free. I presume OpenAI only relies on terms of service to dictate what users are allowed to do, but they can't own the images, a…

If this were a concern, a user can easily bypass this by having a work-for-hire person add a minor transform layer on top of the DALL-E generated images right?

Wouldn't it have to meet the threshold of being a "transformative" work?

https://en.wikipedia.org/wiki/Transformative_use

Re: DALL·E now available in beta

#99

> Reducing bias: We implemented a new technique so that DALL·E generates images of people that more accurately reflect the diversity of the world’s population. This technique is applied at the system level when DALL·E is given a prompt about an individual that does not specify race or gender, like “CEO.” Will it do it "more accurately" as they claim? As in, if 90% of CEOs are male, then the odds of a CEO being male i…

Will it reduce bias across all fields? Or only ones that are desirable? How about historical?

"A photo of a group of soldiers from WW2 celebrating victory over nazi CEOs and plumbers".

Re: DALL·E now available in beta

#100
post #14

One of the commercial use cases this post mentions is authors who want to add illustrations to children's stories. I wonder if there is a way for DALL-E to generate a character, then persist that character over subsequent runs. Otherwise, it would be pretty difficult to generate illustrations that depict a coherent story. Example ... Image 1 prompt: A character named Boop, a green alien with three arms, climbs out of…

You can't do that. I can't see this working well for children's book illustrations unless the story was specifically tailored in a way that makes continuity of style and characters irrelevant.

As an aside, Ursula Vernon did pretty well under the constraint you described. She set a comic in a dreamscape and used AI to generate most of the background imagery: https://twitter.com/UrsulaV/status/1467652391059214337

It's not the "specify the character positions in text" proposed, but still a neat take on using this sort of AI for art.

Post reply on HN