Live data from Hacker News

DALL·E now available in beta

openai.com

121–130 of 579 posts

Re: DALL·E now available in beta

#121
Two questions:

(1) Any opinions on if removing the watermark is possible? Is doing so against the terms of service?

(2) Appears the output is still at 1024x1024 - what are options to upscale the resolution, for example would OpenCV super resolution work?

Re: DALL·E now available in beta

#122

I am thrilled about DALL-E, and the new terms of service. However, how they implemented the improved "diversity" is hilarious. Turns out that they randomly, silently modify your prompt text to append words like "black male" or "female". See https://twitter.com/jd_pressman/status/1549523790060605440 I don't know which emotion I feel more - applause at how glorious this hack is or tears at how ugly it is. Good luck to…

> Turns out that they randomly, silently modify your prompt text to append words like "black male" or "female". I wonder what the distribution of those modifications is?

Today, when DALL-E was still free, my Dad asked me to try a prompt about the Buddha sitting by a river, contemplating. I did about 4 prompt variations, and one of them was an Asian female, if that gives any idea about the frequency (I should note that the depiction was of a young, slim, and attractive female Buddha, so I'm not sure they have the bias thing licked just yet).

Re: DALL·E now available in beta

#123

I am thrilled about DALL-E, and the new terms of service. However, how they implemented the improved "diversity" is hilarious. Turns out that they randomly, silently modify your prompt text to append words like "black male" or "female". See https://twitter.com/jd_pressman/status/1549523790060605440 I don't know which emotion I feel more - applause at how glorious this hack is or tears at how ugly it is. Good luck to…

A dumb solution to a dumber problem.

Re: DALL·E now available in beta

#124
post #92

That's disappointing given up until this point you could have 50 free uses per 24h. I expected it to get monetized eventually, but not so fast and drastically. Well, still had my fun and have to say the creations are so good it's often mind blowing there's an AI behind it.

they're a non-profit so the price is probably still dirt cheap

Not correct. They have a for-profit entity now. That's why there is a huge incentive to monetize. Any for-profit investment gain is capped at 100x, with the rest required to go to their nonprofit. This commercialization is just as I predicted in my substack post 2 days ago that hit the front page of Hacker News: https://aifuture.substack.com/p/the-ai-battle-rages-on

Re: DALL·E now available in beta

#125

> Reducing bias: We implemented a new technique so that DALL·E generates images of people that more accurately reflect the diversity of the world’s population. This technique is applied at the system level when DALL·E is given a prompt about an individual that does not specify race or gender, like “CEO.” Will it do it "more accurately" as they claim? As in, if 90% of CEOs are male, then the odds of a CEO being male i…

If accurately reflects the world population then only one in six pictures will be a white person. Half the pictures will be Asian, another sixth will be Indian.

Slightly more than half of the pictures will be women.

That accurately represents the world's diversity. It won't accurately reflect the world's power balance but that doesn't seem to be their goal.

If you want to say "white male CEO" because you want results that support the existing paradigm it doesn't sound like they'll stop you. I can't imagine a more boring request.

Let's look at interesting questions:

If you ask for "victorian detective" are you going to get a bunch of Asians in deerstalker caps with pipes?

What about Jedi? A lot of the Jedi are blue and almost nobody on Earth is.

Are cartoon characters exempt from the racial algorithm? If I ask for a Smurf surfing on a pizza I don't think that making the Smurf Asian is going to be a comfortable image for any viewer.

What about ageism? 16% of the population is over sixty. Will a request for "superhero lifting a building" have an 16% chance of being old?

If I request a "bad driver peering over a steering wheel" am I still going to get an Asian 50% of the time? Are we ok with that?

I respect the team's effort to create an inclusive and inoffensive tool. I expect it's going to be hard going.

Re: DALL·E now available in beta

#126
post #28

Something I haven’t seen anyone talking about with these huge models: how do future models get trained when more content online is model generated to start with? Presumably you don’t wanna train a model on autogenerated images or text, but you can’t necessarily know which is which.

This precise thing is causing a funny problem in specialty areas. People are using e.g. Google Lens to identify plants, birds and insects, which sometimes returns wrong answers e.g. say it sees a picture of a Summer Tanager and calls it a Cardinal. If the people then post "Saw this Cardinal" and the model picks up that picture/post and incorporates it into its training set, it's just reinforcing the wrong identificat…

That's not really a new problem, though. At one point someone got some bad training data about an old Incan town, the misidentification spread, and nowadays we train new human models to call it Macchu Picchu.

Re: DALL·E now available in beta

#127

Earlier quoted context omitted.

Honestly I would rather that they not try. I don't understand why a computer tool has to be held to a political standard.

It's not a political standard though. There is actual diversity in this world. Why wouldn't you want that in your product?

[deleted]

Re: DALL·E now available in beta

#128

I'm blown away by this: "Starting today, users get full usage rights to commercialize the images they create with DALL·E, including the right to reprint, sell, and merchandise. This includes images they generated during the research preview." I assumed this was going to be the sticking point for wider usage for a long time. They're now saying that you have full rights to sell Dall-E 2 creations?

I think they are reacting to competition. MidJourney is amazing, was easier to get into, gives you commercial rights, and frankly I found more fun to use and even better output in most instances.

MidJourney definitely struggles more with complex prompts from what I saw. If you like the output more, that’s subjective, but I think DALL•E is the leader in the space by a wide margin.

Re: DALL·E now available in beta

#129
post #28

Something I haven’t seen anyone talking about with these huge models: how do future models get trained when more content online is model generated to start with? Presumably you don’t wanna train a model on autogenerated images or text, but you can’t necessarily know which is which.

One interesting comment about this is that some models actually benefit from being fed their own output. Alphafold for instance was fed with its own 'high likelihood' outputs (as demis hassabis described in his lex friedman interview).

Re: DALL·E now available in beta

#130
post #84

Earlier quoted context omitted.

it's a hard problem. at least they tried.

Honestly I would rather that they not try. I don't understand why a computer tool has to be held to a political standard.

There are legitimate reasons to reduce externalizations of societies innate biases.

A mortgage AI that calculates premiums for the public shouldn't bias against people with historically black names, for example.

This problem is harder to tackle because it is difficult to expose and resign the "latent space" that results in these biases; it's difficult to massage the ML algo's to identify and remove the pathways that result in this bias.

It's simply much easier to allow the robot to be bias/racist/reflective of "reality" (its training data), and add a filter / band-aid on top; which is what they've attempted.

when this is appropriate is the more cultured question; I don't think we should attempt to band-aid these models, but for more socially-critical things, it is definitely appropriate.

It's naive on either extreme: do we reject reality, and substitute or own? Or do we call our substitute reality, and hope the zeitgeist follows?

Post reply on HN