Live data from Hacker News

Dall-E 2

openai.com

51–60 of 511 posts

Re: Dall-E 2

#51
post #21

The timing of the Dall-E 2 launch an hour ago seems to correspond with a recent piece of investigative journalism by Buzzfeed News about one of Sam Altman's other ventures, published 15 hours ago and discussed elsewhere actively on HN right now: https://news.ycombinator.com/item?id=30931614 I point this out because while Dall-E 2 seems interesting (I'm out of my depth, so delegating to the conversation taking place h…

Maybe I’m naive, but I see this as a coincidence. If it was an hour later, then maybe there would be something.

If the article GP refers to was posted 16 hours ago instead of 15, would that really make a difference?

Re: Dall-E 2

#52
post #10

A friend of mine was studying graphic design, but became disillusioned and decided to switch to frontend programming after he graduated. His thesis advisor said he should be cautious, because automation/AI will soon take the jobs of programmers, implying that graphic design is a safer bet in this regard. Looks like his advisor is a few years from being proven horribly wrong.

I mean was he really wrong? As models like OpenAI Codex get more powerful over time, they will start eating into large chunks of dev work as well...

Yes. Translating business requirements, customer context, engineering constraints, etc. into usable, practical, functional code, and then maintaining that code and extending it is so far beyond the horizon, that many other skillsets will replaced before programming is. After all, at that point, the AI itself, if it's so smart, should be able to improve itself indefinitely. In which case we're fucked. Programming will be the last thing to be automated before the singularity.

Unlike artwork, precision and correctness is absolutely critical in coding.

Re: Dall-E 2

#53
post #21

The timing of the Dall-E 2 launch an hour ago seems to correspond with a recent piece of investigative journalism by Buzzfeed News about one of Sam Altman's other ventures, published 15 hours ago and discussed elsewhere actively on HN right now: https://news.ycombinator.com/item?id=30931614 I point this out because while Dall-E 2 seems interesting (I'm out of my depth, so delegating to the conversation taking place h…

Maybe I’m naive, but I see this as a coincidence. If it was an hour later, then maybe there would be something.

Another consideration, then: it was published to HN almost instantly after it was released to the world, 52 minutes after the HN post about Worldcoin was submitted and started showing traction.

I don't see the publication of a marketing page (again, not a finished product) for a product founded by someone who's other main venture is being investigated by journalists for misleading claims as being a coincidence, but if the timing matters and 14-15 hours doesn't seem like it works for the assertion in your mind, then perhaps the Dall-E 2 page going live less than an hour after the Worldcoin HN submission fits the bill.

I've got no horse in this race. I'm just drawing attention to familiar PR strategies used for brand risk mitigation, that's all.

Re: Dall-E 2

#54
post #21

The timing of the Dall-E 2 launch an hour ago seems to correspond with a recent piece of investigative journalism by Buzzfeed News about one of Sam Altman's other ventures, published 15 hours ago and discussed elsewhere actively on HN right now: https://news.ycombinator.com/item?id=30931614 I point this out because while Dall-E 2 seems interesting (I'm out of my depth, so delegating to the conversation taking place h…

Genuine question: how are the two stories even related? It’s certainly not apparent from the BuzzFeed article (or at least a quick skim of it).

Sam Altman is OpenAI's CEO.

What I'm submitting for consideration is that the marketing page and associated press blasts (there's a live influencer reaction video airing right now about Dall-E 2, for instance) for Dall-E 2 were potentially pushed up to offset negative press from Worldcoin for their shared founder.

I'd like to be wrong. But it's too well timed.

Re: Dall-E 2

#55
The most interesting item to me is the variations on the garden shop and bathroom sink idea. The realism of these leaks the AI lacking intuition of the requirements. This makes for a number of nonsensical designs that look right at first like: This Sink lacks sensical faucets. https://cdn.openai.com/dall-e-2/demos/variations/modified/ba...

This doorway is downright impossible https://cdn.openai.com/dall-e-2/demos/variations/modified/fl...

Re: Dall-E 2

#56
What happens when they train this thing to make videos? We're about to be dealing with a flood of AI-generated visual/video content. We already have to deal with text bots everywhere... wow.

Re: Dall-E 2

#57

Something about this makes me nauseous. Perhaps is the fact that soon the market value for creatives is going to fall to a hair about zero for all but the most famous. We will be all the poorer for it when 95% of images you see are AI generated. There will be niches of course but in a few short years it'll be over for a huge swathe of creative professionals who are already struggling. Some of the images also hit me w…

By "creatives" you seem to mean "people who drum up the equivalent of elevator music for ads and blogs". This will not remotely replace any working "creative" people that I know.

Except it will only get more powerful with time, probably at an accelerating pace. Everyone always downplays these legitimate fears about AI, pointing out how "it can't do X". They always forget to put the "yet" at the end of that sentence.

Re: Dall-E 2

#58
post #10

A friend of mine was studying graphic design, but became disillusioned and decided to switch to frontend programming after he graduated. His thesis advisor said he should be cautious, because automation/AI will soon take the jobs of programmers, implying that graphic design is a safer bet in this regard. Looks like his advisor is a few years from being proven horribly wrong.

I mean was he really wrong? As models like OpenAI Codex get more powerful over time, they will start eating into large chunks of dev work as well...

No worry, the one thing humans can do that robots can't (yet) is fill spare time with ever more work https://en.wikipedia.org/wiki/Parkinson's_law

Re: Dall-E 2

#59
I'm only part way through the paper, but what struck me as interesting so far is this:

In other text-to-image algorithms I'm familiar with (the ones you'll typically see passed around as colab notebooks that people post outputs from on Twitter), the basic idea is to encode the text, and then try to make an image that maximally matches that text encoding. But this maximization often leads to artifacts - if you ask for an image of a sunset, you'll often get multiple suns, because that's even more sunset-like. There's a lot of tricks and hacks to regularize the process so that it's not so aggressive, but it's always an uphill battle.

Here, they instead take the text embedding, use a trained model (what they call the 'prior') to predict the corresponding image embedding - this removes the dangerous maximization. Then, another trained model (the 'decoder') produces images from the predicted embedding.

This feels like a much more sensible approach, but one that is only really possible with access to the giant CLIP dataset and computational resources that OpenAI has.

Re: Dall-E 2

#60

Something about this makes me nauseous. Perhaps is the fact that soon the market value for creatives is going to fall to a hair about zero for all but the most famous. We will be all the poorer for it when 95% of images you see are AI generated. There will be niches of course but in a few short years it'll be over for a huge swathe of creative professionals who are already struggling. Some of the images also hit me w…

Not exactly. All the ideas put forth in these demos are really arbitrary, with nothing whatsoever to say. Generating crap art becomes more and more effortless: we've seen this in music as well.

Jumping out of the conceptual box to generate novel PURPOSE is not the domain of a Dall-E 2. You've still gotta ask it for things. It's a paintbrush. Without a coherent story, it's an increasingly impressive stunt (or a form of very sophisticated 'retouching brush').

If you can imagine better than the next guy, Dall-E 2 is your new tool for expression. But what is 'better'?

Post reply on HN