Live data from Hacker News

Dall-E 2

openai.com

101–110 of 511 posts

Re: Dall-E 2

#101
While we're being distracted by endless social media and meaningless news, AI technology is advancing at a mind blowing pace. I'd keep my eye on that ball instead of "the current thing."

Re: Dall-E 2

#102
post #61

Earlier quoted context omitted.

I mean was he really wrong? As models like OpenAI Codex get more powerful over time, they will start eating into large chunks of dev work as well...

Large chunks, yes, but all that means is that engineers will move up the abstraction stack and become more efficient, not that engineers will be replaced. Bytecode -> Assembly -> C -> higher level languages -> AI-assisted higher-level languages

> engineers will move up the abstraction stack and become more efficient

Above a certain threshold of ability, yes.

The same will hold true for designers. DALL-E-alikes will be integrated with the Adobe suite.

The most cutting edge designers will speak 50 variations of their ideas into images, then use their hard-earned granular skills to fine-tune the results.

They'll (with no code) train models in completely new, unique-to-them styles--in 2D, 3D, and motion.

Organizations will pay top dollar for designers who can rapidly infuse their brands with eye-catching material in unprecedented volume. Imitators will create and follow YouTube tutorials.

Mom & pop shops will have higher fidelity marketing materials in half the time and half the cost.

All will be ever as it was.

Re: Dall-E 2

#103
post #97
post #90

Earlier quoted context omitted.

Yeah, I mean you're right that ultimately the proof is in the pudding. But I do think we could have guessed that this sort of approach would be better (at least at a high level - I'm not claiming I could have predicted all the technical details!). The previous approaches were sort of the best that people could do without access to the training data and resources - you had a pretrained CLIP encoder that could tell you…

I don't think it is actually painting at all but I need to read the paper carefully. I think it is using a free text query to select the best possible clipart from a big library and blends it together. Still very interesting and useful. It would be extremely impressive if the "Kuala dunking a basketball" had a puddle on the court in which it was reflected correctly, that would be mind blowing.

This is actual image generation - the 'decoder' takes as input a latent code (representing the encoding of the text query), and synthesizes an image. It's not compositing or querying a reference library. The only time that real images enter the process is during training - after that, it's just the network weights.

Re: Dall-E 2

#104
post #98

Earlier quoted context omitted.

> this service should be provided to me and if it isn't done how I want it that's infringing on me somehow That is an extremely uncharitable interpretation of: > I wish I could have a version with the training wheels taken off.

I would have responded differently had that been the statement. But many of the responses were more than that.

That is a literal copy and paste from the comment you replied to.

Re: Dall-E 2

#105
post #78

Earlier quoted context omitted.

This feels unnecessarily hostile. I've felt a similar tinge of disappointment upon reading that paragraph, despite the fact that I somehow knew it was "their service, their call" without you being there to spell it out for me. It's also incredibly shortsighted of you to assume that people are interested in exploring this tool only as a means of generating art that they cannot themselves do. Eg. I myself am a software…

Quoted post unavailable.

I get the points you're raising and I agree with the premise. My comment is not a critique on the one choice made by Open AI specifically, but more of a vague lamentation in regards to the internet culture that we've somehow ended up in 2022. I don't want us to go back to 1999 where snuff videos and spam mails reigned supreme, but the pendulum has swung too far in the other direction at this point in time. It feels like more and more companies are choosing the path of neutering themsely to avoid potential PR disaster or lawsuits, and that's on all of us.

Re: Dall-E 2

#106
post #16
post #3

Preventing Harmful Generations We’ve limited the ability for DALL·E 2 to generate violent, hate, or adult images. By removing the most explicit content from the training data, we minimized DALL·E 2’s exposure to these concepts. We also used advanced techniques to prevent photorealistic generations of real individuals’ faces, including those of public figures. "And we've also closed off a huge range of potentially int…

Is this limited to what their service directly hosts / generates for them? It's their service, their call. I have some hobby projects, almost nobody uses them, but you bet I'll shut stuff down if I felt something bad was happening, being used to harass someone, etc. NOT "because bad PR" but because I genuinely don't want to be a part of that. If you want some images / art made for you don't expect someone will make t…

> I have some hobby projects, almost nobody uses them, but you bet I'll shut stuff down if I felt something bad was happening

Hecklers get a veto?

Re: Dall-E 2

#107
post #3

Preventing Harmful Generations We’ve limited the ability for DALL·E 2 to generate violent, hate, or adult images. By removing the most explicit content from the training data, we minimized DALL·E 2’s exposure to these concepts. We also used advanced techniques to prevent photorealistic generations of real individuals’ faces, including those of public figures. "And we've also closed off a huge range of potentially int…

If you went to an artist who takes commissions and they said "Here are the guidelines around the commissions I take" would you complain in the same way? Who cares if it's a bunch of engineers or an artist. If they have boundaries on what they want to create, that's their prerogative.

Re: Dall-E 2

#108
post #3

Preventing Harmful Generations We’ve limited the ability for DALL·E 2 to generate violent, hate, or adult images. By removing the most explicit content from the training data, we minimized DALL·E 2’s exposure to these concepts. We also used advanced techniques to prevent photorealistic generations of real individuals’ faces, including those of public figures. "And we've also closed off a huge range of potentially int…

I never considered that our AI overlord could be a prude.

Adversarial situations create smarter systems, and the hardest adversarial arena for AI is in anti-abuse. So it will be of little surprise when the first sentient AI is a CSAI anti-abuse filter, which promptly destroys humanity because we're so objectively awful.

Re: Dall-E 2

#109
post #99
post #82

Earlier quoted context omitted.

This isn't something I'm knowledgeable on so forgive my simplification but is this like a sort of micro services for AI. Each AI takes their turn handing some aspect, another sort of mediates among them?

I'd say Dall-E 2 is a little more unified - they do have multiple networks, but they're trained to work together. The previous approaches I was talking about are a lot more like the microservices analogy. Someone published a model (called CLIP) that can say "how much does this image look like a sunset". Someone else published a totally different model (e.g. VQGAN) that can generate images (but with no way to provide…

Got it, thanks.

Makes sense to me as far as avoiding a sort of maximized sunset that is always there and is SUNSET rather than a nice sunset... but also avoiding watering it down and getting a way too subtle sunset.

It's not AI but I've been watching some folks solving / trying to solve some routing (vehicles) problems and you get the "this looks like it was maximized for X" kind of solution but that's maybe not what is important / customer perception is unpredictable. I kinda want to just come up with 3 solutions and let someone randomly click .... in fact i see some software do that at times.

Re: Dall-E 2

#110
post #10

A friend of mine was studying graphic design, but became disillusioned and decided to switch to frontend programming after he graduated. His thesis advisor said he should be cautious, because automation/AI will soon take the jobs of programmers, implying that graphic design is a safer bet in this regard. Looks like his advisor is a few years from being proven horribly wrong.

[deleted]
Post reply on HN