Live data from Hacker News

Imagen, a text-to-image diffusion model

gweb-research-imagen.appspot.com

501–510 of 661 posts

Re: Imagen, a text-to-image diffusion model

#501
post #496
post #491

Seeing the artificial restrictions to this model as well as to DALL-E 2, I can't help but ask myself why the porn industry isn't driving its own research. Given the size of that industry and the sheer abundance of training material, it seems just a matter of time until you can create photo realistic images of yourself with your favourite celebrity for a small fee. Is there anything I am missing? Can you only do this…

Transfer learning is a thing. But I have not tried making generative models with out-of-distribution data before. Distributions other than main training data. There are several indie attempts that I am aware of. Mentioning them to the reply of this comment. (In case the comment gets deleted) The first layers should be general. But the later layers should not behave well to porn images. As they are more specialist lay…

1. DeepCreamPy draws over hentai sensor bars if you direct it where the bar is: https://github.com/gguilt/DeepCreamPy

2. hentAI automates the process: https://github.com/natethegreate/hent-AI

3. [NSFW] Should look at this person on Twitter: https://twitter.com/nate_of_hent_ai

4. [NSFW] PornHub released vintage porn videos upscaled to 4k with AI a while back. The called it the "Remastured Project": https://www.pornhub.com/art/remastured

5. [NSFW] This project shows the limit of AI-wthout-big-tech-or-corporate-support projects. This project creates female genitalia that don't exist in the real world. Project is "This Vagina Does Not Exist": https://thisvaginadoesnotexist.com/about.html

Re: Imagen, a text-to-image diffusion model

#502
post #494
post #491

Seeing the artificial restrictions to this model as well as to DALL-E 2, I can't help but ask myself why the porn industry isn't driving its own research. Given the size of that industry and the sheer abundance of training material, it seems just a matter of time until you can create photo realistic images of yourself with your favourite celebrity for a small fee. Is there anything I am missing? Can you only do this…

Porn is actually a really good litmus test to see if a money/media transfer technology has real promise. Pornography needs exactly 2 things to work well - a way to deliver media, and a way to collect money. If you truly have a system that can do one of those two things better than we currently can, and it's not just empty hype, it will be used for porn. "Empty hype" won't touch that stuff, but real-world usecases wil…

Wow. The last paragraph of my comment looked nearly identical to yours, but I deleted it before submitting because I didn't want to derail. Exactly my thoughts...

Re: Imagen, a text-to-image diffusion model

#503

Would be fascinated to see the DALL-E output for the same prompts as the ones used in this paper. If you've got DALL-E access and can try a few, please put links as replies!

Posting a few comparisons here. https://twitter.com/joeyliaw/status/1528856081476116480?s=21...

Imagen seems more realistic where Dall-E2 is more feel-good.

That is what I feel personally.

Re: Imagen, a text-to-image diffusion model

#505

Earlier quoted context omitted.

The AI ethics thing is just a PR larp at this point. “Oh our tech is so dangerous and amazing it could turn the world upside down” yet we hand it to random Bluechecks on Twitter. It’s just marketing

You know Twitter and Google are different companies, right?

[deleted]

Re: Imagen, a text-to-image diffusion model

#506
post #437

I have to wonder how much releasing these models will "poison the well" and fill the internet with AI generated images that make training an improved model difficult. After all if every 9/10 "oil painted" image online starts being from these generative models it'll become increasingly difficult to scrape the web and to learn from real world data in a variety of domains. Essentially once these things are widely availa…

Look at carpentry blogs, recipe blogs. Nearly all of it is junk content. I bet if you combined GPT and imagen or dalle2 you could replace all of them. Just provide a betty crocker recipe and let it generate a blog that has weekly updates and even a bunch of images - "happy family enjoying pancakes together" I can see the future as being devoid of any humanity.

Doesn't it increases the value of genuine human-produced content? Or their NFTs!

Re: Imagen, a text-to-image diffusion model

#507
post #437

I have to wonder how much releasing these models will "poison the well" and fill the internet with AI generated images that make training an improved model difficult. After all if every 9/10 "oil painted" image online starts being from these generative models it'll become increasingly difficult to scrape the web and to learn from real world data in a variety of domains. Essentially once these things are widely availa…

I also worry about the potential to further stifle human creativity, e.g. why paint that oil painting of a panda riding a bicycle when I could generate one in seconds?

Our imaginations are gigantic. We'll find something else impressive and engaging to do. Or not care. I'm not worried. Watch children: they find a way to play even when there is nothing.

Re: Imagen, a text-to-image diffusion model

#508
post #328

Earlier quoted context omitted.

They're expensive to train, but not awfully expensive to use. Especially if you have hundreds of images you want to generate (due to the way compute devices tend to get much more efficiency with a large batch size). Google could totally afford it, especially if the feature was hidden behind a button the user had to click, and not just run for every image search.

The input control is pretty hard - it kinda needs an AGI :). How do you stop undesirable images being created?

If I were running Google, I would release it with a disclaimer, and not do anything technical to prevent undesirable images being created.

How does Adobe prevent Photoshop being used to draw offensive images? They don't... People understand that a tool can be used for good and bad.

Re: Imagen, a text-to-image diffusion model

#509
post #473

One thing that no one predicted in AI development was how good it would become at some completely unexpected tasks while being not so great at the ones we supposed/hoped it would be good. AI was expected to grow like a child. Somehow blurting out things that would show some increasing understanding on a deep level but poor syntax. In fact we get the exact opposite. AI is creating texts that are syntaxically correct a…

I doubt 99% of humans can draw a ”chess game with a puzzle where white mates in 4 moves”

They can with computer assistance, and this AI sort of has that in that it’s both some “intelligence” and a whole lot of memorized internet, with the issue that we don’t know how to separate those things.

Re: Imagen, a text-to-image diffusion model

#510
post #390

Earlier quoted context omitted.

Running inference on one of these models takes like a GPU minute, so they can't just let the public use them.

Can't be this; Google Colab gives out tons of free GPU usage.

Google has a lot of GPUs, but even so Colab seems like it’s a lot cheaper than it should be. You can get some very good GPUs on the paid plan.
Post reply on HN