One thing that no one predicted in AI development was how good it would become at some completely unexpected tasks while being not so great at the ones we supposed/hoped it would be good. AI was expected to grow like a child. Somehow blurting out things that would show some increasing understanding on a deep level but poor syntax. In fact we get the exact opposite. AI is creating texts that are syntaxically correct a…
Imagen, a text-to-image diffusion model
511–520 of 661 posts
Re: Imagen, a text-to-image diffusion model
#512One thing that no one predicted in AI development was how good it would become at some completely unexpected tasks while being not so great at the ones we supposed/hoped it would be good. AI was expected to grow like a child. Somehow blurting out things that would show some increasing understanding on a deep level but poor syntax. In fact we get the exact opposite. AI is creating texts that are syntaxically correct a…
Syntactically* I know it's the most trivial of things, but in case you were curious as I often am!
Re: Imagen, a text-to-image diffusion model
#513Re: Imagen, a text-to-image diffusion model
#514Interesting and cool technology - but I can't seem to ignore that every high-quality AI art application is always closed, and I don't seem to buy the ethics excuse for that. The same was said for GPT, yet I see nothing but creativity coming out from its users nowadays.
check out open source alternative dalle-mini: https://huggingface.co/spaces/dalle-mini/dalle-mini
Re: Imagen, a text-to-image diffusion model
#515I have to wonder how much releasing these models will "poison the well" and fill the internet with AI generated images that make training an improved model difficult. After all if every 9/10 "oil painted" image online starts being from these generative models it'll become increasingly difficult to scrape the web and to learn from real world data in a variety of domains. Essentially once these things are widely availa…
Re: Imagen, a text-to-image diffusion model
#516I have to wonder how much releasing these models will "poison the well" and fill the internet with AI generated images that make training an improved model difficult. After all if every 9/10 "oil painted" image online starts being from these generative models it'll become increasingly difficult to scrape the web and to learn from real world data in a variety of domains. Essentially once these things are widely availa…
People training newer models just have to look for the "Imagen" tag or the Dall-E2 rainbow at the corner and heuristically exclude images having these. This is trivial. Unless you assume there are bad actors who will crop out the tags. Not many people now have access to Dall-E2 or will have access to Imagen. As someone working in Vision, I am also thinking about whether to include such images deliberately. Using imag…
Re: Imagen, a text-to-image diffusion model
#517Earlier quoted context omitted.
Still has the issue with screwing up mechanical objects. In their demo checkout the wheels on the skateboards, all over the place.
For comparison, most humans can't draw a bicycle: https://www.wired.com/2016/04/can-draw-bikes-memory-definite...
Re: Imagen, a text-to-image diffusion model
#518Would be fascinated to see the DALL-E output for the same prompts as the ones used in this paper. If you've got DALL-E access and can try a few, please put links as replies!
Posting a few comparisons here. https://twitter.com/joeyliaw/status/1528856081476116480?s=21...
Re: Imagen, a text-to-image diffusion model
#519Earlier quoted context omitted.
Posting a few comparisons here. https://twitter.com/joeyliaw/status/1528856081476116480?s=21...
Imagen seems more realistic where Dall-E2 is more feel-good . That is what I feel personally.
Re: Imagen, a text-to-image diffusion model
#520Earlier quoted context omitted.
Posting a few comparisons here. https://twitter.com/joeyliaw/status/1528856081476116480?s=21...
Looking at these… I can’t help but wonder if these are literal examples of AI imagination?