Live data from Hacker News

Spent $15 in DALL·E 2 credits creating this AI image

pub.towardsai.net

41–50 of 153 posts

Re: Spent $15 in DALL·E 2 credits creating this AI image

#41

DALL-E is truly magic. It got me believing we are close to AGI. I wonder what Gary Marcus or Filip Pieknewski think about it. Surely they must be eating crow.

Yesterday I saw one of Gandalf eating samples at Costco. I was laughing hysterically for a minute. AI is not supposed to have a sense of humor. That was supposed to be the last province of the human, but it is quite awhile since a human made me laugh like that.

What was the prompt for that image?

What wrote the prompt?

Re: Spent $15 in DALL·E 2 credits creating this AI image

#42
post #38

-- spent a day with DALL-E - here are some of my favorites: https://imgur.com/a/uD5yjV3 --

You like your lobsters

-- they're the little lobsters we have over here (アカザエ)! - quite expensive - very good =) - https://en.wikipedia.org/wiki/Metanephrops_japonicus --

Re: Spent $15 in DALL·E 2 credits creating this AI image

#43

A lot of these posts showing up on HN. I wonder - is it because it is so new, or is it because the ways in which we are to use this technology are so nascent that we are discovering how to use it more precisely daily?

I believe it’s for a few reasons. First, it is jaw dropping incredible for most people in tech who have at least a hint of how most ML works. Second, the AI image generation field is racing ahead, in academics and new trained models, so there’s lots of new news. Thirdly some really great models like Dall-e have been opened for wider access and lots of everyday users are discovering its capabilities and doing blog write-up’s which are not news, but are surely interesting to most.

Re: Spent $15 in DALL·E 2 credits creating this AI image

#44

DALL-E is truly magic. It got me believing we are close to AGI. I wonder what Gary Marcus or Filip Pieknewski think about it. Surely they must be eating crow.

Machine learning just glues together existing things, which is how art is created. As amusing these pictures are, it's us humans who bring meaning to them, both when producing what these algorithms use as input and when consuming their output. We are the actual magic behind DALL-E.

An AGI wouldn't need us to this extent, or at all. An AGI would also be able to come up with new ways to represent ideas, even ways that are foreign to us.

Re: Spent $15 in DALL·E 2 credits creating this AI image

#45
post #41

Earlier quoted context omitted.

Yesterday I saw one of Gandalf eating samples at Costco. I was laughing hysterically for a minute. AI is not supposed to have a sense of humor. That was supposed to be the last province of the human, but it is quite awhile since a human made me laugh like that.

What was the prompt for that image? What wrote the prompt?

But the prompt was not funny, only the image.

Re: Spent $15 in DALL·E 2 credits creating this AI image

#46

I wonder how this would play out with the new Stable Diffusion

I've tried out a couple of prompts from the post in Stable Diffusion and as expected the results were much weaker. It has drawn some alpacas and basketballs with little relation between the objects.

I've been playing with Stable Diffusion a lot, and in my experience its results are much weaker then what's shown in this post. The artistic pictures that it generates are beautiful, often more beautiful then Dalle-2 ones. But it has a real problem understanding the basic concepts of anything that is not the simplest task like "draw a character in this or that style". And explaining the situations in detail doesn't help - the AI just stumbles upon basic requests.

Seems like Stable Diffusion has a much more shallow understanding of what it draws and can only produce good result for things very similar to the images it learned from. For example, it could generate really good dutch still life paintings for me - with fruits, bottles and all the regular expected objects for this genre of painting. But when I've asked it to add some unusual objects to the painting (like a Nintendo switch, or a laptop) - it couldn't grasp this concept and just added more warbled fruit. Even though the system definitely knows how a Switch looks like.

The results in the post are much more impressive. I doubt that Dalle-2 saw a lot of similar images in training, but in all of the styles and examples it definitely understood how a llama would interact with a basketball, what are their relative sizes and stuff like that. On surface results from different engines might look similar, but to me this is an enormous difference in quality and sophistication.

Re: Spent $15 in DALL·E 2 credits creating this AI image

#49
I tried a number of these generators a week ago (or so), all with the same prompt: "A child looking longingly at a lollipop on the top shelf" with pretty abysmal (and sometimes horrifying) results. I'm not sure if my expectations are too high, but maybe I was doing it wrong?

Re: Spent $15 in DALL·E 2 credits creating this AI image

#50
>the ball is positioned in such a way that the llama has no real hope of making the shot

I love that we're at the level where the physical "realism" of correctly representing quadrupedals playing basketball is a thing now. I suppose the next level AI will be expected to model a full 3d environment with physical assumptions based on the prompt and then run the simulation

Post reply on HN