Live data from Hacker News

Spent $15 in DALL·E 2 credits creating this AI image

pub.towardsai.net

141–150 of 153 posts

Re: Spent $15 in DALL·E 2 credits creating this AI image

#141

>the ball is positioned in such a way that the llama has no real hope of making the shot I love that we're at the level where the physical "realism" of correctly representing quadrupedals playing basketball is a thing now. I suppose the next level AI will be expected to model a full 3d environment with physical assumptions based on the prompt and then run the simulation

The goalposts are practically galloping down the field.

Re: Spent $15 in DALL·E 2 credits creating this AI image

#142
post #14

> it was difficult to find images where the entire llama fit within the frame I had the same trouble. In my experiment I wanted to generate a Porco Rosso style seaplane. illustration. Sadly none of the generated pictured had the whole of the airplane in them. The wingtips or the tail always got left off. I found this method to be a reliable workaround: I have downloaded the image I liked the most. Used an image editi…

Very nice result. But the plane doesn't look very seaplane-y to me. Did you also try it with a plain plane?

Re: Spent $15 in DALL·E 2 credits creating this AI image

#143

> DALL·E 2 struggles to generate realistic faces. According to some sources, this may have been a deliberate attempt to avoid generating deepfakes. That might be true, but after experimenting with DALL·E 2 last week (and spending more than $15), I have a different theory. My tests focused on how well it could create art works around three common themes: still life, landscape, and portrait. For the first two categorie…

It's only small faces that are distorted, and they are often heavily distorted, it's not an "uncanny valley effect", they look like disfigured pieces of meat and skin. It's the same in dalle-mini.

Dalle2 can clearly generate super-realistic faces without any problem, if you look at most of the posts at r/dalle2

The issue with small faces might be architectural if there is context-aware upscaling going on in the network, where a face needs to start larger than some smallest scale or it won't survive that process. That in turn might be an issue of too little training. A small face in a photo in the training data won't generate as much error gradient if it goes wrong as a larger face, but as you suggest we as viewers are much more prone to scrutinize faces even though they are small.

Re: Spent $15 in DALL·E 2 credits creating this AI image

#144

I was curious to compare results with Craiyon.ai Here is "llama in a jersey dunking a basketball like Michael Jordan, shot from below, tilted frame, 35°, Dutch angle, extreme long shot, high detail, dramatic backlighting, epic, digital art": https://imgur.com/a/7LoAtRx Here is "Llama in a jersey dunking a basketball like Michael Jordan, screenshots from the Miyazaki anime movie", much worst: https://imgur.com/a/g99G7…

Fascinating, are there any other similar products in this same category as DALL.E and Craiyon?

After using all of the different models extensively, Stable Diffusion is currently state-of-the-art.

Images are more artistic and less clip art-like than Dall-E, but also don’t have a house style like Midjourney. It’s stunningly good - and open source.

What’s really cool is that the devs have worked hard to optimise the model, so after being trained on 1000 A100s it’ll run happily on an 8gb graphics card or M2 Mac.

Re: Spent $15 in DALL·E 2 credits creating this AI image

#145
post #132

Earlier quoted context omitted.

How long does it take the prompt engineer to make a design though?

Not a professional graphic designer, but did some graphic design classes and photography/photo editing classes in college, and still do it as a hobby. Things that at one time took days can now be done in minutes with some skillful use of Dall-E + Photoshop. IMO, any image editing software that incorporates a similar technology will take over the market and it'll be one of the most important features in any graphic de…

I think the more likely case is you'll get artists who sketch out a concept and use AI to generate the image, and then photoshop the rest.

Re: Spent $15 in DALL·E 2 credits creating this AI image

#146

> DALL·E 2 struggles to generate realistic faces. According to some sources, this may have been a deliberate attempt to avoid generating deepfakes. That might be true, but after experimenting with DALL·E 2 last week (and spending more than $15), I have a different theory. My tests focused on how well it could create art works around three common themes: still life, landscape, and portrait. For the first two categorie…

That's probably not the reason. Generating faces was one of the first things GANs were ever used for. They can make near perfect faces because the internet is flooded with images of faces, often high quality celebrity shots.

The reason it can't do faces well are very likely due to the filters being applied to try and stop people making pictures of real people. This is probably also the explanation for the random misses where it paints pictures of something that's not a llama. OpenAI is rewriting queries to make them more "diverse" i.e. acceptable to leftist ideology, and their rewriting logic seems to be completely broken. There have been many reports of people requesting something without even any humans in it at all, and discovering black/asian/arab people cropping up in it. At least earlier versions of the filter involved simply stuffing words onto the end as proven by people requesting "Person holding a sign that says " and getting back signs saying "black female" etc.

Man asks for a cowboy + a cat and gets a portrait of an Asian girl. Gwern comments with an explanation:

https://www.reddit.com/r/dalle2/comments/w7qvgl/comment/ihm6...

"tldr: it's the diversity stuff. Switch "cowboy" to "cowgirl", which would disable the diversity stuff because it's now explicitly asking for a 'girl', and OP's prompt works perfectly."

Big discussion thread where people discuss the problem and (of course) the censorship that tries to hide what's happening:

https://www.reddit.com/r/dalle2/comments/w944fa/there_is_evi...

"I once tried some food photography and received a cheese with a guys face for no reason."

"This has been mentioned on this sub multiple times, but those threads have consistently been removed by the mods - as will this one."

"There was a thread about that prompt and, yes, the person did get diverse [sumo wrestlers]"

"Been doing women images and seeing the article decided to try narrowing the results to "caucasian woman". Still gave me diversity. Whether you want it, or not, you're getting diversity"

Re: Spent $15 in DALL·E 2 credits creating this AI image

#147

I recently made PromptWiki[0] to try to document useful prompts and examples. I think we're at the beginning of exploring what these image models can do and what the best ways to work with them are. [0] https://promptwiki.com

you should check out these amazing art studies by @proximasan, @EErratica, @KyrickYoung, and @sureailabs (twitter) https://proximacentaurib.notion.site/proximacentaurib/parrot...

Re: Spent $15 in DALL·E 2 credits creating this AI image

#148
post #14

> it was difficult to find images where the entire llama fit within the frame I had the same trouble. In my experiment I wanted to generate a Porco Rosso style seaplane. illustration. Sadly none of the generated pictured had the whole of the airplane in them. The wingtips or the tail always got left off. I found this method to be a reliable workaround: I have downloaded the image I liked the most. Used an image editi…

I think "fitting the entire X within the image" is not done on purpose. The results are more aesthetically pleasing when the subject is large, even if a part of it is missing.

Re: Spent $15 in DALL·E 2 credits creating this AI image

#150
post #136
post #64

Earlier quoted context omitted.

I had similar problems trying to get the whole of a police car overgrown with weeds. https://imgur.com/a/U5Hl2gO I was testing to see how close I could get to replicating a t-shirt graphic concept I saw. I had been using ~"A telephoto shot of A neglected police car from the 1980s Viewed from a 3/4 angle sits in the distance. The entire vehicle is visible but it is overgrown with grass and flowery vines" This process…

Original: https://foreveryonecollective.com/products/abolition-is-crea...

That’s right!
Post reply on HN