Live data from Hacker News

Using Stable Diffusion's img2img on some old Sierra titles

sciprogramming.com

51–60 of 70 posts

Re: Using Stable Diffusion's img2img on some old Sierra titles

#52

Earlier quoted context omitted.

Compared to low res Sierra graphics, what's been shown here is nothing short of astounding, if you ask me. Look at the skull! It's a masterpiece.

> Look at the skull! It's a masterpiece. There was also a time when I (unironically) classified McDonald's food a delicacy (at some point when I was younger than ten).

I didnt ever classify that hot garbage as a delicacy.

Re: Using Stable Diffusion's img2img on some old Sierra titles

#54
post #53

Here's more, characters from old DOS games: https://old.reddit.com/r/StableDiffusion/comments/x2wwxx/usi... and https://old.reddit.com/r/StableDiffusion/comments/x5qrje/usi... I love what the AI did with the hot tub girl from Leisure Suit Larry!

I'm mostly impressed by the guy from Dune 2, he looks like a real person!

Re: Using Stable Diffusion's img2img on some old Sierra titles

#55
This is cool, but img2img likely couldn't be easily used to make images in such games, because every scene would look very different. The problem is that you can't maintain the exact look between images. It could be used for artistic ideas and raw material though.

Re: Using Stable Diffusion's img2img on some old Sierra titles

#57
post #46
post #9

Earlier quoted context omitted.

Weird until I read your comment I was blown away. Then I had another proper look at the first image and in many ways I had to turn off my brains amazing ‘upscaling’ ability. My brain had upscaled that human like blob to a woman spinning around with a sword so her hair covered her face. Looking closely. None of that is there really, just a suggestion of it. And that is enough. The more I learn about vision and sight t…

Artists have for the longest time used our brains ability to upscale. Many paintings, even ones that seem super detailed like those by James Gurney in his Dinotopia series, will have blobs in the background. Our brain will recognize based on silhouette and shape an extraordinary amount of detail that isn’t actually there. Detail such as the type of clothing and the action of a person. But if you look closer it’s a re…

> I’ve been telling anyone who will listen that AI art isn’t stealing much lunch when it comes to professional art.

Yet. These models have been out for only a matter of months. Just last year the state of the art was DALL-e v1, which is a toy in comparison[0] to DALL-e 2/imagen/SD.

Making predictions is perilous but it would be surprising to me if computers did not have fully super-human artistic ability in the next 5 years.

[0] https://openai.com/blog/dall-e/

Re: Using Stable Diffusion's img2img on some old Sierra titles

#58
post #7

Going to have to be the naysayer here. First, I'll say the simple fact that Stable Diffusion produces anything coherent is incredible. I'm blown away by the tech. However, my honest opinion of the results showcased in this article is not positive. Many of the result images contain bizarre distortions or dreamlike artifacts that severely disrupt the flow of the image. Especially the first one. It's clear that there's…

This is the opening act—of course the tech has all sorts of issues.

What I’ve found more amazing than the tech is how rapidly and intensively a community has formed around Stable Diffusion and all the stuff they’re doing with it. I’m more confident than not these issues will get worked out.

This whole thing has been a breath of fresh air. We can run this stuff on high-end workstations and we’re not beholden to tech giants for interesting creative ML applications.

We’re about to see a thousand flowers bloom.

Re: Using Stable Diffusion's img2img on some old Sierra titles

#59
post #2

Is there any service offering user-friendly access to open source DL models? Paid is fine.

If you have over 4MB VRAM you can run it locally. I've been experimenting recently and find that even with 10MB VRAM I can only get 256x256 resolution images. I have a Dockerfile I can share that packages up the install process and removes censorship if anyone is interested. I find the censoring is extremely conservative.

Huh, I've been using the docker container by cmd2 and been doing 512x512 just fine with 8GB. Are you on Windows by any chance?

Re: Using Stable Diffusion's img2img on some old Sierra titles

#60
I'd be interested to know the parameters used, especially prompt_strength

The correspondence to the original image is not especially high: the Leisure Suit Larry image, for example, enhances the original colours of the sea in a nicely realistic way, but all the foreground detail is essentially reinvented from scratch, including some very obvious omissions. In some of them the changes to perspective and more lifelike skull/canyons etc might improve on the original image, but it also flips even pretty basic stuff like which shoulder the woman's hand is placed on (and yes, once you look at that hand closely, the fingers SD has had to add in are all wrong...)

Ideally for this sort of use case you'd want high fidelity to the geometry of the original image but less fidelity to the palette (use more than 256 colours and naturalistic or artistic textures rather than lines and pixel dithering), but I'm not sure SD can manage that at the moment

Post reply on HN