Earlier quoted context omitted.
If you have over 4MB VRAM you can run it locally. I've been experimenting recently and find that even with 10MB VRAM I can only get 256x256 resolution images. I have a Dockerfile I can share that packages up the install process and removes censorship if anyone is interested. I find the censoring is extremely conservative.
Huh, I've been using the docker container by cmd2 and been doing 512x512 just fine with 8GB. Are you on Windows by any chance?
Using Stable Diffusion's img2img on some old Sierra titles
61–70 of 70 posts
Re: Using Stable Diffusion's img2img on some old Sierra titles
#62Earlier quoted context omitted.
If you have over 4MB VRAM you can run it locally. I've been experimenting recently and find that even with 10MB VRAM I can only get 256x256 resolution images. I have a Dockerfile I can share that packages up the install process and removes censorship if anyone is interested. I find the censoring is extremely conservative.
Please do share. I did the same and have it but having trouble deploying it to a prod GPU server, still figuring it out.
You can run it with a command like this (I'm on windows): docker run -it -v :/stable-diffusion/models/ldm/stable-diffusion-v1/model.ckpt -v :/stable-diffusion/outputs -v :/stable-diffusion/inputs -v :/root/.cache --gpus all knightley python /stable-diffusion/scripts/txt2img.py --W 256 --H 256 --prompt "a horse wearing a top hat"
assuming you build the image and tag it "knightley"
Re: Using Stable Diffusion's img2img on some old Sierra titles
#63Going to have to be the naysayer here. First, I'll say the simple fact that Stable Diffusion produces anything coherent is incredible. I'm blown away by the tech. However, my honest opinion of the results showcased in this article is not positive. Many of the result images contain bizarre distortions or dreamlike artifacts that severely disrupt the flow of the image. Especially the first one. It's clear that there's…
Are you naysaying Stable Diffusion or the idea of generative models for art in general? It's hard to look at what's happened in the last few months and not think of it as akin to the invention of the steam engine, but for art. It's not perfect, as early machines had many flaws, were wildly inefficient, produced irregular output. But the innovation that followed created the industrial revolution.
computer generated art *
Re: Using Stable Diffusion's img2img on some old Sierra titles
#64Earlier quoted context omitted.
Weird until I read your comment I was blown away. Then I had another proper look at the first image and in many ways I had to turn off my brains amazing ‘upscaling’ ability. My brain had upscaled that human like blob to a woman spinning around with a sword so her hair covered her face. Looking closely. None of that is there really, just a suggestion of it. And that is enough. The more I learn about vision and sight t…
The key element missing from the generated images is understanding of form. In recognizing objects we're generally relying on shape first(so, strong outlines and silhouettes, blobs of color, and so forth). But only afterwards does our brain start to see forms in perspective. When learning drawing I gradually got a sense of what is really going on is that I'm gaining a more conscious command of different shapes, just…
The current approach creates huge limitation of input/output images being like 512x512 small and a whole load of texture-turning-to-shape and vice versa artifacts.
It could be possibly overcome with a paradigm shift, though.
Re: Using Stable Diffusion's img2img on some old Sierra titles
#65Earlier quoted context omitted.
> Look at the skull! It's a masterpiece. There was also a time when I (unironically) classified McDonald's food a delicacy (at some point when I was younger than ten).
I didnt ever classify that hot garbage as a delicacy.
Never said you did, but there are, ahem, parallels.
Re: Using Stable Diffusion's img2img on some old Sierra titles
#66Earlier quoted context omitted.
Compared to low res Sierra graphics, what's been shown here is nothing short of astounding, if you ask me. Look at the skull! It's a masterpiece.
> Look at the skull! It's a masterpiece. There was also a time when I (unironically) classified McDonald's food a delicacy (at some point when I was younger than ten).
Re: Using Stable Diffusion's img2img on some old Sierra titles
#67Earlier quoted context omitted.
I didnt ever classify that hot garbage as a delicacy.
> I didnt ever classify that hot garbage as a delicacy. Never said you did, but there are, ahem, parallels.
:D
Re: Using Stable Diffusion's img2img on some old Sierra titles
#68Earlier quoted context omitted.
Artists have for the longest time used our brains ability to upscale. Many paintings, even ones that seem super detailed like those by James Gurney in his Dinotopia series, will have blobs in the background. Our brain will recognize based on silhouette and shape an extraordinary amount of detail that isn’t actually there. Detail such as the type of clothing and the action of a person. But if you look closer it’s a re…
> I’ve been telling anyone who will listen that AI art isn’t stealing much lunch when it comes to professional art. Yet . These models have been out for only a matter of months. Just last year the state of the art was DALL-e v1, which is a toy in comparison[0] to DALL-e 2/imagen/SD. Making predictions is perilous but it would be surprising to me if computers did not have fully super-human artistic ability in the next…
Relationships between things is complicated. Someone resting their face on a fence is going to have a huge number of effects on the deformations of the face, especially the eyes and hair depending on how they are resting. It’s not enough for AI to have seen enough pictures of faces on fences to be able to apply that in an image. It needs to understand what pressure and gravity is doing to the underlying structures. That’s how human artists study at least. It’s why they can take that lesson and apply it to learnings about how skin behaves depending on the age of a person. They aren’t copying. They are solving problems by thinking about muscles underneath and how they change depending on any number of factors.
None of this even touches on lighting and colours.
If there’s one prediction of the future I’m willing to make, it’s that until research progresses on teaching computers to apply actual knowledge, AI within the creative space will remain assistive instead of replacing.
Re: Using Stable Diffusion's img2img on some old Sierra titles
#69Earlier quoted context omitted.
Well you’ve gotta either run it locally or pay for it. The GPUs this stuff runs on are too expensive to offer unlimited use for free.
I'm not against paying but it should be somewhat raisonnable. to put things into perspective: Colab pro only costs about $10/month and you will probably be able to generate at the same speed.
Re: Using Stable Diffusion's img2img on some old Sierra titles
#70Earlier quoted context omitted.
Why do people always ask “why?” on these things? Not everything has to be “useful” or have commercial value. The original games aren’t going anywhere. People just enjoy doing this kind of stuff, pushing the technology to make interesting things.
I’m not criticizing it because it’s not useful. I’m criticizing it because it is not art, and it is anti-humanist.