If you have a GPU with >4GB of VRAM and you want to run this locally, here's a fork of the Stable Diffusion repo with a convenient web UI: https://github.com/hlky/stable-diffusion It supports both txt2img and img2img. (Not affiliated.) Edit: Incidentally, I tried running it on a CPU. It is possible, but it took 3 minutes instead of 10 seconds to produce an image. It also required me to hack up the script in a really…
What kind of GPU do you have? It takes several minutes to produce an image on my 1070.
Try Stable Diffusion's Img2Img Mode
71–80 of 166 posts
Re: Try Stable Diffusion's Img2Img Mode
#72If you have a GPU with >4GB of VRAM and you want to run this locally, here's a fork of the Stable Diffusion repo with a convenient web UI: https://github.com/hlky/stable-diffusion It supports both txt2img and img2img. (Not affiliated.) Edit: Incidentally, I tried running it on a CPU. It is possible, but it took 3 minutes instead of 10 seconds to produce an image. It also required me to hack up the script in a really…
What kind of GPU do you have? It takes several minutes to produce an image on my 1070.
Re: Try Stable Diffusion's Img2Img Mode
#73I've been trying to get some sensible images out of my descriptions, but I fail miserably. In this case I had the prompt "cow chewing bone" with 4 squares representing the two pair of feet, the body and the head. None cared about chewing on a bone. With DALL·E 2 I tried to get an image of a little girl building sandcastles and a monster threatening her: "little scared girl building a sandcastle and a big angry monste…
Re: Try Stable Diffusion's Img2Img Mode
#74If you have a GPU with >4GB of VRAM and you want to run this locally, here's a fork of the Stable Diffusion repo with a convenient web UI: https://github.com/hlky/stable-diffusion It supports both txt2img and img2img. (Not affiliated.) Edit: Incidentally, I tried running it on a CPU. It is possible, but it took 3 minutes instead of 10 seconds to produce an image. It also required me to hack up the script in a really…
What kind of GPU do you have? It takes several minutes to produce an image on my 1070.
Re: Try Stable Diffusion's Img2Img Mode
#75Can't help but see this as a friendly herald for a dystopian future.
There you go. We were promised quite far fetched things like 'flying cars', 'time travel', 'life extension' and 'universal basic income'. Instead we get Deep Learning AI's being used for generating and faking sentences, images, videos, voices, code and digital art all being trained on mountains of data in data centers all significantly contributing to the already burning up of the planet to no benefit and no efficien…
My point being, there's enough good things and bad things to fit any mood and worldview. Everyone can basically pick as they'd like.
[0] something like this https://www.reddit.com/r/Futurology/
Re: Try Stable Diffusion's Img2Img Mode
#76I ended up paying the $10 for Google Colab Pro and that's how I've been using this. Maybe I'll figure out how to get this working on my old 1080 TI to see if it's faster.
Anyway, for the one that I'm using which has a web UI, you can use this Colab link. It's pretty great! https://colab.research.google.com/drive/1KeNq05lji7p-WDS2BL-...
What I really wish was that the img2img tool could be used to take a text2img output and then "refine" it further. As it is, the img2img tool doesn't seem particularly great.
People on Reddit are talking about "I just generate 100 images and pick the best one"... but this is incredibly slow on the P100 GPU that Google has me on. Does this just require a monster GPU like a 3080/3090 in order to get any decent results?
Re: Try Stable Diffusion's Img2Img Mode
#77Earlier quoted context omitted.
You must be kidding it’s hard labor creating images and you are complaining that it’s going to be easier in the future? You are complaining that the skill gap is eradicated and it now only depends on your ideas to produce compelling work? That’s absolutely insane everything that takes labor intensive tasks and makes them basically free is the future. This is not dystopian it’s our future, a future in which you and yo…
The skill gap was what made it valuable. There is not that much value in flooding everyone with generated images that have no story, no person behind it, no effort, no reason to exist, out of place. “Art” is as much about the final image as it is about the person who made it, why she made it and how she got there. Also there is pretty dystopian angle because these tools balantly stole all the work of artists by “lear…
Re: Try Stable Diffusion's Img2Img Mode
#78If you upload an image with 2:1 aspect ratio it squishes it to 1:1 instead of letting you crop it. Seems like a basic thing they could add.
Re: Try Stable Diffusion's Img2Img Mode
#79Earlier quoted context omitted.
What kind of GPU do you have? It takes several minutes to produce an image on my 1070.
also on a 1070, I can generate an image in ~15 seconds, surely you're doing something wrong.
Re: Try Stable Diffusion's Img2Img Mode
#80What are your best prompts so far?