DeepFloyd IF: open-source text-to-image model
1–10 of 237 posts
Re: DeepFloyd IF: open-source text-to-image model
#2Re: DeepFloyd IF: open-source text-to-image model
#3Re: DeepFloyd IF: open-source text-to-image model
#4Re: DeepFloyd IF: open-source text-to-image model
#51. A stained glass picture of a woman in a library with a raven on her shoulder with a key in its mouth
2. An oil painting of a man in a factory looking at a cat wearing a top hat
3. A digital art picture of a child riding a llama with a bell on its tail through a desert
4. A 3D render of an astronaut in space holding a fox wearing lipstick
5. Pixel art of a farmer in a cathedral holding a red basketball
Re: DeepFloyd IF: open-source text-to-image model
#616GB VRAM minimum is a bit steep. Sadly excludes my 3080 which is annoying because I'd like something better than Stable Diffusion locally.
Re: DeepFloyd IF: open-source text-to-image model
#7Re: DeepFloyd IF: open-source text-to-image model
#8I'm also very happy for the release of the two upscaler, I can use them to upscale to result of my small 64x64 DDIM models (maybe with some finetuning).
Re: DeepFloyd IF: open-source text-to-image model
#916GB VRAM minimum is a bit steep. Sadly excludes my 3080 which is annoying because I'd like something better than Stable Diffusion locally.
There are multiple ways to speed up the inference time and lower the memory consumption even more with diffusers. To do so, please have a look at the Diffusers docs:
Optimizing for inference time [1]
Optimizing for low memory during inference [2]
[1] https://huggingface.co/docs/diffusers/api/pipelines/if#optim...[2] https://huggingface.co/docs/diffusers/api/pipelines/if#optim...
Re: DeepFloyd IF: open-source text-to-image model
#10So this one can create perfect text in images? If true, that’s insane