Live data from Hacker News

Stable Diffusion XL 1.0

techcrunch.com

21–30 of 182 posts

Re: Stable Diffusion XL 1.0

#21

This explosion of AI-generated imagery will result in an explosion of millions of fake images, obivously. Perhaps in the short-term this is fun, but in the long-term, we will lose a bit more scarcity, which is not that great in my opinion. Isn't the best part of a meal eating after you've not had anything to eat for a while? The best part about a kiss that you've quenched the pain of missing your partner? The best pa…

I want to appreciate your comment, but I can't.

Can you please chisel it on stone tablets for me?

That will really help me appreciate it.

Re: Stable Diffusion XL 1.0

#22
post #19

I hope someday there’s a version of this or something comparable to it that can run on <8gb consumer hardware. The main selling point of Stable Diffusion was its ability to run in that environment.

SDXL 0.9 runs on iPad Pro 8GiB just fine.

Is this using Draw Things, or another app? Did you have to quantize the model first?

Re: Stable Diffusion XL 1.0

#23

I am completely uninformed in this space. Would someone be kind to explain what the current state of the art in image generation is (how does this compare to Midjourney and others)? How do open source models stack up? Also what are the most common use cases for image generation?

[deleted]

Re: Stable Diffusion XL 1.0

#24
post #19

Earlier quoted context omitted.

SDXL 0.9 runs on iPad Pro 8GiB just fine.

Is this using Draw Things, or another app? Did you have to quantize the model first?

Yeah, Draw Things. It will be submitted as soon as SDXL v1.0 weights available. Quantized model should run on iPhones (4GiB / 6GiB models), but we haven't done that yet. So no, these are just typical FP16 weights on iPad.

Re: Stable Diffusion XL 1.0

#25
post #24

Earlier quoted context omitted.

Is this using Draw Things, or another app? Did you have to quantize the model first?

Yeah, Draw Things. It will be submitted as soon as SDXL v1.0 weights available. Quantized model should run on iPhones (4GiB / 6GiB models), but we haven't done that yet. So no, these are just typical FP16 weights on iPad.

Thanks! I guess I'll stick to running it on my Macbook for the time being until the quantized model gets uploaded. What kind of performance are you seeing with the FP16 weights on the iPad? I've run a few SD2.0-based (unquantized) models on my 2020 iPad Pro but it seems like it gets thermally throttled after a while.

Re: Stable Diffusion XL 1.0

#26

I am completely uninformed in this space. Would someone be kind to explain what the current state of the art in image generation is (how does this compare to Midjourney and others)? How do open source models stack up? Also what are the most common use cases for image generation?

SDXL is in roughly the same ballpark as MJ 5 quality-wise, but the main value is in the array of tooling immediately available for it, and the license. You can fine-tune it on your own pictures, use higher order input (not just text), and daisy-chain various non-imagegen models and algorithms (object/feature segmentation, depth detection, processing, subject control etc) to produce complex images, either procedural or one-off. It's all experimental and very improvised, but is starting to look like a very technical CGI field separate from the classic 3D CGI.

Re: Stable Diffusion XL 1.0

#27
post #24

Earlier quoted context omitted.

Yeah, Draw Things. It will be submitted as soon as SDXL v1.0 weights available. Quantized model should run on iPhones (4GiB / 6GiB models), but we haven't done that yet. So no, these are just typical FP16 weights on iPad.

Thanks! I guess I'll stick to running it on my Macbook for the time being until the quantized model gets uploaded. What kind of performance are you seeing with the FP16 weights on the iPad? I've run a few SD2.0-based (unquantized) models on my 2020 iPad Pro but it seems like it gets thermally throttled after a while.

Will be more info upon release. SDXL v0.9 performs generally the same as SD v1 / v2 on the same resolution. But because you tend to run it at larger resolution, you might feel it slower.

Re: Stable Diffusion XL 1.0

#29
post #18

I am completely uninformed in this space. Would someone be kind to explain what the current state of the art in image generation is (how does this compare to Midjourney and others)? How do open source models stack up? Also what are the most common use cases for image generation?

SDXL 0.9 should be the state-of-the-art image generation model (in the open). It generates at 1024x1024 large resolution, with high coherency and good selection of styles out of box. It also has reasonable text-understanding comparing to other models. That has been said, based on the configurations of these models, we are far from saturating what the best model can do. The problem is, FID is terrible metrics to evalu…

Why do you think FID is a terrible metrics? What don't you like in particular about it?
Post reply on HN