Live data from Hacker News

Stable Diffusion XL 1.0

techcrunch.com

71–80 of 182 posts

Re: Stable Diffusion XL 1.0

#71

It’ll be "released" once the model weights show up on the repo or in HuggingFace… for now it’s "announced" It should appear here at some point, currently only the VAE was added: https://huggingface.co/stabilityai

Isn't this it? https://huggingface.co/stabilityai/stable-diffusion-xl-base-...

Re: Stable Diffusion XL 1.0

#72
post #71

It’ll be "released" once the model weights show up on the repo or in HuggingFace… for now it’s "announced" It should appear here at some point, currently only the VAE was added: https://huggingface.co/stabilityai

Isn't this it? https://huggingface.co/stabilityai/stable-diffusion-xl-base-...

Yes, it's now been released

Re: Stable Diffusion XL 1.0

#73

Is this pre-censored like their other later models?

No. From what I’ve gathered was trained on human anatomy, but not straight up porn. What they tried for 2.0/2.1 was way too overdone, to the point where if I prompted “princess Zelda,” the generation would only look mildly like her. Presumably they just didn’t have many images of people in the training. 1.5 and SDXL both work fine of that front.

Fine tuners will quickly take it further, if that’s what you’re after.

Re: Stable Diffusion XL 1.0

#74

It’ll be "released" once the model weights show up on the repo or in HuggingFace… for now it’s "announced" It should appear here at some point, currently only the VAE was added: https://huggingface.co/stabilityai

[deleted]

Re: Stable Diffusion XL 1.0

#75
post #40

I'm out of date on the image-generating side of AI, but I'd like to check things out. What's the best tool for image generation that's available on a website right now? Ie, not a model that I have to run locally.

I've found https://firefly.adobe.com/ pretty good at composing images with multiple subjects. [disclaimer - I work at Adobe, but not in the Creative Cloud] But I wouldn't say it's the "best." Just trained on images that weren't taken from unconsenting artists.

I'm actually a big fan of firefly. It has a different kind of style from the others, presumably due to its training dataset?

Re: Stable Diffusion XL 1.0

#76
post #40

I'm out of date on the image-generating side of AI, but I'd like to check things out. What's the best tool for image generation that's available on a website right now? Ie, not a model that I have to run locally.

If you want to play around with Stable Diffusion XL: https://clipdrop.co

Re: Stable Diffusion XL 1.0

#77

I hope someday there’s a version of this or something comparable to it that can run on <8gb consumer hardware. The main selling point of Stable Diffusion was its ability to run in that environment.

I feel like this is the greatest demand for LLMs at the moment too. It's hard to believe we're only 8 months into this industry, so I imagine we'll start seeing smaller footprints soon.

8 months from what point?

Gpt3 is 36 months old. Dalle-e is 28 months old. Even StableDiffusion is like 11 months old.

Re: Stable Diffusion XL 1.0

#78
post #5

In the meantime I've been getting good mileage out of Kandinsky - anyone got a good sense of how they compare?

This is the first I have heard of Kandinsky. Thanks for the tip. SDXL is a bigger model. There are some subjective comparison posts with SDXL 0.9, but I can't see them since they are on X :/

That sounds so weird, it took me a minute to understand. Go to nitter.net which has no login requirement and no ads, but all the same content that X (tmsfkat) has.

Re: Stable Diffusion XL 1.0

#79

Earlier quoted context omitted.

It's already supported in automatic1111 (see recent updates), and someone in the community will convert it to the automatic1111 format within minutes/hours after it's released on huggingface.

Sort of. IIRC (which may be unlikely) Auto1111 has the base model in the text to image plane, but if you want to use the refiner that is a separate IMG2IMG step/tab. Which would be a pain in the ass imo. The "Comfy" tool is node based and you can string both together which is nice. Although if you aren't confident in your images you don't need the refiner for a bit.

I think the diffusers UIs (like Invoke and VoltaML) are going to implement the refiner soon since HF already has a pipeline for it.

Comfy and A1111 are based around the original SD StabilityAI code, but the implementation must be pretty similar if they could add the base model so quickly.

Re: Stable Diffusion XL 1.0

#80

Earlier quoted context omitted.

This is the first I have heard of Kandinsky. Thanks for the tip. SDXL is a bigger model. There are some subjective comparison posts with SDXL 0.9, but I can't see them since they are on X :/

That sounds so weird, it took me a minute to understand. Go to nitter.net which has no login requirement and no ads, but all the same content that X (tmsfkat) has.

Two Nitter instances failed to load it, unfortunately.

And yeah, X is weird to type out too.

Post reply on HN