Live data from Hacker News

Easy Stable Diffusion XL in your device, offline

noiselith.com

101–110 of 169 posts

Re: Easy Stable Diffusion XL in your device, offline

#101
post #69
post #14

Just installed, this is very cool. Local AI is the future I want (and what I'm working on too). A few notes using it... Pros: - seems pretty self contained - built in model installer works really well and helps you download anything from CivitAI (I installed https://civitai.com/models/183354/sdxl-ms-paint-portraits ) - image generation is high quality and stable - shows intermediate steps during generation Cons: - do…

> not open source like competitors Who are the competitors?

I'd also recommend InvokeAI, an open source offering which has a very nice editable canvas and is very performant with diffusers.

https://github.com/invoke-ai/InvokeAI

Re: Easy Stable Diffusion XL in your device, offline

#102
post #83

Sales prompt: "Young woman with blonde curls in front of a fantasy world background, come hither eyes, sitting with her legs spread, wearing a white shirt and jeans hot pants." I mean, really??

Glad I’m not the only one who found it inappropriate. Feels very much like a dog whistle.

Re: Easy Stable Diffusion XL in your device, offline

#103
post #14

Just installed, this is very cool. Local AI is the future I want (and what I'm working on too). A few notes using it... Pros: - seems pretty self contained - built in model installer works really well and helps you download anything from CivitAI (I installed https://civitai.com/models/183354/sdxl-ms-paint-portraits ) - image generation is high quality and stable - shows intermediate steps during generation Cons: - do…

+1 for asking download location.

Re: Easy Stable Diffusion XL in your device, offline

#104
post #83

Sales prompt: "Young woman with blonde curls in front of a fantasy world background, come hither eyes, sitting with her legs spread, wearing a white shirt and jeans hot pants." I mean, really??

Glad I’m not the only one who found it inappropriate. Feels very much like a dog whistle.

What's subtle about it? In the dog whistle analogy, who are they who cannot hear the whistle?

To me this is more like yelling "ROVER! COME HERE BOY!" at the top of your lungs.

Re: Easy Stable Diffusion XL in your device, offline

#105
post #78

Earlier quoted context omitted.

automatic1111: great for the fast implementation of the most recent generative features comfyui: excellent for workflows and recalling the workflows, as they're saved into the resulting image metadata (i.e. sharing images, shares the image generation pipeline) InvokeAI: Great UX and community, arguably were a bit behind in features as they were focused on making the UI work well. Now at the stage of bringing in the b…

> recalling the workflows, as they're saved into the resulting image metadata (i.e. sharing images, shares the image generation pipeline) Doesn't a1111 already do this? Theres a PNG Info tab where you can drag and drop a PNG and it will pull all the prompt, inverse prompt, model, etc. And then a button to send it to the main generation tab. It doesn't automatically load the model, but that may be intentional because…

Comfy is node based. The saved metadata pulls up the full nodal workflow.

Re: Easy Stable Diffusion XL in your device, offline

#106

There are already a number of local, inference options that are (crucially) open-source, with more robust feature sets. And if the defense here is "but Auto1111 and Comfy don't have as user-friendly a UI", that's also already covered. https://github.com/invoke-ai/InvokeAI

Yeah "Run Stable Diffusion locally" is a weird pitch since that's already easy to do tbh.

Re: Easy Stable Diffusion XL in your device, offline

#107
post #87

Earlier quoted context omitted.

"You should check out this thing" has a very different implied context than "You should check out this thing I made". The first sounds like a recommendation from an enthusiastic user, not from the the author. Because of this, discovering that you are the author makes your recommendation feel deceptive.

I am sorry if you feel that way. I joined HN when it was a small tight-knit community without much of marketing presence. The "obvious" comment is more like "people know other people" kind of thing. I didn't try to deceive anyone to use the app (and why should I?). If you feel this is unannounced self-promotion, yes, it is, and can be done better. --- Also, for the "objective" comment, it meant to say "the original c…

I think it was obvious. That said, thank you so much for your labor of love! The app is amazing! Any plans for SDXL Turbo support?

Re: Easy Stable Diffusion XL in your device, offline

#108

I would highly recommend Fooocus to anyone who hasn't tried: https://github.com/lllyasviel/Fooocus There are a bajillion local SD pipelines, but this one is, by far , the one with the highest quality output out-of-the-box, with short prompts. Its remarkable. And thats because it integrates a bajillion SDXL augmentations that other UIs do not implement or enable by default. I've been using stable diffusion since 1.5 c…

I was afraid of the Python setup (even though I'm a Python developer), but yep: Make the virtualenv, install the dependencies, done. This is amazing, the images it generates are immediately beautiful.

It does look bad that it bundles GTM, though, as a sibling commenter says.

Samples:

https://imgz.org/i9oicVqo/

https://imgz.org/i8Ur3WjW/

https://imgz.org/i5j6r6TZ/

Re: Easy Stable Diffusion XL in your device, offline

#109
post #78

Earlier quoted context omitted.

automatic1111: great for the fast implementation of the most recent generative features comfyui: excellent for workflows and recalling the workflows, as they're saved into the resulting image metadata (i.e. sharing images, shares the image generation pipeline) InvokeAI: Great UX and community, arguably were a bit behind in features as they were focused on making the UI work well. Now at the stage of bringing in the b…

> recalling the workflows, as they're saved into the resulting image metadata (i.e. sharing images, shares the image generation pipeline) Doesn't a1111 already do this? Theres a PNG Info tab where you can drag and drop a PNG and it will pull all the prompt, inverse prompt, model, etc. And then a button to send it to the main generation tab. It doesn't automatically load the model, but that may be intentional because…

> Doesn't a1111 already do this?

Not that provides the same thing, no, largely because of fundamental design differences.

> Theres a PNG Info tab where you can drag and drop a PNG and it will pull all the prompt, inverse prompt, model, etc. And then a button to send it to the main generation tab.

A1111 by nature, has a bunch of disconnected operations in separate tabs and scripts. Even if the PNG captures all of a generation operation that would be executed by a single launch-button click, its not really equivalent to capturing a whole ComfyUI workflow, which can be the equivalent of a process which would be numerous different tasks in A1111 with manually shuttling data between tabs and scripts.

A1111 has a bunch of manual "send to X" buttons to do with the output of runs, so that they can be the input of another task, wherein in Comfy those operations are part of one workflow with a pipeline connecting the output of one to the input of another. And when saving generation data, those manual shuttle points in A1111 are barriers as to what is part of a single generation that can be saved.

Re: Easy Stable Diffusion XL in your device, offline

#110

Earlier quoted context omitted.

If you want an absolute beast, especially for this stuff, you probably want Intel + Nvidia. Apple Silicon is a beast in power efficiency but a top of the line M3 does not come close to the top of the line Intel + Nvidia combo.

Well this would just be the excuse. I'm typing this on a Ryzen 5950X w/32 GB of RAM and a 4090. So I guess I already have the beast?

I guess EPYC and a few H100s is the next big step, at a much higer price point...
Post reply on HN