Live data from Hacker News

OnnxStream: Stable Diffusion XL 1.0 Base on a Raspberry Pi Zero 2

github.com

81–88 of 88 posts

Re: OnnxStream: Stable Diffusion XL 1.0 Base on a Raspberry Pi Zero 2

#81
post #78
post #77

Earlier quoted context omitted.

> Maybe you want to try dating? It wouldn't be an app but a subscription you can join and the AI will handle the UX. A subscription to.. an app?

To a service. No more apps anymore, the app is the AI who can mediate as a UI to anything.

That sounds… exhausting?

Re: OnnxStream: Stable Diffusion XL 1.0 Base on a Raspberry Pi Zero 2

#82

Earlier quoted context omitted.

> I found this claiming an A100 can generate 1 image/s. The article you linked is over a year old. Needless to say there have been a LOT of optimizations in the last year. Back then it was common to use 50+ steps for many of the common samplers. Current methods use a few steps like 1. This OnnxStream are using SDXL-turbo, and you can combine LCM and a few other methods to go very fast. The reason it's so much faster…

> Back then it was common to use 50+ steps for many of the common samplers. Current methods use a few steps like 1. The "look how fast we can go" method (turbo model with 1 step and without CFG) is blindingly fast, but the quality is...nothing close to what was being done in normal 50+ steps with normal setitngs gens. Realistically, even with Turbo+LCM, you're still going to 4+ steps (often 8+), with CFG, for reasona…

> Realistically, even with Turbo+LCM, you're still going to 4+ steps (often 8+), with CFG, for reasonable one-generation quality anywhere close to the images people generated at 50+ steps without Turbo/LCM.

For sure the only reason I considered comparing it that way was because the linked repo appears to also be going for a similar approach with 1 step/image on the pi.

From my own experience I've had a hard ever getting a decent image below 6~8steps, but this repo seems more focused on getting it to run in a reasonable amount of time at all, which understandably requires the minimal "maybe passable" settings.

Re: OnnxStream: Stable Diffusion XL 1.0 Base on a Raspberry Pi Zero 2

#83
post #80
post #79

Earlier quoted context omitted.

Music will still be like Spotify IMHO, you will pay a monthly subscription. There wouldn't be an app, like everything else.

How about something like photoshop? Or a text editor for when I want to write a book?

These are already under threat, people are editing images using text for a year now - no tools needed, you can describe what you want or tap on an area of the image and describe what changes you want. These days they started doing videos too.

Re: OnnxStream: Stable Diffusion XL 1.0 Base on a Raspberry Pi Zero 2

#85
post #83
post #80

Earlier quoted context omitted.

How about something like photoshop? Or a text editor for when I want to write a book?

These are already under threat, people are editing images using text for a year now - no tools needed, you can describe what you want or tap on an area of the image and describe what changes you want. These days they started doing videos too.

They’re still apps, and nobody that does anything minimally professional uses “no tools” (in fact it’s quite the opposite, AI added a mindnumbing number of tools to the toolset). I really don’t see how these will just “disappear” and become some amorphous “talk to the computer” interface

Re: OnnxStream: Stable Diffusion XL 1.0 Base on a Raspberry Pi Zero 2

#86
post #84
post #81

Earlier quoted context omitted.

That sounds… exhausting?

Why?

Talking is waaaaay more time consuming than tapping or typing, having to memorize what I’m able to do on my device instead of just taking a glance at its screen is a lot of cognitive overload, not having any shared UI with others is a recipe for eternal confusion…

I really don’t see any clear advantages of having a one-fits all shapeless interface that’s driven by prompts (verbal or not). Doesn’t feel like a UX that makes sense to me, and I haven’t heard a case where it does yet tbh (I do think the “personal assistant” makes sense, and might replace some stuff, but don’t see how it’d become the next interface for everything)

Re: OnnxStream: Stable Diffusion XL 1.0 Base on a Raspberry Pi Zero 2

#87
post #86
post #84

Earlier quoted context omitted.

Why?

Talking is waaaaay more time consuming than tapping or typing, having to memorize what I’m able to do on my device instead of just taking a glance at its screen is a lot of cognitive overload, not having any shared UI with others is a recipe for eternal confusion… I really don’t see any clear advantages of having a one-fits all shapeless interface that’s driven by prompts (verbal or not). Doesn’t feel like a UX that…

There's no reason why AI app wouldn't be able to interact with you through the most suitable UI.

For example, when playing music can show you the basic buttons but also let you type or speak for more advanced stuff like "let's do karaoke" or "why don't we go through move soundtracks by showing me iconic scenes from each movie while playing the songs".

Re: OnnxStream: Stable Diffusion XL 1.0 Base on a Raspberry Pi Zero 2

#88
post #85
post #83

Earlier quoted context omitted.

These are already under threat, people are editing images using text for a year now - no tools needed, you can describe what you want or tap on an area of the image and describe what changes you want. These days they started doing videos too.

They’re still apps, and nobody that does anything minimally professional uses “no tools” (in fact it’s quite the opposite, AI added a mindnumbing number of tools to the toolset). I really don’t see how these will just “disappear” and become some amorphous “talk to the computer” interface

It's not on professional level yet, but it was barely O.K. a few months ago. It's moving very fast.

The gist is, if something is learnable by practice the AI can do it because by training these machines we actually teach them patterns and methods. Any "blue collar" job like editing images is going away.

Post reply on HN