I'm running Asahi Linux on a 32GB M1 Pro. Any chance of being able to run text-to-image models locally? I've had some success with LLMs, but only the smaller models. No idea where to start with images, everything seems geared towards msft+nvda.
FLUX1.1 [pro] – New SotA text-to-image model from Black Forest Labs
91–100 of 159 posts
Re: FLUX1.1 [pro] – New SotA text-to-image model from Black Forest Labs
#92It doesn’t get piano keyboards right, but it’s the first image generator I’ve tried that sometimes get “someone playing accordion” mostly right. When I ask for a man playing accordion, it’s usually a somewhat flawed piano accordion, but If I ask for a woman playing accordion, it’s usually a button accordion. I’ve also seen a few that are half-button, half-piano monstrosities. Also, if I ask for “someone playing accor…
Periodic data is always hard for generative image systems - particularly if that "cycle" window is relatively large (as would be the case for octaves of a piano).
Re: FLUX1.1 [pro] – New SotA text-to-image model from Black Forest Labs
#93Pretty smart model. Here's one I made: https://replicate.com/p/6ez0x8xqvsrga0cjadg8m7bah0
Yet, it doesn't seem to know how a Tektronix 4010 actually looks like... ;) I had similar issues trying to paint a "I cast non-magic missile" meme with a fantasy wizard using a missile launcher. No model out there (I've tried SD, SDXL, FLUX.1dev and now this FLUX1.1pro) knows how a missile launcher looks like (neither as a generic term, nor any specific systems) and even has no clue how it's held, so they all draw re…
Re: FLUX1.1 [pro] – New SotA text-to-image model from Black Forest Labs
#94Pretty smart model. Here's one I made: https://replicate.com/p/6ez0x8xqvsrga0cjadg8m7bah0
flux is amazing, but I find it requires a very literal description, which pushes the "creative work" back to the text itself. Which can certainly be a good thing, just a bit less gratifying to non visual types like myself. :)
I wonder, only somewhat jokingly, if one could make text generators which "imagine" detailed fantastical scenes, suitable for feeding to a text to image model.
Re: FLUX1.1 [pro] – New SotA text-to-image model from Black Forest Labs
#95I'm running Asahi Linux on a 32GB M1 Pro. Any chance of being able to run text-to-image models locally? I've had some success with LLMs, but only the smaller models. No idea where to start with images, everything seems geared towards msft+nvda.
DiffusionBee will let you do this quite easily. edit: nevermind, it's a macos app
Re: FLUX1.1 [pro] – New SotA text-to-image model from Black Forest Labs
#96Pretty smart model. Here's one I made: https://replicate.com/p/6ez0x8xqvsrga0cjadg8m7bah0
It's quite good at following a detailed paragraph long description of an scene, which is a double edged sword. A lot of the fun for me with early text to image models was underspecifying an image and then enjoying how the model "invents" it. "Steampunk spaceship", "communist bear", "glass city". flux is amazing, but I find it requires a very literal description, which pushes the "creative work" back to the text itsel…
Re: FLUX1.1 [pro] – New SotA text-to-image model from Black Forest Labs
#97Re: FLUX1.1 [pro] – New SotA text-to-image model from Black Forest Labs
#98Are there any projects that allow for easy setup and hosting Flux locally? Similar to SD projects like InvokeAI or a1111
Re: FLUX1.1 [pro] – New SotA text-to-image model from Black Forest Labs
#99Pretty smart model. Here's one I made: https://replicate.com/p/6ez0x8xqvsrga0cjadg8m7bah0
This is miles ahead of most other image generation models available today.