Flux is so frustrating to me. Really good prompt adherence, strong ability to keep track of multiple parts of a scene, it's technically very impressive. However it seems to have had no training on art-art. I can't get it to generate even something that looks like Degas, for instance. And, I can't even fine tune a painterly art style of any sort into Flux dev. I get that there was working, living artist backlash at SD…
>However it seems to have had no training on art-art. I can't get it to generate even something that looks like Degas, for instance It feels like they just removed names from the datasets to make it worse at recreating famous people and artists.
FLUX1.1 [pro] – New SotA text-to-image model from Black Forest Labs
81–90 of 159 posts
Re: FLUX1.1 [pro] – New SotA text-to-image model from Black Forest Labs
#82Flux is so frustrating to me. Really good prompt adherence, strong ability to keep track of multiple parts of a scene, it's technically very impressive. However it seems to have had no training on art-art. I can't get it to generate even something that looks like Degas, for instance. And, I can't even fine tune a painterly art style of any sort into Flux dev. I get that there was working, living artist backlash at SD…
>but, it's a real loss. Both in terms of human knowledge of, say composition, emotion, and so on, but also for style diversity But that real art still exists, and can still be found, so what exactly is the loss here?
Re: FLUX1.1 [pro] – New SotA text-to-image model from Black Forest Labs
#83Flux is so frustrating to me. Really good prompt adherence, strong ability to keep track of multiple parts of a scene, it's technically very impressive. However it seems to have had no training on art-art. I can't get it to generate even something that looks like Degas, for instance. And, I can't even fine tune a painterly art style of any sort into Flux dev. I get that there was working, living artist backlash at SD…
I’ve had the same problem with photography styles, even though the photographer I’m going for is Prokudin-Gorskii who used emulsion plates in the 1910s and the entire Library of Congress collection is in the public domain. I’m curious how they even managed to remove them from the training data since the entire LoC is such an easy dataset to access.
Re: FLUX1.1 [pro] – New SotA text-to-image model from Black Forest Labs
#84Earlier quoted context omitted.
Doesn't take a lot of effort to get Flux dev/schnell to run on 3090s unquantized, but I agree that 24gb is the consumer GPU memory limit and there are many with less than that. Flux runs great on modern Mac hardware as well, if you have at least 32gb of unified memory.
Really? I tried using it in ComfyUI on my Mac Studio, failed, went searching for answers and all I could find said that something something fp8 can't run on a Mac, so I moved on.
Re: FLUX1.1 [pro] – New SotA text-to-image model from Black Forest Labs
#85Earlier quoted context omitted.
>but, it's a real loss. Both in terms of human knowledge of, say composition, emotion, and so on, but also for style diversity But that real art still exists, and can still be found, so what exactly is the loss here?
We may differ on our take about the usefulness of diffusion models, but I'd say it's a loss in that many of the visuals humans will see in the next ten years are going to be generated by these models, and I for one wish they weren't just trained on weeb shit.
And between 1995 and 2022 the amount of Art produced surpasses the cumulative output of all other periods of human history.
Re: FLUX1.1 [pro] – New SotA text-to-image model from Black Forest Labs
#86Earlier quoted context omitted.
I think that's part of what makes FLUX.1 so good: the content it's trained on is very similar . Diversity is a double-edged sword. It's a desirable feature where you want it, and an undesirable feature everywhere else. If you want an impressionist painting, then it's good to have Monet and Degas in the training corpus. On the other hand, if you want a photograph of water lilies, then it's good to keep Monet out of th…
DALL-E3 doesn't struggle with this. It's just opinions. There's no technical limitation. They chose to weaken the model in this regard.
Re: FLUX1.1 [pro] – New SotA text-to-image model from Black Forest Labs
#87Earlier quoted context omitted.
I've had a similar experience, incredible at generating a very specific style of image, but not great at generating anything with a specific style. I suspect we'll see the answer to this is LoRAs. Two examples that stick out are: - Flux Tarot v1 [0] - Flux Amateur Photography [1] Both of these do a great job of combining all the benefits of Flux with custom styles that seem to work quite well. [0] https://huggingface…
I like those, and there's an electroshock lora that's just awesome out there. That said, Tarot and others like it are "illustrator" type styles with extra juice. I have not successfully trained a LoRa for any painting style, Flux does not seem to know about painting.
Here are a few I've recently trained: https://civitai.com/user/dvyio
Re: FLUX1.1 [pro] – New SotA text-to-image model from Black Forest Labs
#88Earlier quoted context omitted.
I've had a similar experience, incredible at generating a very specific style of image, but not great at generating anything with a specific style. I suspect we'll see the answer to this is LoRAs. Two examples that stick out are: - Flux Tarot v1 [0] - Flux Amateur Photography [1] Both of these do a great job of combining all the benefits of Flux with custom styles that seem to work quite well. [0] https://huggingface…
I like those, and there's an electroshock lora that's just awesome out there. That said, Tarot and others like it are "illustrator" type styles with extra juice. I have not successfully trained a LoRa for any painting style, Flux does not seem to know about painting.
Re: FLUX1.1 [pro] – New SotA text-to-image model from Black Forest Labs
#89Earlier quoted context omitted.
We may differ on our take about the usefulness of diffusion models, but I'd say it's a loss in that many of the visuals humans will see in the next ten years are going to be generated by these models, and I for one wish they weren't just trained on weeb shit.
Just think that before 1995 (and in reality, decades later than that) most of the world would never have access to 99% of the worlds art. And between 1995 and 2022 the amount of Art produced surpasses the cumulative output of all other periods of human history.
Re: FLUX1.1 [pro] – New SotA text-to-image model from Black Forest Labs
#90I'm running Asahi Linux on a 32GB M1 Pro. Any chance of being able to run text-to-image models locally? I've had some success with LLMs, but only the smaller models. No idea where to start with images, everything seems geared towards msft+nvda.
"Draw Things" is a native Mac app for text to image. It's a a lot more advanced than DiffusionBee, it will download the models for you, and it's free. It's also available for iOS. (!)