Live data from Hacker News

FLUX1.1 [pro] – New SotA text-to-image model from Black Forest Labs

replicate.com

31–40 of 159 posts

Re: FLUX1.1 [pro] – New SotA text-to-image model from Black Forest Labs

#31

I'm worried about what happens when more people find out about Ideogram. There are a lot of things that don't appear in ELO scores. For one, they will not reflect that you cannot prompt women's faces in Flux. We can only speculate why.

How locked down is it? My problem with a lot of these is I like to make really ridiculous meme type images, but I run into walls for dumb reasons. Like if I want to make something thats "copyrighted" like a mix of certain characters from one franchise or whatever, I cannot sometimes I get told that the model cannot generate copyrighted content, even though courts ruled that AI generated stuff cannot be copyrighted ei…

I've been using Venice.ai which offers afaik the most uncensored service currently available, outside of running your own instances. No problem with prompts that include copyrighted terms.

Re: FLUX1.1 [pro] – New SotA text-to-image model from Black Forest Labs

#32
post #8

Are there any projects that allow for easy setup and hosting Flux locally? Similar to SD projects like InvokeAI or a1111

Forge

https://github.com/lllyasviel/stable-diffusion-webui-forge

https://www.reddit.com/r/StableDiffusion/comments/1esxkk8/ho...

Re: FLUX1.1 [pro] – New SotA text-to-image model from Black Forest Labs

#33
It doesn’t get piano keyboards right, but it’s the first image generator I’ve tried that sometimes get “someone playing accordion” mostly right.

When I ask for a man playing accordion, it’s usually a somewhat flawed piano accordion, but If I ask for a woman playing accordion, it’s usually a button accordion. I’ve also seen a few that are half-button, half-piano monstrosities.

Also, if I ask for “someone playing accordion”, it’s always a woman.

Re: FLUX1.1 [pro] – New SotA text-to-image model from Black Forest Labs

#34
post #12

I'm worried about what happens when more people find out about Ideogram. There are a lot of things that don't appear in ELO scores. For one, they will not reflect that you cannot prompt women's faces in Flux. We can only speculate why.

What do you mean? FLUX.1 prompts women or women faces just fine? Do you mean the skin texture is unrealistic or some other artifacts?

Flux tends to gravitate towards a single face archetype for both sexes. For women it's a narrow face with a very slightly cleft chin. Men almost always appear with a very short cut beard or stubble. r/stablediffusion calls it the "flux face", and there are several LoRAs that aim to steer the model away from them.

Re: FLUX1.1 [pro] – New SotA text-to-image model from Black Forest Labs

#35

"state of the art" has become such tired marketing jargon. "our most advanced and efficient model yet" "a significant step forward in our mission to empower creators" I get it, you can't sell things if you don't market them, and you can't make a living making things if you don't sell them, but it's exhausting.

Agreed, but the flux dev model is easily the best model out there in terms of overall prompt adherence that can also be run locally.

Some comparisons against DALL-E 3.

https://mordenstar.com/blog/flux-comparisons

Re: FLUX1.1 [pro] – New SotA text-to-image model from Black Forest Labs

#36
Far more interesting will be when pony diffusion V7 launches.

No one in the image space wants to admit it, but well over half of your user base wants to generate hardcore NSFW with your models and they mostly don’t care about any other capabilities.

Re: FLUX1.1 [pro] – New SotA text-to-image model from Black Forest Labs

#37

Flux is so frustrating to me. Really good prompt adherence, strong ability to keep track of multiple parts of a scene, it's technically very impressive. However it seems to have had no training on art-art. I can't get it to generate even something that looks like Degas, for instance. And, I can't even fine tune a painterly art style of any sort into Flux dev. I get that there was working, living artist backlash at SD…

I’ve had the same problem with photography styles, even though the photographer I’m going for is Prokudin-Gorskii who used emulsion plates in the 1910s and the entire Library of Congress collection is in the public domain. I’m curious how they even managed to remove them from the training data since the entire LoC is such an easy dataset to access.

Re: FLUX1.1 [pro] – New SotA text-to-image model from Black Forest Labs

#38

Flux is so frustrating to me. Really good prompt adherence, strong ability to keep track of multiple parts of a scene, it's technically very impressive. However it seems to have had no training on art-art. I can't get it to generate even something that looks like Degas, for instance. And, I can't even fine tune a painterly art style of any sort into Flux dev. I get that there was working, living artist backlash at SD…

And I can't imagine there's a real copyright (or ethical) issue with including artwork in the public domain because the artist died over a century ago.

Re: FLUX1.1 [pro] – New SotA text-to-image model from Black Forest Labs

#39
post #8

Are there any projects that allow for easy setup and hosting Flux locally? Similar to SD projects like InvokeAI or a1111

The answer is it really depends on your hardware, but the nice thing is that you can split out the text encoder when using ComfyUI. On a 24gb VRAM card I can run the Q8_0 GGUF version of flux-dev with the T5 FP16 text encoder. The Q8_0 gguf version in particular has very little visual difference from the original fp16 models. A 1024x1024 image takes about 15 seconds to generate.

Re: FLUX1.1 [pro] – New SotA text-to-image model from Black Forest Labs

#40

"state of the art" has become such tired marketing jargon. "our most advanced and efficient model yet" "a significant step forward in our mission to empower creators" I get it, you can't sell things if you don't market them, and you can't make a living making things if you don't sell them, but it's exhausting.

Flux is state of the art. You can see an ELO-scored leaderboard here:

https://huggingface.co/spaces/ArtificialAnalysis/Text-to-Ima...

Post reply on HN