Live data from Hacker News

FLUX1.1 [pro] – New SotA text-to-image model from Black Forest Labs

replicate.com

51–60 of 159 posts

Re: FLUX1.1 [pro] – New SotA text-to-image model from Black Forest Labs

#51

Flux is so frustrating to me. Really good prompt adherence, strong ability to keep track of multiple parts of a scene, it's technically very impressive. However it seems to have had no training on art-art. I can't get it to generate even something that looks like Degas, for instance. And, I can't even fine tune a painterly art style of any sort into Flux dev. I get that there was working, living artist backlash at SD…

I think that's part of what makes FLUX.1 so good: the content it's trained on is very similar.

Diversity is a double-edged sword. It's a desirable feature where you want it, and an undesirable feature everywhere else. If you want an impressionist painting, then it's good to have Monet and Degas in the training corpus. On the other hand, if you want a photograph of water lilies, then it's good to keep Monet out of the training data.

Re: FLUX1.1 [pro] – New SotA text-to-image model from Black Forest Labs

#52
post #48

Pretty smart model. Here's one I made: https://replicate.com/p/6ez0x8xqvsrga0cjadg8m7bah0

That is astoundingly good adherence to the description. I already liked and was impressed by Flux1 but that is perhaps the most impressive image generation I've ever seen.

Re: FLUX1.1 [pro] – New SotA text-to-image model from Black Forest Labs

#53

Earlier quoted context omitted.

The point is that the metrics say the thing, this stuff doesn't say actually anything. What does "state of the art" mean? That it's using the latest "cutting edge" model technology? When Apple releases a new iPhone Pro Max, it's "state of the art". When they release a new iPhone SE, there's an argument to be made that it's not because it uses 2 year old chips. But what would it even mean for BFL to release a model wh…

Holy shit the level of pedantry. State of the art in this context means it out performs all other models to date on standard evaluations, which is precisely what it does. Did you miss the first flux release? Black forest labs aren't screwing around. The team consists of many of the _actual_ originators of Stable Diffusion's research (which was effectively co-opted by Emad Mostaque who is likely a sociopath).

> State of the art in this context means it out performs all other models to date on standard evaluations, which is precisely what it does.

That's not what "state of the art" means, and if it did it would still be hollow marketing jargon, because there are specific and meaningful ways to say that FLUX1.1 [pro] outperforms all competitors (and they do say so, later in the press release)

Your confusion about what "state of the art" means is exactly why marketers still use the phrase even though it has been overused and worn out since at least the 1980's. State of the art means something is "new", and that it is the "latest development", and that it incorporates "cutting edge" technology. The implication is that new is better, and that the "state of the art" is an improvement over what came before. (And to be clear, that's often true! Including in this case!) But that's not what the phrase actually means, it just means that something is new. And every press release is about something new.

FLUX1.1 [pro] would be state of the art even if it was worse than the previous version. Stable Diffusion 2.0 was state of the art when it was released.

Re: FLUX1.1 [pro] – New SotA text-to-image model from Black Forest Labs

#55
post #34
post #12

Earlier quoted context omitted.

What do you mean? FLUX.1 prompts women or women faces just fine? Do you mean the skin texture is unrealistic or some other artifacts?

Flux tends to gravitate towards a single face archetype for both sexes. For women it's a narrow face with a very slightly cleft chin. Men almost always appear with a very short cut beard or stubble. r/stablediffusion calls it the "flux face", and there are several LoRAs that aim to steer the model away from them.

[deleted]

Re: FLUX1.1 [pro] – New SotA text-to-image model from Black Forest Labs

#56
post #48

Pretty smart model. Here's one I made: https://replicate.com/p/6ez0x8xqvsrga0cjadg8m7bah0

Yet, it doesn't seem to know how a Tektronix 4010 actually looks like... ;)

I had similar issues trying to paint a "I cast non-magic missile" meme with a fantasy wizard using a missile launcher. No model out there (I've tried SD, SDXL, FLUX.1dev and now this FLUX1.1pro) knows how a missile launcher looks like (neither as a generic term, nor any specific systems) and even has no clue how it's held, so they all draw really weird contraptions.

Re: FLUX1.1 [pro] – New SotA text-to-image model from Black Forest Labs

#57
post #12

I'm worried about what happens when more people find out about Ideogram. There are a lot of things that don't appear in ELO scores. For one, they will not reflect that you cannot prompt women's faces in Flux. We can only speculate why.

What do you mean? FLUX.1 prompts women or women faces just fine? Do you mean the skin texture is unrealistic or some other artifacts?

what they really mean is that it's not useful for generating lewd imagery of women. It was likely nerfed in this regard on purpose because BFL didn't want to be associated with that (however legal it may be).

Re: FLUX1.1 [pro] – New SotA text-to-image model from Black Forest Labs

#58
post #30
post #13

Earlier quoted context omitted.

Flux genuinely is the best model I’ve tried though. If there is a better one I’d love to know.

Have you tried Ideogram v2?

Have you run Ideogram offline?

Re: FLUX1.1 [pro] – New SotA text-to-image model from Black Forest Labs

#59

Earlier quoted context omitted.

Holy shit the level of pedantry. State of the art in this context means it out performs all other models to date on standard evaluations, which is precisely what it does. Did you miss the first flux release? Black forest labs aren't screwing around. The team consists of many of the _actual_ originators of Stable Diffusion's research (which was effectively co-opted by Emad Mostaque who is likely a sociopath).

> State of the art in this context means it out performs all other models to date on standard evaluations, which is precisely what it does. That's not what "state of the art" means, and if it did it would still be hollow marketing jargon, because there are specific and meaningful ways to say that FLUX1.1 [pro] outperforms all competitors (and they do say so, later in the press release) Your confusion about what "stat…

I said in this context for a reason. That's how state of the art has been used (in papers, not copy) with regard to deep learning since well before DALL-E 1. I maintain that you're being pedantic about appropriating a term of art to mean something else. Everyone else here knows what the meaning is in context. Just not you.

Re: FLUX1.1 [pro] – New SotA text-to-image model from Black Forest Labs

#60

Earlier quoted context omitted.

Flux is more weird than old SD projects since Flux is extremely resource dependant and won't run on most hardware.

Doesn't take a lot of effort to get Flux dev/schnell to run on 3090s unquantized, but I agree that 24gb is the consumer GPU memory limit and there are many with less than that. Flux runs great on modern Mac hardware as well, if you have at least 32gb of unified memory.

Really? I tried using it in ComfyUI on my Mac Studio, failed, went searching for answers and all I could find said that something something fp8 can't run on a Mac, so I moved on.
Post reply on HN