Flux: Open-source text-to-image model with 12B parameters
1–10 of 239 posts
Re: Flux: Open-source text-to-image model with 12B parameters
#2Re: Flux: Open-source text-to-image model with 12B parameters
#3Re: Flux: Open-source text-to-image model with 12B parameters
#4Result (distilled schnell model) for
"Photo of Criminal in a ski mask making a phone call in front of a store. There is caption on the bottom of the image: "It's time to Counter the Strike...". There is a red arrow pointing towards the caption. The red arrow is from a Red circle which has an image of Halo Master Chief in it."
Re: Flux: Open-source text-to-image model with 12B parameters
#5[flagged]
Re: Flux: Open-source text-to-image model with 12B parameters
#6You can try the models on replicate https://replicate.com/black-forest-labs . Result (distilled schnell model) for "Photo of Criminal in a ski mask making a phone call in front of a store. There is caption on the bottom of the image: "It's time to Counter the Strike...". There is a red arrow pointing towards the caption. The red arrow is from a Red circle which has an image of Halo Master Chief in it." https://www.re…
Re: Flux: Open-source text-to-image model with 12B parameters
#7[flagged]
Re: Flux: Open-source text-to-image model with 12B parameters
#8Re: Flux: Open-source text-to-image model with 12B parameters
#9I have seen a lot of promises made by diffusion models.
This is in a whole different world. I legitimately feel bad for the people still a StabilityAI.
The playground testing is really something else!
The licensing model isn’t bad, although I would like to see them promise to open up their old closed source models under Apache when they release new API versions.
The prompt adherence and the breadth of topics it seems to know without a finetune and without any LORAs, is really amazing.
Re: Flux: Open-source text-to-image model with 12B parameters
#10It is very fast and very good at rendering text, and appears to have a text encoder such that the model can handle both text and positioning much better: https://x.com/minimaxir/status/1819041076872908894
A fun consequence of better text rendering is that it means text watermarks from its training data appear more clearly: https://x.com/minimaxir/status/1819045012166127921