Live data from Hacker News

Flux: Open-source text-to-image model with 12B parameters

blog.fal.ai

171–180 of 239 posts

Re: Flux: Open-source text-to-image model with 12B parameters

#171
post #110

Earlier quoted context omitted.

https://replicate.com/black-forest-labs/flux-dev is working very nicely. No sign-up.

My go-to test for these tools so far has been the seven horned, seven eyed lamb mentioned in the Book of Revelation. Every tool I've tried has failed at this task.

Can you share the exact prompt you used?

Re: Flux: Open-source text-to-image model with 12B parameters

#172

Earlier quoted context omitted.

My go-to test for these tools so far has been the seven horned, seven eyed lamb mentioned in the Book of Revelation. Every tool I've tried has failed at this task.

Can you share the exact prompt you used?

Sure. Normally I try a few variants, but "lamb with seven horns" was what I tried when I made that post.

For what it's worth, I've previously asked in the Stable Diffusion Discord server for help generating a "lamb with seven horns and seven eyes" but the members there were also unsuccessful.

Re: Flux: Open-source text-to-image model with 12B parameters

#173
post #105

Earlier quoted context omitted.

I think we've generally run out of names to give projects and need to start reusing names. Maybe use letters to disambiguate them. Flux A is the ML library Flux B is the T2I model Flux C is the React library Flux D is the physics concept of power per unit area Flux E is the goo you put on solder

Don't forget Fl.ux, which was a very popular way to make "night shift" happen for more than a decade

f.lux

Re: Flux: Open-source text-to-image model with 12B parameters

#174

You can try the models here: (available without sign-in) FLUX.1 [schnell] (Apache 2.0, open weights, step distilled): https://fal.ai/models/fal-ai/flux/schnell (requires sign-in) FLUX.1 [dev] (non-commercial, open weights, guidance distilled): https://fal.ai/models/fal-ai/flux/dev FLUX.1 [pro] (closed source [only available thru APIs], SOTA, raw): https://fal.ai/models/fal-ai/flux-pro

What is the difference between schnell and dev? Just the kind of distillation?

>FLUX.1 [dev]: The base model

>FLUX.1 [schnell]: A distilled version of the base model that operates up to 10 times faster

It should also be noted that "schnell" is the German word for "fast".

Re: Flux: Open-source text-to-image model with 12B parameters

#175

hi friends! burkay from fal.ai here. would like to clarify that the model is NOT built by fal. all credit should go to Black Forest Labs ( https://blackforestlabs.ai/ ) which is a new co by the OG stable diffusion team. what we did at fal is take the model and run it on our inference engine optimized to run these kinds of models really really fast. feel free to give it a shot on the playgrounds. https://fal.ai/models…

You also might want to "clarify" that it is not open source (and neither are any of the other "open source" models). If you want to call it something, try "open weights", although the usage restrictions make even that a HUGE FUCKING STRETCH. Also, everybody should remember that these models are not copyrightable and you should never agree to any license for them...

When I read "open source" i thought they actually are doing open source instead of "open weights" this time. Surely they would expect to be called out on hackernews if they label it incorrectly...

Thanks for pointing that out @Hizonener

Re: Flux: Open-source text-to-image model with 12B parameters

#176

Am I missing something? The beach image they give still fails to follow the prompt in major ways.

The quality is difficult to judge consistently as there's variants among seed with the same prompt. And then there's the problem of cherry picked examples making the news. So I'm building a community gallery to generate Pro images for free, hope this at least increases the sample size https://fluxpro.art/

Re: Flux: Open-source text-to-image model with 12B parameters

#177

Earlier quoted context omitted.

As far as I know, none have been released. And it doesn't even really make sense, because, as I said, the models aren't copyrightable to begin with and therefore aren't licensable either. However, plenty of open source software exists. The fact that open source models don't exist doesn't excuse attempts to falsely claim the prestige of the phrase "open source".

> models aren't copyrightable to begin with You are wrong about that. It's a file with numbers. Which makes it a database or dataset and very much protected by copyright. That's why licenses are needed. For the phone book, things like open street maps, and indeed AI models. > The fact that open source models don't exist The fact that many people (myself included) routinely download and use models distributed under OS…

> You are wrong about that. It's a file with numbers. Which makes it a database or dataset and very much protected by copyright. That's why licenses are needed. For the phone book, things like open street maps, and indeed AI models.

This is only true in jurisdictions that follow the sweat of the brow doctrine, where effort alone without creativity is considered enough for copyright. In other places, such as the USA, collections of facts are not copyrightable and a minimal amount of creativity is required for something to qualify as copyrightable. The phone book is an example that is often used, actually, to demonstrate the difference.

https://en.wikipedia.org/wiki/Sweat_of_the_brow

Re: Flux: Open-source text-to-image model with 12B parameters

#178
post #110

Earlier quoted context omitted.

https://replicate.com/black-forest-labs/flux-dev is working very nicely. No sign-up.

My go-to test for these tools so far has been the seven horned, seven eyed lamb mentioned in the Book of Revelation. Every tool I've tried has failed at this task.

Ah. I try the following:

> A Gary Larsen, "Far Side" comic of a racoon disguising itself by wearing a fedora and long trench coat. The raccoon's face is mostly hidden by the fedora. There are extra paws sticking out of the front of the trench coat from between the buttons, suggesting that the racoon is in fact a stack of several raccoons.

Every human I've ever described this to has no problem picturing what I mean. It's a classic comic trope. AIs still struggle.

Re: Flux: Open-source text-to-image model with 12B parameters

#179
post #142

Earlier quoted context omitted.

> it is not open source It would be nice here if you give some examples of what you call open source model. Please ;) Because the impression is that these things do not exist, it's just a dream which does not deserve such a nice term..

I'm personally comfortable calling a model "open source" if the license is compatible with the https://opensource.org/ definition. The Llama models aren't. Some of the Mistral models are (the Apache 2 ones). Microsoft Phi-3 is - it's MIT.

Open source must include source material so that another can reproduce that the model. I would expect that to be a minimum.

Re: Flux: Open-source text-to-image model with 12B parameters

#180
post #110

Earlier quoted context omitted.

The playground is a drag. After accepting being forced to sign up, attach my GitHub, and hand over my email address, I entered the desired prompt and waited with anticipation.. Only to see a black screen and how much it's going to cost per megapixel. Bummer. After seeing what was generated in the blog post I was excited to try it! Now feeling disappointed. I was hoping it'd be more like https://play.go.dev . Good luc…

https://replicate.com/black-forest-labs/flux-dev is working very nicely. No sign-up.

Is there a noscript/basic (x)html prompt?
Post reply on HN