Live data from Hacker News

Flux: Open-source text-to-image model with 12B parameters

blog.fal.ai

181–190 of 239 posts

Re: Flux: Open-source text-to-image model with 12B parameters

#181
post #110

Earlier quoted context omitted.

The playground is a drag. After accepting being forced to sign up, attach my GitHub, and hand over my email address, I entered the desired prompt and waited with anticipation.. Only to see a black screen and how much it's going to cost per megapixel. Bummer. After seeing what was generated in the blog post I was excited to try it! Now feeling disappointed. I was hoping it'd be more like https://play.go.dev . Good luc…

https://replicate.com/black-forest-labs/flux-dev is working very nicely. No sign-up.

also no signup for the optimized endpoint on fal https://fal.ai/models/fal-ai/flux/schnell

Re: Flux: Open-source text-to-image model with 12B parameters

#182
post #105

Earlier quoted context omitted.

I think we've generally run out of names to give projects and need to start reusing names. Maybe use letters to disambiguate them. Flux A is the ML library Flux B is the T2I model Flux C is the React library Flux D is the physics concept of power per unit area Flux E is the goo you put on solder

Don't forget Fl.ux, which was a very popular way to make "night shift" happen for more than a decade

And flux https://fluxcd.io/ and flux https://formulae.brew.sh/formula/flux

Re: Flux: Open-source text-to-image model with 12B parameters

#185

Earlier quoted context omitted.

You also might want to "clarify" that it is not open source (and neither are any of the other "open source" models). If you want to call it something, try "open weights", although the usage restrictions make even that a HUGE FUCKING STRETCH. Also, everybody should remember that these models are not copyrightable and you should never agree to any license for them...

It's certainly not true that models are not copyrightable; databases have copyright protection if creativity was involved in creating them. That said, I don't think outputs of the model are derivative works of it, any more than the model is a derivative of its training data, so it's not clear to me they can actually enforce what you do with them.

> It's certainly not true that models are not copyrightable; databases have copyright protection if creativity was involved in creating them.

Are you talking about https://en.wikipedia.org/wiki/Database_right or plain old copyright?

I'm no IP lawyer, but I've always thought that copyright put "requirements" on the artefact (i.e the threshold of originality), not the process.

In my jurisdiction we have database rights, meaning that you get IP protections for the artefact based on the work put into the process. For example a database of distances between adress pairs or something is probably not copyrightable, but can be protected under database rights if enough work was done to compile the data.

EDIT: Saw in another place in thread speaking about the https://en.wikipedia.org/wiki/Sweat_of_the_brow doctrine, relates to Database rights. (Neither of which notably are not applicable in the U.S)

Re: Flux: Open-source text-to-image model with 12B parameters

#186
post #178

Earlier quoted context omitted.

My go-to test for these tools so far has been the seven horned, seven eyed lamb mentioned in the Book of Revelation. Every tool I've tried has failed at this task.

Ah. I try the following: > A Gary Larsen, "Far Side" comic of a racoon disguising itself by wearing a fedora and long trench coat. The raccoon's face is mostly hidden by the fedora. There are extra paws sticking out of the front of the trench coat from between the buttons, suggesting that the racoon is in fact a stack of several raccoons. Every human I've ever described this to has no problem picturing what I mean. I…

A rough rule of thumb is that if a text-generator AI model of some size would struggle to understand your sentence, then an image-generator model a couple of times the size or even bigger would also struggle.

The intelligence just doesn't "fit" in there.

Personally I'm curious to see what would happen if someone burnt $100M of compute time on training a truly enormous image generator model, something the same-ish size as GPT4...

Re: Flux: Open-source text-to-image model with 12B parameters

#187
post #110

Earlier quoted context omitted.

https://replicate.com/black-forest-labs/flux-dev is working very nicely. No sign-up.

My go-to test for these tools so far has been the seven horned, seven eyed lamb mentioned in the Book of Revelation. Every tool I've tried has failed at this task.

That classic “bowl of ramen without chopsticks “ also fails. Haven’t seen any get that right yet either.

Re: Flux: Open-source text-to-image model with 12B parameters

#188
post #178

Earlier quoted context omitted.

My go-to test for these tools so far has been the seven horned, seven eyed lamb mentioned in the Book of Revelation. Every tool I've tried has failed at this task.

Ah. I try the following: > A Gary Larsen, "Far Side" comic of a racoon disguising itself by wearing a fedora and long trench coat. The raccoon's face is mostly hidden by the fedora. There are extra paws sticking out of the front of the trench coat from between the buttons, suggesting that the racoon is in fact a stack of several raccoons. Every human I've ever described this to has no problem picturing what I mean. I…

>Every human I've ever described this to has no problem picturing what I mean. It's a classic comic trope. AIs still struggle.

But AIs learn and therefore create in exactly the same way as humans, ostensibly on the same data. How can this be possible? /s

Re: Flux: Open-source text-to-image model with 12B parameters

#189

Earlier quoted context omitted.

My go-to test for these tools so far has been the seven horned, seven eyed lamb mentioned in the Book of Revelation. Every tool I've tried has failed at this task.

That classic “bowl of ramen without chopsticks “ also fails. Haven’t seen any get that right yet either.

Negation is a known weak spot. Aren't you just retesting that again and again? Does it tell you much beyond that?

Re: Flux: Open-source text-to-image model with 12B parameters

#190

Earlier quoted context omitted.

You also might want to "clarify" that it is not open source (and neither are any of the other "open source" models). If you want to call it something, try "open weights", although the usage restrictions make even that a HUGE FUCKING STRETCH. Also, everybody should remember that these models are not copyrightable and you should never agree to any license for them...

A personal bugbear is the AI fascination with calling themselves open source, virtue signalling I guess. Open weights is exactly right. Source code and arguably more important datasets are both required to replicate the work, which is more in the spirit of open source (and science). I think Meta is especially egregious here, given their history. Never underestimate the value of getting hordes of unpaid workers to ref…

> virtue signalling

I'd prefer "false advertising" - it's more direct and without the culture war baggage.

Post reply on HN