It's incredible how negative and dismissive the comments in here are while here I am thinking the model actually looks impressively capable. But then again I heard the downers have always been the first to leave their dung comments here so let's see...
Flux 3
141–150 of 151 posts
Re: Flux 3
#142Earlier quoted context omitted.
Flux2.dev and even klein 9b are extremely close to sota. People who are saying otherwise probably haven't used them very much.
I run a fairly high-traffic site for generative image models focusing on complex prompt adherence. Flux.2 doesn’t score anywhere near SOTA proprietary models. Klein 9b is decent for image-to-image, but when used for pure generative purposes brings back SDXL levels of body horror (have some Gattica pianists). If you can get past the annoying JSON structuring, Ideogram 4 is probably the best option in the open‑weights…
I use flux2.dev on my 5090 with a prompt upsampler (that I host on my 3080ti, it's been so long since I looked at it, its either some qwen or Mistral model). I use the upsampler to produce json structure based on this prompting guide: https://docs.bfl.ml/guides/prompting_guide_flux2
For being able to run on my 5090, it produces incredibly consistent and stable results that work for my use case so well, I find myself reaching for it over proprietary models.
I'm always interested to see objective comparisons though. It's clear from yours that flux lags. But the fact that I can use it locally over closed source sota models is very compelling (to me).
You've convinced me to checkout ideogram4 though!
Re: Flux 3
#143Earlier quoted context omitted.
I run a fairly high-traffic site for generative image models focusing on complex prompt adherence. Flux.2 doesn’t score anywhere near SOTA proprietary models. Klein 9b is decent for image-to-image, but when used for pure generative purposes brings back SDXL levels of body horror (have some Gattica pianists). If you can get past the annoying JSON structuring, Ideogram 4 is probably the best option in the open‑weights…
I suppose "close" could be interpreted subjectively. I use flux2.dev on my 5090 with a prompt upsampler (that I host on my 3080ti, it's been so long since I looked at it, its either some qwen or Mistral model). I use the upsampler to produce json structure based on this prompting guide: https://docs.bfl.ml/guides/prompting_guide_flux2 For being able to run on my 5090, it produces incredibly consistent and stable resu…
I personally use Qwen3 8b LLM for this purpose, but I think that the KJ prompt builder node handles this out of the box in ComfyUI.
Re: Flux 3
#144These people are hiring in... Freiburg im Breisgau? Wonder how hiring is working out for them there.
You are completely ignorant. This is one of the best places to live and work on earth. Nonetheless, they also have positions in San Francisco for what is worth.
Your post would have been just fine without that first bit.
Re: Flux 3
#145Earlier quoted context omitted.
filmmakers have the most to gain here, now anyone with a good idea for a story can just prompt out their own full length movie or tv show.
Before AI: 90% of everything was crap After AI: 99% of everything is crap ... but there will be >10x more of it, so we'll ultimately end up with more good stuff.
Re: Flux 3
#146Re: Flux 3
#147Lots of words about multi-modal but then this: > our mission to develop real-world visual intelligence Visual is mono-modal, isn't it?
its doing video, audio, images and motion. I think that counts as multimodal.
And I'm not sure video, images and motion are actually 3 different things. Images are just still motion and videos capture motion, so it's really just "video and audio" of which one is visual, the other is not, thus my confused/surprised comment.
Re: Flux 3
#148Earlier quoted context omitted.
Before AI: 90% of everything was crap After AI: 99% of everything is crap ... but there will be >10x more of it, so we'll ultimately end up with more good stuff.
But since our capacity to experience stuff can't keep up with the sheer quantity of stuff, 99% of the stuff we see will be crap, even if there's technically a larger amount of good stuff in absolute terms.
There will come a day where some piece of AI content brings a smile to your lips before a tear to your eye.