Live data from Hacker News

FLUX.2 [Klein]: Towards Interactive Visual Intelligence

bfl.ai

51–59 of 59 posts

Re: FLUX.2 [Klein]: Towards Interactive Visual Intelligence

#52
post #13

Earlier quoted context omitted.

A few days ago people were replying to every image on Twitter saying "Grok, put him/her/it in a bikini" and Grok would just do it. It was minimum effort, maximum damage trolling and people loved it.

Nah it's been happening for months and involved kids, over and over, albeit for the same reasoning, lulz & totally based. I am a bit surprised that you thought this was just a PG-rated stunt on X for a couple days, it's been in the news for weeks, including on HN.

I see absolutely no citations. Can you point to anything that shows a specific Grok issue vs generally people doing icky things with photo generation software?

Because, as I remember you said “post-pedo Grok”.

Re: FLUX.2 [Klein]: Towards Interactive Visual Intelligence

#53

I haven’t gotten around to adding Klein to my GenAI Showdown site yet, but if it’s anything like Z-Image Turbo, it should perform extremely well. For reference, Z-Image Turbo scored 4 out of 15 points on GenAI Showdown. I’m aware that doesn’t sound like much, but given that one of the largest models, Flux.2 (32b), only managed to outscore ZiT (a 6b model) by a single point and is significantly heavier-weight, that’s…

I think it shows problems with your tests tbh. The bigger models are way more capable than you make them out to be. They are also better in training and understanding of CGI render outputs as reference like normal maps or id-masks. Your testing suite is the perfect example that structured data implies false confidence. Pure t2i is not a good benchmark anymore.

Thanks for the feedback.

> The bigger models are way more capable than you make them out to be.

No test suite is ever going to be perfect. GenAI Showdown was started with the goal of focusing on a very narrow spectrum of testing (prompt adherence) because as a creator that's the one of the most interest to me.

> Pure t2i is not a good benchmark anymore

Just FYI Image Editing is already a separate benchmark (see the navbar at the top).

> Your testing suite is the perfect example that structured data implies false confidence

Again - the headline is "Specific prompts and challenges with a strong emphasis placed on adherence". If I tried to capture every possible aspect of GenAI models (multimodal, texture maps, periodic motion, tiling, etc) - I'd be at it until the heat death of the universe.

Incidentally - which model (specifically) do you think is ranked unfairly? While Flux.2 [dev] did only score a single point above ZiT, it's weighted score is much higher (1442 points vs 911 points).

Re: FLUX.2 [Klein]: Towards Interactive Visual Intelligence

#54

Earlier quoted context omitted.

Nah it's been happening for months and involved kids, over and over, albeit for the same reasoning, lulz & totally based. I am a bit surprised that you thought this was just a PG-rated stunt on X for a couple days, it's been in the news for weeks, including on HN.

I see absolutely no citations. Can you point to anything that shows a specific Grok issue vs generally people doing icky things with photo generation software? Because, as I remember you said “post-pedo Grok”.

You can Google whatever you need yourself at this point, you told the world I was operating in bad faith based off one sentence from a stranger. You ignored my reply to you. And now you are engaging with me on another reply as if my claim was Grok is uniquely capable of this, when I in fact said the opposite, and the interesting part of the discussion was me pointing out all can do this. Have a good day!

Re: FLUX.2 [Klein]: Towards Interactive Visual Intelligence

#55

Earlier quoted context omitted.

I see absolutely no citations. Can you point to anything that shows a specific Grok issue vs generally people doing icky things with photo generation software? Because, as I remember you said “post-pedo Grok”.

You can Google whatever you need yourself at this point, you told the world I was operating in bad faith based off one sentence from a stranger. You ignored my reply to you. And now you are engaging with me on another reply as if my claim was Grok is uniquely capable of this, when I in fact said the opposite, and the interesting part of the discussion was me pointing out all can do this. Have a good day!

“Post-pedo grok”

Just admit you’re very accustomed to shitting on x, grok, whatever Musk is associated with as a reinforcement to your political ideology.

Your comments weren’t about AI, thy were about Grok, and then you were incapable of defending that claim.

Re: FLUX.2 [Klein]: Towards Interactive Visual Intelligence

#56

Earlier quoted context omitted.

You can Google whatever you need yourself at this point, you told the world I was operating in bad faith based off one sentence from a stranger. You ignored my reply to you. And now you are engaging with me on another reply as if my claim was Grok is uniquely capable of this, when I in fact said the opposite, and the interesting part of the discussion was me pointing out all can do this. Have a good day!

“Post-pedo grok” Just admit you’re very accustomed to shitting on x, grok, whatever Musk is associated with as a reinforcement to your political ideology. Your comments weren’t about AI, thy were about Grok, and then you were incapable of defending that claim.

I am of no party or clique, why would Elon be doing moderation anyway? He has better things to do. If anything, sounded understaffed and thus taken advantage of by ne’er do wells - you can check if I’m pivoting by noting I noted in my original post every model can do this and Grok being focused on was a strange aberration.

I feel pathetic defending myself to someone who keeps reading my mind in the blandest way possible, then accuses me of wrongthought I must have had, based on things I never said. Hard to believe you’re living up to your ideals in this moment if you’re a fellow advocate for truth seekers and great men. I respect interlocution, but not repeated personal attacks based on thoughts projected and things unsaid. That’s not truth seeking behavior.

Re: FLUX.2 [Klein]: Towards Interactive Visual Intelligence

#59
post #26

Earlier quoted context omitted.

Quality is increasing, but these small models have very little knowledge compared to their big brothers (Qwen Image/Full size Flux 2). As in characters, artists, specific items, etc.

That's what LoRAs are for. And small models are also much easier to fine tune than large ones.

I hate that excuse. I want the model to know who the Paw Patrol is without either finding a lora (which probably won't exist because they're mostly porn) or needing to make a dataset, tag it, and then train it myself.
Post reply on HN