Live data from Hacker News

Flux 3 X Mimic: The Next Generation of Video-Action Models

bfl.ai

51–57 of 57 posts

Re: Flux 3 X Mimic: The Next Generation of Video-Action Models

#51
post #29

Earlier quoted context omitted.

time is the great filter

What's a good recent movie I could watch? Today is Friday

Do you like horror? That's the only genre where risky projects are getting consistent funding. Obsession and Backrooms are both really good.

Re: Flux 3 X Mimic: The Next Generation of Video-Action Models

#52
post #24

Earlier quoted context omitted.

We’re just pulling signs of LLM touched writing out of our ass now. Might be time to move on from the accusations, assume all writing is at least LLM assisted and judge it purely on the quality.

Yeah, as if people haven't used words like "irregardless" way before LLM's already

After all, the LLM's learned it from somewhere

Re: Flux 3 X Mimic: The Next Generation of Video-Action Models

#53

Ok the phrasing here, it’s, it’s just - > However, compared to more specialized approaches for representation learning they produce less disentangled representations, which puts a ceiling on their usefulness for tasks that require world understanding. Only an LLM would use a less disentangled representation of the concept “more entangled” when trying to explain to people in the real world why less disentangled repres…

> Ok the phrasing here, it’s, it’s just - That dash definitely means an LLM wrote this comment. /s

Nah, that was an n-dash, which is widely respected as humans-only.

Re: Flux 3 X Mimic: The Next Generation of Video-Action Models

#54
post #28
post #23

Earlier quoted context omitted.

Survivorship bias, possible that you just aren’t watching all the shitty ones that nobody remembers.

That's exactly the ones I like! On Saturdays I watch cheesy horror movies from the 80s. But with my GF I watch whatever is rated >= 8 on Imdb and the 8s from years ago are ALWAYS much, much better than the recent ones. Maybe I just don't like the style. All the over explaining, etc.

Still survivorship bias. Those movies are curated! Well curated if you like them I might add.

Re: Flux 3 X Mimic: The Next Generation of Video-Action Models

#56
post #9

Really interesting. Upshot: a well trained multimodal video generation model has a world representation model trained inside it. They’ve done some work lifting this world model out and deploying it to robots, where it seems to work well. On the one hand, this isn’t a new idea, and the quality video models certainly have understanding of materials, light, the world (at least in an Occam’s razor sense of understanding)…

Rhoda Ai also does vidéo to robot models
Post reply on HN