Live data from Hacker News

Flux 3

bfl.ai

1–10 of 151 posts

Re: Flux 3

#3
Lots of words about multi-modal but then this:

> our mission to develop real-world visual intelligence

Visual is mono-modal, isn't it?

Re: Flux 3

#4
post #3

Lots of words about multi-modal but then this: > our mission to develop real-world visual intelligence Visual is mono-modal, isn't it?

Is this really the value-add comment you’re going with?

Re: Flux 3

#5
Sorry because pointing this is a bit tired by now, but reading already the first two paragraph thete is this unmistakable stench of LLM slop writing. Immediately disengaged.

Re: Flux 3

#7
> It jointly learns from images, videos, and audio within a unified architecture, because what it needs to learn is not any one of these elements in isolation.

I'm confused, videos contain images and audio ...?

Re: Flux 3

#9
Open-weight plans are near the bottom (Launch section):

    - Video and audio generation and editing through APIs and private weight access. (“FLUX 3 Video”)
    - Action prediction through selected research and commercial partners, beginning with mimic robotics (“FLUX-mimic and FLUX 3 Action”)
    - Image synthesis and editing through APIs and private weight access. (“FLUX 3 Image”)
    - Open-weight access to a multimodal backbone, for content creation (video, audio and image) and action prediction. (“FLUX 3 Dev”)

Re: Flux 3

#10
I hope the open-weight versions will be SOTA.

> Over the next few weeks and months, we will make the following capabilities available

> Open-weight access to a multimodal backbone, for content creation (video, audio and image) and action prediction. (“FLUX 3 Dev”)

> We will also release more technical details on the underlying approach.

Post reply on HN