Live data from Hacker News

Qwen3-Omni: Native Omni AI model for text, image and video

github.com

91–100 of 152 posts

Re: Qwen3-Omni: Native Omni AI model for text, image and video

#92

Earlier quoted context omitted.

government hardly has the capacity to ban foreign weights

The danger is that lawmakers, confused about the difference between foreign weights and foreign APIs, accidentally ban both.

There will be no confusion whatsoever, and no accidents. They don't write those bills themselves.

Whatever a given bill does is precisely what its authors, who are almost never elected by any constituency on Earth, intend for it to do.

Re: Qwen3-Omni: Native Omni AI model for text, image and video

#94
post #59

Earlier quoted context omitted.

same, did you figure it out?

I think so, you need to click the big jagged audio icon to start a voice session.

Is the Qwen3-Omni-Flash the same as Qwen3-Omni-30B-A3B, or is the Omni-Flash a different closed-source model?

Re: Qwen3-Omni: Native Omni AI model for text, image and video

#96

Interesting, the pacing seemed very slow when conversing in english, but when I spoke to it in spanish, it sounded much faster. It's really impressive that these models are going to be able to do real time translation and much more. The Chinese are going to end up owning the AI market if the American labs don't start competing on open weights. Americans may end up in a situation where they have some $1000-2000 device…

> Americans may end up in a situation where they have some $1000-2000 device at home with an open Chinese model running on it, if they care about privacy or owning their data. I think HN vastly overestimates the market for something like this. Yes, there are some people who would spend $2,000 to avoid having prompts go to any cloud service. However, most people don’t care. Paying $20 per month for a ChatGPT subscript…

The reason people will pay $2,000 for a private at home AI is porn.

Re: Qwen3-Omni: Native Omni AI model for text, image and video

#97

Interesting, the pacing seemed very slow when conversing in english, but when I spoke to it in spanish, it sounded much faster. It's really impressive that these models are going to be able to do real time translation and much more. The Chinese are going to end up owning the AI market if the American labs don't start competing on open weights. Americans may end up in a situation where they have some $1000-2000 device…

[deleted]

Re: Qwen3-Omni: Native Omni AI model for text, image and video

#98
post #29

Earlier quoted context omitted.

Americans may end up in a situation where they have some $1000-2000 device at home with an open Chinese model running on it Wouldn't worry about that, I'm pretty sure the government is going to ban running Chinese tech in this space sooner or later. And we won't even be able to download it. Not saying any of the bans will make any kind of sense, but I'm pretty sure they're gonna say this is a "strategic" space. And e…

Eh this is the internet. There's always a way. They couldn't ban piracy either.

Correction: they couldn't enforce a ban

Re: Qwen3-Omni: Native Omni AI model for text, image and video

#99
post #86

Earlier quoted context omitted.

sitting here in the US, reading that China is strongly urging the adoption of Linux and pushing for open CPU architectures like RISC-V and also self-hosted open models are we the baddies??

If there is a walled garden, and you aren't in it, you'll probably push for the walls to come down. No moral basis needed.

Could you elaborate on what you mean by "moral basis" in your comment?

Re: Qwen3-Omni: Native Omni AI model for text, image and video

#100

Does this support realtime speech to speech via API? If so, where is this hosted/documented? I wasn’t able to see any info. I’d love to use this in lieu of OAIs (expensive) real time speech to speech offering.

https://www.alibabacloud.com/help/en/model-studio/realtime?s...

Thank you. I'm looking for this as well. The realtime model is a closed-source model and it's different than the open Qwen3-Omni-30B-A3B, right?

I wonder how hard is it to turn the open-source model to be a realtime model.

Post reply on HN