Live data from Hacker News

Qwen3-TTS family is now open sourced: Voice design, clone, and generation

qwen.ai

31–40 of 229 posts

Re: Qwen3-TTS family is now open sourced: Voice design, clone, and generation

#31

Qwen team, please please please, release something to outperform and surpass the coding abilities of Opus 4.5. Although I like the model, I don't like the leadership of that company and how close it is, how divisive they're in terms of politics.

Well DeepSeek V4 is rumored to be in that range and will be released in 3 weeks.

Re: Qwen3-TTS family is now open sourced: Voice design, clone, and generation

#32

Qwen team, please please please, release something to outperform and surpass the coding abilities of Opus 4.5. Although I like the model, I don't like the leadership of that company and how close it is, how divisive they're in terms of politics.

The Chinese labs distill the SOTA models to boost the performance of theirs. They are a trailer hooked up (with a 3-6 month long chain) to the trucks pushing the technology forwards. I've yet to see a trailer overtake it's truck. China would need an architectural breakthrough to leap American labs given the huge compute disparity.

I have seen indeed a trailer overtake its truck. Not a beautiful view.

Re: Qwen3-TTS family is now open sourced: Voice design, clone, and generation

#33
If you want to try out the voice cloning yourself you can do that an this Hugging Face demo: https://huggingface.co/spaces/Qwen/Qwen3-TTS - switch to the "Voice Clone" tab, paste in some example text and use the microphone option to record yourself reading that text - then paste in other text and have it generate a version of that read using your voice.

I shared a recording of audio I generated with that here: https://simonwillison.net/2026/Jan/22/qwen3-tts/

Re: Qwen3-TTS family is now open sourced: Voice design, clone, and generation

#36
post #28
post #9

Earlier quoted context omitted.

Have you tried the new GLM 4.7?

I've been using GLM 4.7 alongside Opus 4.5 and I can't believe how bad it is. Seriously. I spent 20 minutes yesterday trying to get GLM 4.7 to understand that a simple modal on a web page (vanilla JS and HTML!) wasn't displaying when a certain button was clicked. I hooked it up to Chrome MCP in Open Code as well. It constantly told me that it fixed the problem. In frustration, I opened Claude Code and just typed "Why…

I've used a bunch of the SOTA models (via my work's Windsurf subscription) for HTML/CSS/JS stuff over the past few months. Mind you, I am not a web developer, these are just internal and personal projects.

My experience is that all of the models seem to do a decent job of writing a whole application from scratch, up to a certain point of complexity. But as soon as you ask them for non-trivial modifications and bugfixes, they _usually_ go deep into rationalized rabbit holes into nowhere.

I burned through a lot of credits to try them all and Gemini tended to work the best for the things I was doing. But as always, YMMV.

Re: Qwen3-TTS family is now open sourced: Voice design, clone, and generation

#37

Earlier quoted context omitted.

>how divisive they're in terms of politics What do you mean by this?

Dario said not nice words about China and open models in general: https://www.bloomberg.com/news/articles/2026-01-20/anthropic...

I mean, there's no way it's about this right?

Being critical of favorable actions towards a rival country shouldn't be divisive, and if it is, well, I don't think the problem is in the criticism.

Also the link doesn't mention open source? From a google search, he doesn't seem to care much for it.

Re: Qwen3-TTS family is now open sourced: Voice design, clone, and generation

#38

Has anyone successfully run this on a Mac? The installation instructions appear to assume an NVIDIA GPU (CUDA, FlashAttention), and I’m not sure whether it works with PyTorch’s Metal/MPS backend.

I recommend using modal for renting the metal.

Re: Qwen3-TTS family is now open sourced: Voice design, clone, and generation

#39

Earlier quoted context omitted.

>how divisive they're in terms of politics What do you mean by this?

Dario said not nice words about China and open models in general: https://www.bloomberg.com/news/articles/2026-01-20/anthropic...

From the perspective of competing against China in terms of AI the argument against open models makes sense to me. It’s a terrible problem to have really. Ideally we should all be able to work together in the sandbox towards a better tomorrow but thats not reality.

I prefer to have more open models. On the other hand China closes up their open models once they start to show a competitive edge.

Post reply on HN