Live data from Hacker News

Show HN: Nari Qwen3-TTS and Qwen3-ASR – High accuracy, low latency and cost

narilabs.com

31–37 of 37 posts

Re: Show HN: Nari Qwen3-TTS and Qwen3-ASR – High accuracy, low latency and cost

#31
post #27

This is awesome. Thanks for pushing the audio pareto frontier forward. Probably far fetched for now, but I think the next big evolution is building the pareto/much cheaper alternative to GPT-Live-1. The STT/TTS market is quite saturated, while today, there's almost no cheap/open source alternative to GPT-Live-1.

Is this really something people want? Honestly you can properly lower the pricing at least by 50%+.

Getting something conversationally better has been done, the tool calling will likely be worse though.

The infrastructure for real time is really annoying though.

Re: Show HN: Nari Qwen3-TTS and Qwen3-ASR – High accuracy, low latency and cost

#32
post #17
post #5

If you're going to announce a TTS model, service, or whatever, you really need demos.

hey sorry about that, you can here a few of our voices here: https://narilabs.com/product/qwen3-tts/

The horizontal moving elements of examples become stuck and unable to be scolled once one of them is played. I'm using Vivaldi (chrome based) on Android

Re: Show HN: Nari Qwen3-TTS and Qwen3-ASR – High accuracy, low latency and cost

#34
post #27

This is awesome. Thanks for pushing the audio pareto frontier forward. Probably far fetched for now, but I think the next big evolution is building the pareto/much cheaper alternative to GPT-Live-1. The STT/TTS market is quite saturated, while today, there's almost no cheap/open source alternative to GPT-Live-1.

Is this really something people want? Honestly you can properly lower the pricing at least by 50%+. Getting something conversationally better has been done, the tool calling will likely be worse though. The infrastructure for real time is really annoying though.

if we can lower the pricing by not 50% but 10x, then I think it would be something people want. we are taking the bet that OSS models will take a huge chunk of market share not just in LLMs but in multimodal as well

Re: Show HN: Nari Qwen3-TTS and Qwen3-ASR – High accuracy, low latency and cost

#37
post #34

Earlier quoted context omitted.

Is this really something people want? Honestly you can properly lower the pricing at least by 50%+. Getting something conversationally better has been done, the tool calling will likely be worse though. The infrastructure for real time is really annoying though.

if we can lower the pricing by not 50% but 10x, then I think it would be something people want. we are taking the bet that OSS models will take a huge chunk of market share not just in LLMs but in multimodal as well

If audio only then 10x is 100% possible right now based on math of the services.

+ Video is unlikely unless they are willing to give up margins.

Didn't take long at all for Google to smash out with 50% lol

Post reply on HN