Live data from Hacker News

Launch HN: Retell AI (YC W24) – Conversational Speech API for Your LLM

news.ycombinator.com

121–130 of 182 posts

Re: Launch HN: Retell AI (YC W24) – Conversational Speech API for Your LLM

#121

Is there societal value that this product is harming?

We are dedicated to preventing that from happening. Spam calling and identity theft are key areas we will build guardrails around. Feel free to let us know if you think of any other case.

Re: Launch HN: Retell AI (YC W24) – Conversational Speech API for Your LLM

#122
One use case that I'd be interested in for this is training it to use my voice a la Descript. It's been really nice that our Head of Marketing can iterate on video voiceovers using my voice and all I had to do was read 30 seconds of copy.

Any plans for something like that? It'd be really interesting to have prospects what with an AI version of me to answer some basic questions.

Re: Launch HN: Retell AI (YC W24) – Conversational Speech API for Your LLM

#123

I'm always curious for things like this where people get training data.

If you are referring to the LLM used in the demo, it's a simple GPT. If you are referring to audio data, there are some (not a lot) public datasets, although be careful of the license of the dataset. To get more data, you might want to build a studio to collect from contracted voice actors, or you can purchase from other sources.

Re: Launch HN: Retell AI (YC W24) – Conversational Speech API for Your LLM

#124

One use case that I'd be interested in for this is training it to use my voice a la Descript. It's been really nice that our Head of Marketing can iterate on video voiceovers using my voice and all I had to do was read 30 seconds of copy. Any plans for something like that? It'd be really interesting to have prospects what with an AI version of me to answer some basic questions.

Voice cloning is definitely on our roadmap.

Re: Launch HN: Retell AI (YC W24) – Conversational Speech API for Your LLM

#127
post #66

Earlier quoted context omitted.

For TTS, we are currently integrating with different providers like Elevenlabs, Openai TTS, etc. We do have plans down the road to train our own TTS model.

Ah thank you! What's the lowest latency option you have found so far?

Deepgram TTS is pretty fast, but they have not publicly launched yet.

Re: Launch HN: Retell AI (YC W24) – Conversational Speech API for Your LLM

#128
Commercially, I struggle to see how this fits in to the most natural application, a contact center.

For example, take Amazon Connect pricing:

There is an Amazon Connect service usage charge, based on end-customer call duration. At $0.018 per minute * 7 minutes = $0.126

There is an inbound call per minute charge for German DID numbers. At $0.0040 per minute * 7 minutes = $0.0280

If I were to bring retell into the loop, I'm changing my self-service per minute cost from .018 + .004 = .0184 per minute to .1184 per minute at the cheapest setting. And from that, I don't have a clear case for the impact on deflection, and because of that I don't have clear ROI.

This isn't saying that the product isn't great -- but at the current price I struggle to see how anyone at scale can use it without eating the double whammy of LLM + Retell costs.

Re: Launch HN: Retell AI (YC W24) – Conversational Speech API for Your LLM

#129
Awesome product. We would love to use this for our app, but it wouldn't make sense economically. At $0.10 per minute it would cost significantly more than our existing TTS and SST solution. We've manually added a VAD and will have to add a way of handling interruption. All-in-all it roughly costs us $0.01 per minute and we just can't afford a 10X increase in costs.

Guessing you guys have found a use case with higher margins than ours which'll explain the price. Great work. Hope we can afford this one day.

Post reply on HN