English sounds really great, congrats! other languages I've tried doesn't sound that good, you can hear a strong english accent
German sounds okay.
The voice selection matters a lot for this research preview
41–50 of 168 posts
English sounds really great, congrats! other languages I've tried doesn't sound that good, you can hear a strong english accent
German sounds okay.
The voice selection matters a lot for this research preview
I did not see an British accent example. Generally it appears the TTS systems all do US accents and the British accent tends to sound like Frasier - an American faking an British accent.
We have lots of great British voices in our voice library! Or if you want to hear an american trying to do a british accent add "[British accent]" at the start of the generation
Their non-English (automated?) localization of the front page is ridiculously badly translated.
I didn't see anything about this in the documentation or prompting guide, but... is it supposed to be able to sing? Since I am a fundamentally unserious person, I copied in the Friends theme song lyrics into the demo and what came out was a singing voice with guitar. In another test, I added [verse] and [chorus] labels and it's singing acappella. [1] and [2] were prompted with just the lyrics. [3] was with the verse/…
They have some singing in their demo! So I’m guessing that’s baked into the model
The (American English) voices are absolutely amazing but the tags for laughs still feel more like an "inserted dedicated laugh section" than a "laugh at this point in speaking" type thing. I.e. it can't seem to reliably know when to giggle while saying a word, "just" giggle leading up to a word.
They're also still too expensive, and that's creating a lot of opportunity for other players. Even though ElevenLabs remains the quality leader, the others aren't that far behind. There are even a bunch of good TTS models being released as fully open source, especially by cutting-edge Chinese labs and companies. Perhaps in a bid to cut off the legs of American AI companies or to commoditize their compliment. Whatever…
We have a curated list of v3 voices in the library, but feel free to try others to find what works. Make sure language voice language match.
I didn't see anything about this in the documentation or prompting guide, but... is it supposed to be able to sing? Since I am a fundamentally unserious person, I copied in the Friends theme song lyrics into the demo and what came out was a singing voice with guitar. In another test, I added [verse] and [chorus] labels and it's singing acappella. [1] and [2] were prompted with just the lyrics. [3] was with the verse/…
From the example: "Oh no, I'm really sorry to hear you're having trouble with your new device. That sounds frustrating." Being patronized by a machine when you just want help is going to feel absolutely terrible. Not looking forward to this future.
quick note that that voice selection matters a lot with our new v3 model, especially voice language! We have a curated list of v3 voices in the library, but feel free to try others to find what works. Make sure language voice language match.