Earlier quoted context omitted.
So it seems to be using AWS Polly's "standard" Joanna voice. However, you can also use "neural" voices (basically enhanced with ML) which sound somewhat better. Here's Joanna in neural mode: http://no.gd/voice2.mp3 .. and here's a British female voice which sounds pretty good: http://no.gd/voice1.mp3 Neural voices are 4x the price of standard ones though! Standard is $4 per 1 million characters, neural is $16. https:…
I really wish I could use the Neural voices on this, but at $4 per million characters, I'm just beyond breaking even. Since this is a self-funded side project, I can't run at a loss, and I don't think you can reasonably charge more for the kind of service ListenLater.fm provides, but I'm really excited about the prices coming down in this area. For what it's worth, Google, Amazon, and basically everywhere else I've l…
It probably pays to get a foot in the door with this idea now because if the prices do eventually collapse (or, better, high quality open source TTS makes huge advances) you'll be well positioned to take advantage of it.