Earlier quoted context omitted.
As imagined by Marilyn Manson...
Pretty much! It shows off how the codec works to a great extent though as it seems to be misinterpreting parts of the music to be the pitch of the speech, so Paul's voice sounds weird at the start of most lines but okay throughout the lines. I've also run a BBC news report through the program with better results although it demonstrates that any background noise at all can throw things off significantly: https://twit…
When it comes to noisy speech, it should be possible to improve things by actually training on noisy speech (the current model is trained only on clean speech). Stay tuned :-)