Live data from Hacker News

Viewing profile — audiohermit

audiohermit

HN member
Joined
Fri, Jun 05, 2020, 7:05 PM UTC
HN karma
395
Public activity
6 items

About audiohermit

No profile information was provided.

Recent public activity

  1. comment
    Comment #23493524

    Not really, this is the only thing I know of in terms of collection: https://www.isca-speech.org/archive/Interspeech_2018/pdfs/24... Usually you're basing your recipe off of those …

  2. comment
    Comment #23493416

    Agreed. I didn't have a better comparison at hand. I'm looking at you GAN papers.

  3. comment
    Comment #23491459

    Much of the work in speech synthesis has been about closing the gap in vocoders, which take a generated spectrogram and output a waveform. There's a clear gap between practical onl…

  4. comment
    Comment #23490751

    I'll push back on this. The quality of the read speech should be a higher concern than having parallel data. Unless OP's wife is a teacher or actor/voice actor, if LibriSpeech tran…

  5. comment
    Comment #23490637

    I work in pathological speech processing/synthesis so I'm unfortunately familiar with your father's position. It really sucks that these people didn't know that archiving their voi…

  6. comment
    Comment #23490394

    Hey, speech ML researcher here. Make sure you have different recordings of different contexts. fifteen.ai's best TTS voices use ~90 min of utterances, some separated by emotion. If…