Viewing profile — audiohermit
audiohermit
HN member- Joined
- Fri, Jun 05, 2020, 7:05 PM UTC
- HN karma
- 395
- Public activity
- 6 items
- HN profile
- View on Hacker News ↗
About audiohermit
No profile information was provided.
Recent public activity
-
comment
Comment #23493524
Not really, this is the only thing I know of in terms of collection: https://www.isca-speech.org/archive/Interspeech_2018/pdfs/24... Usually you're basing your recipe off of those …
-
comment
Comment #23493416
Agreed. I didn't have a better comparison at hand. I'm looking at you GAN papers.
-
comment
Comment #23491459
Much of the work in speech synthesis has been about closing the gap in vocoders, which take a generated spectrogram and output a waveform. There's a clear gap between practical onl…
-
comment
Comment #23490751
I'll push back on this. The quality of the read speech should be a higher concern than having parallel data. Unless OP's wife is a teacher or actor/voice actor, if LibriSpeech tran…
-
comment
Comment #23490637
I work in pathological speech processing/synthesis so I'm unfortunately familiar with your father's position. It really sucks that these people didn't know that archiving their voi…
-
comment
Comment #23490394
Hey, speech ML researcher here. Make sure you have different recordings of different contexts. fifteen.ai's best TTS voices use ~90 min of utterances, some separated by emotion. If…