Viewing profile — mazoza
mazoza
HN member- Joined
- Thu, Apr 22, 2021, 12:02 AM UTC
- HN karma
- 21
- Public activity
- 16 items
- HN profile
- View on Hacker News ↗
About mazoza
No profile information was provided.
Recent public activity
-
comment
Comment #42050799
I dont actually see any tokens used in the model. It seems like the model actually predicts latents and then VAE converts back to audio. More like Tortoise or XTTS
-
comment
Comment #42044396
Can one of the authors explain what this actually means from the post? hertz-vae: a 1.8 billion parameter transformer decoder which acts as a learned prior for the audio VAE. The m…
- story
-
comment
Comment #38340254
meh this is not that good. Sounds quite boring.
- story
- story
-
comment
Comment #35272123
I've been using Coqui Studio for a while now to generate AI voices for my video game project. It delivers most that is in the blogpost. I like it but I'd ask for a better API suppo…
- story
-
comment
Comment #32386053
It is a lot faster, has more languages and models and can run inference on even CPU with many of their models.
-
comment
Comment #29194544
Coqui is hosting a week-long competition/hackathon for STT in long-tail languages. You will be able to train STT models for your favorite languages using CoquiSTT ...and you get fr…
- story
-
comment
Comment #28085237
I means it is faster than real time almost 10x So it is the contrary
-
comment
Comment #28074569
https://github.com/coqui-ai/STT
-
comment
Comment #28074555
I know the old speech team continues as Coqui https://github.com/coqui-ai/
-
comment
Comment #26897006
It is sad seeing another Mozilla project going down to the hole. Also, why do they pay people to develop on something that is not being maintained?
- story