Live data from Hacker News

Viewing profile — mazoza

mazoza

HN member
Joined
Thu, Apr 22, 2021, 12:02 AM UTC
HN karma
21
Public activity
16 items

About mazoza

No profile information was provided.

Recent public activity

  1. comment
    Comment #42050799

    I dont actually see any tokens used in the model. It seems like the model actually predicts latents and then VAE converts back to audio. More like Tortoise or XTTS

  2. comment
    Comment #42044396

    Can one of the authors explain what this actually means from the post? hertz-vae: a 1.8 billion parameter transformer decoder which acts as a learned prior for the audio VAE. The m…

  3. story
  4. comment
    Comment #38340254

    meh this is not that good. Sounds quite boring.

  5. story
  6. story
  7. comment
    Comment #35272123

    I've been using Coqui Studio for a while now to generate AI voices for my video game project. It delivers most that is in the blogpost. I like it but I'd ask for a better API suppo…

  8. story
  9. comment
    Comment #32386053

    It is a lot faster, has more languages and models and can run inference on even CPU with many of their models.

  10. comment
    Comment #29194544

    Coqui is hosting a week-long competition/hackathon for STT in long-tail languages. You will be able to train STT models for your favorite languages using CoquiSTT ...and you get fr…

  11. story
  12. comment
    Comment #28085237

    I means it is faster than real time almost 10x So it is the contrary

  13. comment
    Comment #28074569

    https://github.com/coqui-ai/STT

  14. comment
    Comment #28074555

    I know the old speech team continues as Coqui https://github.com/coqui-ai/

  15. comment
    Comment #26897006

    It is sad seeing another Mozilla project going down to the hole. Also, why do they pay people to develop on something that is not being maintained?

  16. story