Live data from Hacker News

Viewing profile — iceychris

iceychris

HN member
Joined
Wed, Jul 25, 2018, 8:29 PM UTC
HN karma
153
Public activity
15 items

About iceychris

No profile information was provided.

Recent public activity

  1. comment
    Comment #28599702

    I'm using NixOS with i3 as my daily driver, can recommend.

  2. comment
    Comment #25117176

    Depending on your objective, noisy data might be useful. I'd like LibreASR to also work in noisy environments, so training on data that is noisy should already help a bit with that…

  3. comment
    Comment #25110498

    Audio boards based on ESP32 boards are quite under the radar and have lovely features for just a few bucks. Running LibreASR on a RPi should also be feasible soon. Thank you for yo…

  4. comment
    Comment #25101390

    Data and compute are the largest hurdles. I only have one GPU and training one model takes 3+ days, so I am limited by that. Also, scraping from YouTube takes time and a lot of sto…

  5. comment
    Comment #25100420

    I haven't trained on LibriSpeech exclusively, but yes, the perf on LibriSpeech dev is quite bad, around ~60.0 WER. If the poor alignment of yt captions is the issue, maybe concaten…

  6. comment
    Comment #25100331

    Yes, probably. The data I trained on mostly reflects UK and US accents.

  7. comment
    Comment #25100253

    Hey blackcat! Your project [0] helped me a lot! Pre-training the encoder sounds great, I'll maybe add it in the future. [0] https://github.com/theblackcat102/Online-Speech-Recognit…

  8. comment
    Comment #25100223

    I have not yet trained a french model. Also, the gif shows Macron speaking to the congress with his english accent [0] [0] https://www.youtube.com/watch?v=RqUc1h7bZQ4

  9. comment
    Comment #25100064

    Right, fixed it, thank you :D

  10. comment
    Comment #25100007

    As I commented above, very poorly. It's still early days.

  11. comment
    Comment #25099954

    LibriSpeech, Tatoeba, Common Voice and scraped YouTube videos.

  12. comment
    Comment #25099938

    The upper transcript is YouTube's automatic transcription. Below is the web app transcribing live. And yes, it is actually missing a few words.

  13. comment
    Comment #25099920

    Hey HN! I've been working on this for a while now. While there are other on-premise solutions using older models such as DeepSpeech [0], I haven't found a deployable project suppor…

  14. story
  15. comment
    Comment #17612525

    username: iceychris Thank you very much :)