Live data from Hacker News

Viewing profile — jilijeanlouis

jilijeanlouis

HN member
Joined
Fri, Sep 24, 2021, 8:23 AM UTC
HN karma
18
Public activity
50 items

About jilijeanlouis

No profile information was provided.

Recent public activity

  1. story
  2. comment
    Comment #48846143

    [flagged]

  3. story
  4. comment
    Comment #48817194

    Pretty cool !

  5. story
  6. comment
    Comment #47595870

    did you try gladia: ranking #1 on STT blind test https://compare-stt.com/

  7. comment
    Comment #47595863

    same for gladia it's ranked top 1 in the STT blind tests: https://compare-stt.com/

  8. comment
    Comment #47502030

    Author here. We built this because we kept seeing different word error rates (WER) for the same models depending on who was testing and how. Normalization rules ended up being a bi…

  9. story
  10. story
  11. comment
    Comment #40703720

    Having worked with Whisper for quite some time now, it's true that hallucinations can be a real pain point. Long pauses between sentences / silence and background noise make it wor…

  12. comment
    Comment #40294270

    This is really easy to do: it's just an embedding of your voice. So typically like 10/30 sec max of your voice to configure this. You already do a similar setup for faceId. I agree…

  13. story
  14. comment
    Comment #40271741

    Our API, Gladia, supports speaker diarization. We use a hybrid enterprise-grade ASR system for speech-to-text, with our own version of Whisper at its core, and state-of-the-art ope…

  15. comment
    Comment #40247143

    Language detection in the presence of strong accents is, in my opinion, one of the most under-discussed biases in AI. Traditional ASR systems struggle when English (or any language…

  16. comment
  17. comment
    Comment #40178199

    Definitely twitter. This is where everything is announced and commented

  18. comment
    Comment #40178176

    There are actually a big opportunity for companies like loccus: https://www.loccus.ai/

  19. comment
    Comment #40178145

    Did you try other providers such as 11labs or open source like voice craft or openvoice

  20. comment
    Comment #40178074

    It fun that this question comes today, last night I was saying to myself that all audiobooks from audible sounded the same. But AWS TTS quality is bad so the closest would be play.…

  21. comment
    Comment #40178045

    Despite regulation I don’t Believe it will actually get implemented even governments have regulations around audio accessibility for education in particular, for years now, and sti…

  22. comment
  23. comment
    Comment #40177995

    BTW I’ve heard from researchers from a big GAFAM (can’t name) that they actually use gpt to do this on the ground truth and spot that correction and do a second pass of human label…

  24. comment
    Comment #40177979

    Why would you use a dedicated app? Does it have to be natively embedded in android ?

  25. comment
    Comment #40177974

    That’s a bit problem with Arabic especially because of dialects and the lack of good datasets (when I say I mean robust - including noisy dirty audios).Would you mind testing gladi…