Viewing profile — jilijeanlouis
jilijeanlouis
HN member- Joined
- Fri, Sep 24, 2021, 8:23 AM UTC
- HN karma
- 18
- Public activity
- 50 items
- HN profile
- View on Hacker News ↗
About jilijeanlouis
No profile information was provided.
Recent public activity
- story
-
comment
Comment #48846143
[flagged]
- story
-
comment
Comment #48817194
Pretty cool !
- story
-
comment
Comment #47595870
did you try gladia: ranking #1 on STT blind test https://compare-stt.com/
-
comment
Comment #47595863
same for gladia it's ranked top 1 in the STT blind tests: https://compare-stt.com/
-
comment
Comment #47502030
Author here. We built this because we kept seeing different word error rates (WER) for the same models depending on who was testing and how. Normalization rules ended up being a bi…
- story
- story
-
comment
Comment #40703720
Having worked with Whisper for quite some time now, it's true that hallucinations can be a real pain point. Long pauses between sentences / silence and background noise make it wor…
-
comment
Comment #40294270
This is really easy to do: it's just an embedding of your voice. So typically like 10/30 sec max of your voice to configure this. You already do a similar setup for faceId. I agree…
- story
-
comment
Comment #40271741
Our API, Gladia, supports speaker diarization. We use a hybrid enterprise-grade ASR system for speech-to-text, with our own version of Whisper at its core, and state-of-the-art ope…
-
comment
Comment #40247143
Language detection in the presence of strong accents is, in my opinion, one of the most under-discussed biases in AI. Traditional ASR systems struggle when English (or any language…
-
comment
Comment #40199841
[dead]
-
comment
Comment #40178199
Definitely twitter. This is where everything is announced and commented
-
comment
Comment #40178176
There are actually a big opportunity for companies like loccus: https://www.loccus.ai/
-
comment
Comment #40178145
Did you try other providers such as 11labs or open source like voice craft or openvoice
-
comment
Comment #40178074
It fun that this question comes today, last night I was saying to myself that all audiobooks from audible sounded the same. But AWS TTS quality is bad so the closest would be play.…
-
comment
Comment #40178045
Despite regulation I don’t Believe it will actually get implemented even governments have regulations around audio accessibility for education in particular, for years now, and sti…
-
comment
Comment #40178029
+1000
-
comment
Comment #40177995
BTW I’ve heard from researchers from a big GAFAM (can’t name) that they actually use gpt to do this on the ground truth and spot that correction and do a second pass of human label…
-
comment
Comment #40177979
Why would you use a dedicated app? Does it have to be natively embedded in android ?
-
comment
Comment #40177974
That’s a bit problem with Arabic especially because of dialects and the lack of good datasets (when I say I mean robust - including noisy dirty audios).Would you mind testing gladi…