Live data from Hacker News

Complete silence is always hallucinated as "ترجمة نانسي قنقر" in Arabic

github.com

61–70 of 354 posts

Re: Complete silence is always hallucinated as "ترجمة نانسي قنقر" in Arabic

#61

Earlier quoted context omitted.

> I wonder if the ZDF gave its approval for it being used for LLM training though? I am pretty sure they didn't get asked.

Just like the people forced to pay for ZDF under threat of imprisonment.

Just like any other public service paid for with public funds?

Re: Complete silence is always hallucinated as "ترجمة نانسي قنقر" in Arabic

#62
post #33

Earlier quoted context omitted.

ُThe Arabic text is the translator's self credit "Translated by Nancy Qanfar"

And the German is “subtitles of [public broadcaster] for [content network], 2017 I'm not sure this is really overfitting, the network does exactly what the training data demands. According to the training data silence art the end transcribes to a copyright notice or subtitle credits

fitting on noise in the training data is exactly what overfitting is. underfitting is smoothing out signal

Re: Complete silence is always hallucinated as "ترجمة نانسي قنقر" in Arabic

#65
This is a nice reminder that there is no real reasoning in the "AI" it is just still guessing the next word. After being trained on subtitle files which I guess is actually a clever idea as they convey real conversations without pirating, subtitles are freely distributed after all by dedicated translators. Good to see they're the ones getting credit though!

Re: Complete silence is always hallucinated as "ترجمة نانسي قنقر" in Arabic

#66
Using Whisper to sub Japanese vtuber concerts for my enjoyment, I've noticed a similar trend. Not one specific phrase, but several. Some are strange ("I'm going to make a hole in the back of the head"), some are clearly from lyrics websites.

Re: Complete silence is always hallucinated as "ترجمة نانسي قنقر" in Arabic

#67
post #56

Earlier quoted context omitted.

Indeed, with another model I would get persistent transcriptions of silent parts into 'Thanks for watching!' or '[MUSIC]'. Pretty dumb that this failure mode wasn't caught in some QA process, and there are now multiple transcription models suffering from the same issue. Having silent parts in your input audio seems like it should be a very common occurrence...

When I was taught mathematics, the zero value was always considered the most important edge case. You prove something for N=0 (or N=1), then for N=M+1. It's even more important in audio DSP: processing near-zeroes can end up being extremely CPU intensive, look up denormal/subnormal floats.

Yeah, I studied mathematics (algebra and number theory) and zero is the point, often sporting discontinuities, or weird asymptotic behavior.

Quite a lot of algorithms use some form of division and zero is the only number in our typical structures (Z, Q, R, C), that cannot be used to divide with.

Re: Complete silence is always hallucinated as "ترجمة نانسي قنقر" in Arabic

#70
post #56

Earlier quoted context omitted.

When I was taught mathematics, the zero value was always considered the most important edge case. You prove something for N=0 (or N=1), then for N=M+1. It's even more important in audio DSP: processing near-zeroes can end up being extremely CPU intensive, look up denormal/subnormal floats.

Yeah, I studied mathematics (algebra and number theory) and zero is the point, often sporting discontinuities, or weird asymptotic behavior. Quite a lot of algorithms use some form of division and zero is the only number in our typical structures (Z, Q, R, C), that cannot be used to divide with.

Well, now in this brave new age of AI we can enjoy computer programs crashing with an

    Error: division by please upvote, share and like!
Post reply on HN