Live data from Hacker News

High-fidelity simultaneous speech-to-speech translation

arxiv.org

11–20 of 59 posts

Re: High-fidelity simultaneous speech-to-speech translation

#11
This is so cool. The future is cool!

I wonder how it will work on languages that have different grammatical structure than french/english? Like Finno-Ugric languages which have sort of a Yoda speech to them. Edit: In Finno-Ugric languages words later on in a sentence can completely change the meaning. Will be interesting to look at.

It's considerate of them to name it after my favourite whisky.

Re: High-fidelity simultaneous speech-to-speech translation

#12

This is why I wonder about the value of language learning for reasons other than “I’m really passionate about it.” We are so close to interfaces that reduce the language barrier by a lot…

What about brain development and general intelligence. Knowledge will always have a value, or else we become slaves to the machine.

Re: High-fidelity simultaneous speech-to-speech translation

#14

This is so cool. The future is cool! I wonder how it will work on languages that have different grammatical structure than french/english? Like Finno-Ugric languages which have sort of a Yoda speech to them. Edit: In Finno-Ugric languages words later on in a sentence can completely change the meaning. Will be interesting to look at. It's considerate of them to name it after my favourite whisky.

The alignment between source and target is automatically inferred, basically by searching when the uncertainty over a given output word reduces the most once enough input words are seen. This is then lifted to the audio domain. In theory the same trick should work even with longer grammatical inversions between languages, although this will lead to larger delays. To be tested!

Re: High-fidelity simultaneous speech-to-speech translation

#18
post #9

Nice. I'm impressed. Translator jobs are going to go poof! overnight. Just sayin'.

Translators sure, interpreters no.

Interpreters also have to factor in cultural context and customs, ensuring that meaning is conveyed without offence being given in formal contexts.

Re: High-fidelity simultaneous speech-to-speech translation

#19

This is why I wonder about the value of language learning for reasons other than “I’m really passionate about it.” We are so close to interfaces that reduce the language barrier by a lot…

Well if you take a look ... at the Multistream Visualization examples provided in the demo page, it's jus ... t the same as existing human provided interpretation solution at best. Constant 3-5s delays, random pauses, and likely lots of omissions here and there to absorb differences in sentence structures. I'd argue this only nullified another one of excuses to not learn a language.
Post reply on HN