This is so cool. The future is cool! I wonder how it will work on languages that have different grammatical structure than french/english? Like Finno-Ugric languages which have sort of a Yoda speech to them. Edit: In Finno-Ugric languages words later on in a sentence can completely change the meaning. Will be interesting to look at. It's considerate of them to name it after my favourite whisky.
If Finnish is not widely known, German is more familiar, and there you can put the "nicht" at the very end of a sentence, reversing its meaning. Also, the verb may come close to the end, after an extended description of the subject / object; in English, you want the verb early. Human translators somehow handle that; machines would likely exhibit a similar delay.
High-fidelity simultaneous speech-to-speech translation
41–50 of 59 posts
Re: High-fidelity simultaneous speech-to-speech translation
#42All these Japanese project names and no Japanese support (ToT)
Re: High-fidelity simultaneous speech-to-speech translation
#43You can try it out here (select translation instead of transcription) https://soniox.com/
Disclaimer: I work at Soniox.
Re: High-fidelity simultaneous speech-to-speech translation
#44Earlier quoted context omitted.
I don't know if you're multilingual, but some concepts are just legitimately easier to express in some languages; and the different grammatical structures that languages have can be useful for emphasising certain things, or to express subtle relationships between concepts. I'm not a particularly fluent speaker of Japanese and Russian, but I still find it helpful to drop into them sometimes when speaking with someone…
I have to second this. I study Japanese myself and the entire way the Japanese communicate is reflected so deeply in the language. There is so so much nuance to pretty much every sentence they speak and there are certain grammar points that carry more meaning in three syllables than what can be expressed in English or German in a full sentence. And ok turn this way of communicating shapes their culture too I believe.…
Re: High-fidelity simultaneous speech-to-speech translation
#45This is so cool. The future is cool! I wonder how it will work on languages that have different grammatical structure than french/english? Like Finno-Ugric languages which have sort of a Yoda speech to them. Edit: In Finno-Ugric languages words later on in a sentence can completely change the meaning. Will be interesting to look at. It's considerate of them to name it after my favourite whisky.
Re: High-fidelity simultaneous speech-to-speech translation
#46Earlier quoted context omitted.
Translators sure, interpreters no. Interpreters also have to factor in cultural context and customs, ensuring that meaning is conveyed without offence being given in formal contexts.
I don't see why software couldn't do that, if you give them the context.
Re: High-fidelity simultaneous speech-to-speech translation
#47For anyone else looking for examples: https://huggingface.co/spaces/kyutai/hibiki-samples
Re: High-fidelity simultaneous speech-to-speech translation
#48Earlier quoted context omitted.
If Finnish is not widely known, German is more familiar, and there you can put the "nicht" at the very end of a sentence, reversing its meaning. Also, the verb may come close to the end, after an extended description of the subject / object; in English, you want the verb early. Human translators somehow handle that; machines would likely exhibit a similar delay.
Vaguely related anecdote: have you ever dictated a number to a French speaker? When you say “forty-two” or “seventy-six”, an English speaker will start writing the 4 or the 7 the moment they hear the “forty” or the “seventy”. The French speaker will also write the 4 the moment they hear the “quarante” in “quarante-deux” (40+2), but when you say “soixante-seize” (60+16), they will (without thinking about it!) only sta…
Re: High-fidelity simultaneous speech-to-speech translation
#49Earlier quoted context omitted.
Translators sure, interpreters no. Interpreters also have to factor in cultural context and customs, ensuring that meaning is conveyed without offence being given in formal contexts.
That seems like something LLMs could eventually get good at
Re: High-fidelity simultaneous speech-to-speech translation
#50Earlier quoted context omitted.
I think it'll greatly increase cultural learning, by increasing the opportunity to interact with people. I've traveled to a lot of countries, and never learned more than a handful of words in each, primarily related to basic service interactions. I enjoyed talking to locals when they spoke English. I couldn't interact in any meaningful way with the vast majority of people, though. Learning languages is great. If you…
Thanks. Wonderful take and optimistic. You are correct I think.