Live data from Hacker News

Meta AI announces Massive Multilingual Speech code, models for 1000+ languages

github.com

41–50 of 232 posts

Re: Meta AI announces Massive Multilingual Speech code, models for 1000+ languages

#41

Code: https://github.com/facebookresearch/fairseq/tree/main/exampl... Blog Post: https://ai.facebook.com/blog/multilingual-model-speech-recog... Paper: https://research.facebook.com/publications/scaling-speech-te... Languages coverage: https://dl.fbaipublicfiles.com/mms/misc/language_coverage_mm...

I loaded the language coverage into Datasette Lite and added some facets here:

https://lite.datasette.io/?json=https://gist.github.com/simo...

Here's how I did that: https://gist.github.com/simonw/63aa33ec827b093f9c6a2797df950...

Here are the top 20 represented language families:

    Niger-Congo 1,019
    Austronesian 609
    Sino-Tibetan 288
    Indo-European 278
    Afro-Asiatic 222
    Trans-New Guinea 219
    Otomanguean 149
    Nilo-Saharan 131
    Austro-Asiatic 100
    Dravidian 60
    Australian 51
    Creole 45
    Kra-Dai 43
    Uto-Aztecan 41
    Quechuan 36
    Language isolate 35
    Torricelli 32
    Maipurean 31
    Mayan 30
    Sepik 30

Re: Meta AI announces Massive Multilingual Speech code, models for 1000+ languages

#42
post #15

Come to think about it, Meta is a much better name for an AI company than a VR company.

It's not VR, it's Metaverse; VR might involve experiences that are not bullshit.

The metaverse, if we define it as an information layer that exists in parallel with the physical 'stuff' universe and is a seamless, effortless, and essential part of what we experience as reality will be an enormous part of our future. Meta might just be a few centuries ahead of the curve, which is just as bad as being a few centuries late.

Re: Meta AI announces Massive Multilingual Speech code, models for 1000+ languages

#43
post #40

I was kind of hoping for Interlingue, but was surprised to not even see Esperanto on the list.

Random thought ; could Esperanto (as a crypto-language) be turned into any sort of programming language. Could one, conceivably program in Esperanto in any meaningful way?

Re: Meta AI announces Massive Multilingual Speech code, models for 1000+ languages

#44

meta is doing more for open ai than openai

For OpenAI, AI is the revenue source. For Meta, it's a tool used to build products. If OpenAI gives their stuff away, they lose customers. If Meta does it, they can have community around it, and have joint effort about improving tools that then they'll use for their internal products. OpenAI is modern (and most likely - very short lived) Microsoft in the AI space, while Meta tries to replicate Linux in the AI space.

[deleted]

Re: Meta AI announces Massive Multilingual Speech code, models for 1000+ languages

#45

Code: https://github.com/facebookresearch/fairseq/tree/main/exampl... Blog Post: https://ai.facebook.com/blog/multilingual-model-speech-recog... Paper: https://research.facebook.com/publications/scaling-speech-te... Languages coverage: https://dl.fbaipublicfiles.com/mms/misc/language_coverage_mm...

Based on the availability of STT, TTS, and translation models available to download in that github repo, a real-life babelfish is 'only' some glue code away. Wild times we're living in.

AFAIK STT is still very bad without speaker-specific fine-tuning, so it's not going to be a literal babelfish (translating in the ear of the receiver), but it could make you speak many languages.

Re: Meta AI announces Massive Multilingual Speech code, models for 1000+ languages

#46
post #39
post #36

The problem with all these model releases is they have no demos or even video of it working. It’s all just download it and run it, like it’s an app.

I'd argue that having a "download and run" approach is so much better than videos or demos. Why do you think this is a problem?

Because to download and run it you need to have a laptop nearby to download and run it on, with the correct operating system and often additional dependencies too.

I do most of my research reading on my phone. I want to be able to understand things without breaking out a laptop.

Re: Meta AI announces Massive Multilingual Speech code, models for 1000+ languages

#48
post #40

I was kind of hoping for Interlingue, but was surprised to not even see Esperanto on the list.

Random thought ; could Esperanto (as a crypto-language) be turned into any sort of programming language. Could one, conceivably program in Esperanto in any meaningful way?

I think you are more likely to do so in a more logical and structured language like lojban

Re: Meta AI announces Massive Multilingual Speech code, models for 1000+ languages

#49
post #46
post #39

Earlier quoted context omitted.

I'd argue that having a "download and run" approach is so much better than videos or demos. Why do you think this is a problem?

Because to download and run it you need to have a laptop nearby to download and run it on, with the correct operating system and often additional dependencies too. I do most of my research reading on my phone. I want to be able to understand things without breaking out a laptop.

There's a paper associated with it that you can read on your phone. And I don't think demo videos are really associated with "research". I agree they could've added both, but let's be honest here you'll have demo videos on this in the next 12 hours for sure.

Re: Meta AI announces Massive Multilingual Speech code, models for 1000+ languages

#50
post #39
post #36

The problem with all these model releases is they have no demos or even video of it working. It’s all just download it and run it, like it’s an app.

I'd argue that having a "download and run" approach is so much better than videos or demos. Why do you think this is a problem?

It's a lot bigger investment. It'd be nice to at least see a video; that's easier than expecting people to download and run something just to see it.
Post reply on HN