Live data from Hacker News

Fast and accurate language identification using fastText

fasttext.cc

11–13 of 13 posts

Re: Fast and accurate language identification using fastText

#12
post #5

Why is it just 93% accurate on Wikipedia? Is it that hard to identify languages?

I suspect it's due to mixed-language content on Wikipedia. A lot of Wikipedia articles talk about foreign language art and culture, this is one of the largest (if not the largest) single categories of content on non-English Wikipedias.

Yes, it's not so good on the samples with several languages
Post reply on HN