Earlier quoted context omitted.
That makes assumptions about the language, for the languages I know: German has three definite articles, Latin doesn't have any, so it is not obvious what looking for "the" would result in either.
I think people (here and below) are getting hung up on definite articles, but Zipf's Law makes no such observation. It says only that a word's frequency in a natural language corpus tends to be in inverse proportion to its rank in a frequency table. In English, the most frequent words are articles, but the general observation about word frequency holds across languages (whether those languages have articles or not).
Minoan language Linear A linked to Linear B in new research
41–50 of 92 posts
Re: Minoan language Linear A linked to Linear B in new research
#42This article is a reasonable summary of the status quo in Linear A studies since 1956, but the reporter seemingly deliberately obfuscates the nature and scope of Dr. Ester Salgarella's new contribution. It appears to be the creation of an online corpus in collaboration Dr. Simon Castellan, linked near the bottom of the article: https://sigla.phis.me/ This will be a great resource and is an important work, but it appe…
I don't know about "deliberately". Might be just a confused, inarticulate piece by a confused reporter. When an article opens with " the Minoan language known as Linear A " (no, it's a script) and " Linear B developed later in the prehistoric period " (an oxymoron, prehistoric = pre-literary)… you know not to expect much. Did anyone manage to parse out what is even being claimed in this article?
Re: Minoan language Linear A linked to Linear B in new research
#43Earlier quoted context omitted.
can't they use the zipf law ? https://en.wikipedia.org/wiki/Zipf%27s_law The decreasing exponential law should allow to find "the" and some closed form POS words, so yeah determiners, prepositions and conjonctions.
The total extant corpus of Linear A amounts to fewer than 10,000 characters (and this is, I believe, the largest corpus of any undeciphered script). There's not enough text to do statistical analysis.
Even worse than the small number of texts is that all, or almost all, are just bookkeeping records, so they contain few words besides numbers, symbols for useful goods, e.g. wine, olive oil, barley, wool and so on, and proper names of places or people.
So even if there might be a few hundreds of texts, most just reproduce the same phrases, only with different numbers and names substituted in them.
Any statistics on this handful of stereotype phrases will offer no information about the statistics of the words of the Minoan language as used in a normal conversation or story telling.
Re: Minoan language Linear A linked to Linear B in new research
#44Earlier quoted context omitted.
can't they use the zipf law ? https://en.wikipedia.org/wiki/Zipf%27s_law The decreasing exponential law should allow to find "the" and some closed form POS words, so yeah determiners, prepositions and conjonctions.
You really think that in decades of linguists studying Linear A, no one has thought of trying Zipf's law? If scientists have studied something for this long, and you come up with an idea that fits in a single paragraph, it's probably been tried and didn't work. Unless you're the field's leading expert in which case you would be off doing it, not posting it on HN :) Edit: typos
The literature survey section mentions there have been good results using computational methods in 2020 to automatically decipher Linear B. The discussion section mentions "To the best of our knowledge, this is the first study to discuss and show computational analysis of Linear A."
Again, neither of these are my fields, but it looks like if these linguists have tried to use Zipf's law or other computational methods unsuccessfully in deciphering Linear A, the results weren't published. (Or a poor literature survey, or other explanations...) I'm not an academic either, so I don't know what the practices are for publishing unsuccessful results.
Re: Minoan language Linear A linked to Linear B in new research
#45Earlier quoted context omitted.
Can we assume that any of those features would be part of this language?
a language without thoses would be hella weird and primitive, like stereotypical robotic talking. To answer you question, I don't know, do linear B have them?
One of the historical issues with linguistics is that it analyzed every language as if it were Classical Latin or Classical Greek, and if that language had elements that didn't work out... well, that can't be proper then, can it? You still see some residuum of this in English prescriptivist poppycock, like the prohibition against ending sentences in prepositions.
As linguists actually started inventorying world languages, it became more and more clear that there is a very wide dichotomy of grammatical features that don't necessarily translate well to familiar languages. There are vanishingly few features that are actually universal to all languages--the noun may well be the only universal part of speech. That a language doesn't choose to mark a feature in a particular way doesn't make it more primitive than another language. English doesn't have a numerical classifier... is it more primitive than an Australian Aboriginal language? Or is it more primitive than Japanese for not having a way to mark register (~ politeness)?
(FWIW, Linear B is used to write Mycenaean Greek, and this has been known for ~70 years.)
Re: Minoan language Linear A linked to Linear B in new research
#46Earlier quoted context omitted.
can't they use the zipf law ? https://en.wikipedia.org/wiki/Zipf%27s_law The decreasing exponential law should allow to find "the" and some closed form POS words, so yeah determiners, prepositions and conjonctions.
I'm skeptical of the claim that this would work at all even if there was a larger corpus. Let's say you had a million pages of classical Chinese text, but absolutely no context about what the text meant or was about. By looking at it closely and using statistical analysis you could certainly determine various rules of the grammar, and you might even be able to guess that certain characters are grammatical constructio…
Re: Minoan language Linear A linked to Linear B in new research
#47Earlier quoted context omitted.
a language without thoses would be hella weird and primitive, like stereotypical robotic talking. To answer you question, I don't know, do linear B have them?
> a language without thoses would be hella weird and primitive, like stereotypical robotic talking. To answer you question, I don't know, do linear B have them? I think there are very few assumptions of the form " every reasonable language has […]" that hold up even for all current languages, let alone historical ones.
Re: Minoan language Linear A linked to Linear B in new research
#48Earlier quoted context omitted.
I think people (here and below) are getting hung up on definite articles, but Zipf's Law makes no such observation. It says only that a word's frequency in a natural language corpus tends to be in inverse proportion to its rank in a frequency table. In English, the most frequent words are articles, but the general observation about word frequency holds across languages (whether those languages have articles or not).
"The most frequently appearing words in this pile of un-translateable text are the most common words in the language it is written in" seems like it falls somewhere between blindingly obvious, and entirely useless. Unless you have some clue what those words mean, how does that observation help you?
Re: Minoan language Linear A linked to Linear B in new research
#49Earlier quoted context omitted.
I'm skeptical of the claim that this would work at all even if there was a larger corpus. Let's say you had a million pages of classical Chinese text, but absolutely no context about what the text meant or was about. By looking at it closely and using statistical analysis you could certainly determine various rules of the grammar, and you might even be able to guess that certain characters are grammatical constructio…
My guess is that if you had a really big, wide-ranging and high-quality corpus of a completely unknown human language then you probably would be able to decipher and translate it. If you could deduce or guess the grammatical structure the next step might be to look at which nouns can be subjects of which verbs, for example, and it might then be possible to guess which nouns refer to humans and which verbs describe ac…
Re: Minoan language Linear A linked to Linear B in new research
#50Earlier quoted context omitted.
Can we assume that any of those features would be part of this language?
a language without thoses would be hella weird and primitive, like stereotypical robotic talking. To answer you question, I don't know, do linear B have them?
The only thing "robotic" about it is the fact that "robot" is a Czech word that was adopted worldwide.