A lot of loonies make this claim, but Tom's work is credible enough that it's being reviewed by linguistics experts at Rutgers and Cambridge. Additional validation: his approach produces results. He's translated over 300 words, and that's never been done before, and his solution actually solves some problems in Linear B. Tom is an AI engineer, and Claude Code was key to his work. Disclosures: I know Tom socially, and…
You know him socially but is there a reason you’re writing this rather than him? It looks like he has his own web presence. Cynical read would be you’re stealing his thunder a bit by prematurely announcing this before it’s fully confirmed
AI Engineer Claims to Have Cracked Linear A
31–40 of 195 posts
Re: AI Engineer Claims to Have Cracked Linear A
#32Earlier quoted context omitted.
You can use Claude, like the author, to reproduce the result.
This isn't really a reasonable approach, is it? The original prompts aren't provided, nor is the original context; even then, you can't really treat a stochastic system like an LLM as a major component in reproducibility.
Re: AI Engineer Claims to Have Cracked Linear A
#33Earlier quoted context omitted.
You know him socially but is there a reason you’re writing this rather than him? It looks like he has his own web presence. Cynical read would be you’re stealing his thunder a bit by prematurely announcing this before it’s fully confirmed
Isn't it customary for the author of a post shared on HN to leave a comment on the thread?
Re: AI Engineer Claims to Have Cracked Linear A
#34Re: AI Engineer Claims to Have Cracked Linear A
#35If you have a 4k screen, you can fit all remaining Linear A text on your screen at once, in 14pt high font.
Re: AI Engineer Claims to Have Cracked Linear A
#36however, nawaya or what ever examples around it are not part of the Hebrew language.
Re: AI Engineer Claims to Have Cracked Linear A
#37Earlier quoted context omitted.
You know him socially but is there a reason you’re writing this rather than him? It looks like he has his own web presence. Cynical read would be you’re stealing his thunder a bit by prematurely announcing this before it’s fully confirmed
What thunder? Claude did the work and used a human to interface with experience and causality better.
One of the things I find weird with AI is how the dismissals of work that involve AI splits into two camps: like yours, saying the AI did the work while the human played no role and deserves no credit; and those saying the AI rips off its training data while the human using it played no role and deserves no credit.
Re: AI Engineer Claims to Have Cracked Linear A
#38- The "Libation Formula", which the author used as the base for his translations, is the most studied piece of writing in Linear A, because it's the only recurring phrase (with grammatical variation) that we have. The corpus is extremely fragmentary, with just a handful of instances of longer text (and even then, the texts are the length of an average sentence in English). The majority of documents available to us are lists (of inventory, personnel, offerings or something of this sort). The longer texts make use of punctuation marks, likely put in between words. This gives us a non-trivial vocabulary, which still does not match that of any known language.
- With such fragmentary remaining material, we cannot be sure that a) all the texts we call "Linear A" are written in the same language, and b) the recognizable words are not abbreviations, for example.
- The author made an assumption that Linear A symbols which have counterparts in Linear B should have the same phonetic values. This gives us an already known glyph that represented "NA". "Duplicate" glyphs are only found in the P-series, and are assumed to represent syllables which were distinguished by the Linear A language, but not by Greek - such as aspirated/unaspirated P. There is a glyph that stands for "NWA" in Linear B, but instances of it have been found in Linear A as well.
- There are countless words with no known etymology in Ancient Greek, assumed to originate from a substrate language or languages spoken in the area at the time Greeks migrated to their present-day homeland. The language of Linear A would be a likely candidate for such substrate. If Linear A were a Semitic language, then we should already be able to establish Semitic etymologies for those words as they were in Greek. Of course it could also be the case that these words came from an another language which did not adopt writing or its writing did not survive to our times.
Re: AI Engineer Claims to Have Cracked Linear A
#39Sorry but I don’t recognize this as being an achievement by an amateur. This dude had no chance in hell until we trained a model to use his time to suss it out.
Re: AI Engineer Claims to Have Cracked Linear A
#40Is this extendible to a generalizable approach to translate any language pair (without a translation map or translation dataset)?
It's a common misconception that is what happened with Ancient Egyptian with the Rosetta Stone. The Rosetta Stone was just one of the big pieces of the puzzle. The decoding came when people realized that Coptic (a language written alphabetically and still in use in the Coptic Church today) is actually descended from Ancient Egyptian; as Spanish is to Latin, Coptic is to Ancient Egyptian.
Similarly the attempts to decode classical Maya were all dead ends. Until Yuri Knorozov realized that it encoded the ancestor of the Maya languages which are still spoken to this day. (Knorozov's Wikipedia article is worth checking out just for his photo with his cat. [0] IMHO.)
I have written before about the La Mojarra 1 stele in Mexico [1]. It looks a lot like Maya. [2] But it isn't Maya. Maybe the difference like between Russian and Latin writing?
No one can read it. It's undecipherable. There are some attempts to identify it with a proposed ancient language that would have been related to the modern Mixe-Zoque languages: some of the glyphs that are shared with Maya, when read phonetically, start sounding like a Mixe-Zoque language. But no one has proposed a confident decipherment. There probably isn't enough text. La Mojarra 1 is the only long example of the Isthmian script.
Deciphering Akkadian was very difficult, at first. The process started with Persian; old Persian was written in a simplified adapted form of the Mesopotamian cuneiform (wedges on clay). It was a kind of alphabet. And Old Persian was already understood. And there was a bilingual text on a monument carved by Darius I. But even then -- decoding relies so heavily on the fact that Akkadian is a Semitic language distantly related to Hebrew, more distantly, also Ancient Egyptian. So again, we sort of knew what we were looking for.
That is all to say: even if the Voynich manuscript (for example) contains real text in an otherwise completely lost language, I'm not sure it is possible even theoretically to translate it.
[0] https://en.wikipedia.org/wiki/Yuri_Knorozov
[1] https://en.wikipedia.org/wiki/La_Mojarra_Stela_1
[2] https://commons.wikimedia.org/wiki/File:La_Mojarra_Stela_1_S...