Live data from Hacker News

AI Engineer Claims to Have Cracked Linear A

aiclambake.com

51–60 of 195 posts

Re: AI Engineer Claims to Have Cracked Linear A

#51

Earlier quoted context omitted.

You're absolutely right! We've opened a ticket with the Linear A folks, hopefully they'll get back to us soon with an update as to whether we've got it correct or not. Hang tight!

This comment sure is load bearing.

Regardless, we should stand ready, loaded for bear.

Re: AI Engineer Claims to Have Cracked Linear A

#52

The reason linear A is so difficult is that the total remaining corpus of Linear A text is ~7500 characters, spread out over ~1500 inscriptions. If you have a 4k screen, you can fit all remaining Linear A text on your screen at once, in 14pt high font.

An in addition to that, a vast majority of documents are lists which consist of a "header" (1 to 3 words) and word-number pairs afterwards. An another common class are small clay seals with 1, 2 characters carved into them. It's likely that in both cases, we may be dealing with abbreviations.

Some of the lists end with "ku-ro" and a number that's the sum of all the previous numbers, oddly frequently off by one.

Re: AI Engineer Claims to Have Cracked Linear A

#53
post #27

Earlier quoted context omitted.

This isn't really a reasonable approach, is it? The original prompts aren't provided, nor is the original context; even then, you can't really treat a stochastic system like an LLM as a major component in reproducibility.

> even then, you can't really treat a stochastic system like an LLM as a major component in reproducibility. If you had the other things, being "stochastic" is not even remotely a show-stopper. Stochastic processes abound and are the reason the mathematics of statistics was developed in the first place, ultimately allowing us to create such things as LLMs. When all the relevant steps gets published, I absolutely expe…

My issue with this is that it's a form of "soft" reproducibility, where it'll work for many (maybe even most!) people, but that depends on the way the original prompt was formulated (read on) and the state of the random noise in the system.

On the prompt formulation; prompts with very similar formulations (in terms of both semantics, hamming distance, or both) can lead to _wildly divergent_ outputs in my experience. It's not rigourous, and when that divergence happens, it's extremely difficult (arguably impossible, by nature of the architecture of transformers) to identify why the divergence happened and where.

Re: AI Engineer Claims to Have Cracked Linear A

#54
If confirmed this is really cool and impressive work.

Honestly curious how many years before it can be one shotted in a coding harness with Fable.next by someone who’s not a linguistics expert.

Develop, test, and rank hypotheses about the phonetic values, morphology, grammar, and possible language family of Linear A using the full available corpus. Do not assume any decipherment is correct. Treat all candidate readings as hypotheses to be scored…

Re: AI Engineer Claims to Have Cracked Linear A

#55
post #14

Earlier quoted context omitted.

You can use Claude, like the author, to reproduce the result.

somehow I suspect it was a bit more involved than: Claude, please solve Linear A.

A little bit more. If you ask ChatGPT to "solve linear a" it thinks you mean linear algebra. If you specify that it's the Minoan translation problem, you get a table similar to the one that we get a glimpse of in the without access to the paper, we can't say how much more work the paper has than my gist.

https://gist.github.com/fragmede/bbf277d36a2398065f109484f34...

Re: AI Engineer Claims to Have Cracked Linear A

#57

Earlier quoted context omitted.

You can use Claude, like the author, to reproduce the result.

This isn't really a reasonable approach, is it? The original prompts aren't provided, nor is the original context; even then, you can't really treat a stochastic system like an LLM as a major component in reproducibility.

> stochastic system

Every day when you lower your butt onto your chair, you trust a stochastic system enough to assume you'll rest on the chair safely and not spontaneously phase through, which would lead to rather gory and painful terminal experience.

Physics at macro scale is stochastic, which is a good reminder that stochastic != uniformly random. Expected distributions matter.

Re: AI Engineer Claims to Have Cracked Linear A

#58

The reason linear A is so difficult is that the total remaining corpus of Linear A text is ~7500 characters, spread out over ~1500 inscriptions. If you have a 4k screen, you can fit all remaining Linear A text on your screen at once, in 14pt high font.

An in addition to that, a vast majority of documents are lists which consist of a "header" (1 to 3 words) and word-number pairs afterwards. An another common class are small clay seals with 1, 2 characters carved into them. It's likely that in both cases, we may be dealing with abbreviations. Some of the lists end with "ku-ro" and a number that's the sum of all the previous numbers, oddly frequently off by one.

They hadn't yet decided whether to count from 0 or from 1.

Re: AI Engineer Claims to Have Cracked Linear A

#59

A lot of loonies make this claim, but Tom's work is credible enough that it's being reviewed by linguistics experts at Rutgers and Cambridge. Additional validation: his approach produces results. He's translated over 300 words, and that's never been done before, and his solution actually solves some problems in Linear B. Tom is an AI engineer, and Claude Code was key to his work. Disclosures: I know Tom socially, and…

You know him socially but is there a reason you’re writing this rather than him? It looks like he has his own web presence. Cynical read would be you’re stealing his thunder a bit by prematurely announcing this before it’s fully confirmed

Promoting your friends' work is hardly stealing their thunder. It's increasing their thunder!
Post reply on HN