Live data from Hacker News

An entire Herculaneum scroll has been read for the first time

scrollprize.org

41–50 of 404 posts

Re: An entire Herculaneum scroll has been read for the first time

#41

When I read translations like these, I always wonder if the tone is translated. Did the writer mean to convey a very formal “to the utmost”, or was it a more casual “to the max”. How much of the translators bias makes these seem like academic papers instead of social media posts.

[dead]

Re: An entire Herculaneum scroll has been read for the first time

#42

Earlier quoted context omitted.

That's a tough one to give a strong estimate of. Some scrolls are easier or harder to unwrap and read for a multitude of different reasons, mostly due to how damaged the scroll was in the eruption, and how easy or not the ink is to read. IIRC from what we've scanned of the herculaneum collection, none of the ink is easily visible via spectrum alone, so we have to use a lot of ML and physically based rendering techniq…

Do you think this particular scroll is easier or harder to read that the others will be? Or about average?

Pherc1667 was quite small and just so happened to have readable ink, so it was easier than I expect most others to be.

Re: An entire Herculaneum scroll has been read for the first time

#43
post #5

I am on the vesuvius challenge team that did the segmentation, unwrapping, and ink detection, so feel free to ask any questions.

What are the wildest, most exciting but plausible things that might be discovered in these documents?

Here's a list. The scrolls are from a library that burned in 79 AD.

https://en.wikipedia.org/wiki/List_of_lost_literary_works

Re: An entire Herculaneum scroll has been read for the first time

#44

When I read translations like these, I always wonder if the tone is translated. Did the writer mean to convey a very formal “to the utmost”, or was it a more casual “to the max”. How much of the translators bias makes these seem like academic papers instead of social media posts.

Any useful translation of an ancient text is accompanied by the text in the original language, so that the reader may assess how faithful is the translation.

For anyone who wants to read ancient texts, there are bilingual editions, for example those of the "Loeb library".

The translations that omit the original text are just for the people who want to have some idea about the content, but do not care about the correctness of the translation.

With a bilingual edition, it is easy to understand the original text even with relatively little knowledge about the original language.

The original text is important because frequently the translator is forced to introduce inaccuracies in the translation, because of the absence of exact equivalents in the target language, which would require a long explanation of the original meaning, instead of just a translated sentence.

Especially misleading are translations where several distinct ancient words are translated using the same English word, so some nuances are lost.

Equally confusing are the cases when the translator chooses to translate the same ancient word by different English words, because even if the meaning of a word may depend on the context, many translators fail to judge correctly the context, because they may lack specialized knowledge so their guesses are not necessarily better than of the readers who may be less competent in linguistics, but more competent in the science or technology needed to understand the context. Better translators prefer to use a one-to-one mapping between words, which makes it easier for the readers to discover the meaning intended by the ancient writer, after seeing multiple examples of usage.

Re: An entire Herculaneum scroll has been read for the first time

#45
post #5

Earlier quoted context omitted.

What are the wildest, most exciting but plausible things that might be discovered in these documents?

Probably a lot more texts of Epicurean philosophy and not a whole lot else unfortunately according to my papyrologist friend.

Why would Epicurean philosophy be unfortunate?

I was under the impression that there was almost nothing left of that school of thought, and that it’s writings had been destroyed.

What would you like to have instead?

Re: An entire Herculaneum scroll has been read for the first time

#47

Earlier quoted context omitted.

Given the current rate of progress, how long do you think it will take to decipher the entire collection?

That's a tough one to give a strong estimate of. Some scrolls are easier or harder to unwrap and read for a multitude of different reasons, mostly due to how damaged the scroll was in the eruption, and how easy or not the ink is to read. IIRC from what we've scanned of the herculaneum collection, none of the ink is easily visible via spectrum alone, so we have to use a lot of ML and physically based rendering techniq…

Do we known what ink is used?

Re: An entire Herculaneum scroll has been read for the first time

#48

The person who wrote this was was closer in time to the technology that was able to unwind and read burned fragments of their text, than the technology that build the pyramids. pretty wild to think about.

>technology that build the pyramids

You mean ropes and carts?

Re: An entire Herculaneum scroll has been read for the first time

#49
post #39

Earlier quoted context omitted.

Outstanding work! I've participated in the challenge, but didn't get far. One of the questions I had at the time was - if I'm going to use ML to detect ink, could it invent hallucinated letters, or even parts of text, and how to prevent that?

Yes, it's quite possible for ML to hallucinate ink, though it is on a much more local scale, like predicting a slightly longer stroke, filling in more of a character than is actually in the data, etc. Perhaps enough to change a reading of a character or show where ink isnt. It is difficult for ink detection to hallucinate grammatical and idiomatic greek and latin.

What is the input to the ML algorithm? Does it know the surrounding context so that it has a chance to deduce "if this stroke is slightly longer then the end result will be idiomatic greek and latin"?
Post reply on HN