Earlier quoted context omitted.
OK, fair enough, but can you find one in this article that's hard for an LLM? The gnarliest one I saw, 4o handled instantly, and I went back and looked carefully at the image and the text and I'm sold. Like if this is a crowdsourcing project, why not do a first pass with an LLM and present users with both the image and the best-effort LLM pass? Later I signed up, went to the current missions, and they all seem to pos…
My parents have saved letters from their parents which are written in cursive but in two perpendicular layers. Meaning the writing goes horizontally in rows and then when they got to the end of the page it was turned 90 degrees and continued right on top of what was already there for the whole page. This was apparently to save paper and postage. It looks like an unintelligible jumble but my mother can actually deciph…
Can you read this cursive handwriting? The National Archives wants your help
81–90 of 267 posts
Re: Can you read this cursive handwriting? The National Archives wants your help
#82Earlier quoted context omitted.
Go ahead and find something hard, and relate back the steps you took to find it.
> Go ahead and find something hard, and relate back the steps you took to find it. This is a strawman[0] argument. You proclaimed: A lot of what they want transcribed is totally straightforward to OCR And I replied: If it's that easy, then do it and be the hero they want. So do it or do not. Nowhere does my finding "something hard" have any relevance to your proclamation. 0 - https://en.wikipedia.org/wiki/Straw_man
(I don't think you need to Wikipedia-cite "straw man" on HN).
Re: Can you read this cursive handwriting? The National Archives wants your help
#83Earlier quoted context omitted.
There are conceivable reasons why they may be telling a half truth here. Just engaging the public is a worthy goal here.
> There are conceivable reasons why they may be telling a half truth here. Just engaging the public is a worthy goal here. Asserting an ulterior motive without supporting proof is to engage in conspiracy theories. Sometimes a cigar is just a cigar.[0] 0 - https://quoteinvestigator.com/2011/08/12/just-a-cigar/
Re: Can you read this cursive handwriting? The National Archives wants your help
#84Earlier quoted context omitted.
Go ahead and find something hard, and relate back the steps you took to find it.
> Go ahead and find something hard, and relate back the steps you took to find it. This is a strawman[0] argument. You proclaimed: A lot of what they want transcribed is totally straightforward to OCR And I replied: If it's that easy, then do it and be the hero they want. So do it or do not. Nowhere does my finding "something hard" have any relevance to your proclamation. 0 - https://en.wikipedia.org/wiki/Straw_man
That's not a claim that processing the entire archive would be trivial. And even if it was, whether that would make someone the "hero they want" is part of what's being called into question.
So your silly demand going unmet proves nothing.
Also, "give me an example please" is not a strawman!
If you actually want to prove something, you need to show at least one document in the set that a human can do but not a machine, or to really make a good point you need to show that a non-neglibile fraction fit that description.
Re: Can you read this cursive handwriting? The National Archives wants your help
#85* Article doesn't provide a direct link to the topic mission
* Signup is pretty easy. Well organized and even gently requires you to have two forms of 2FA.
* Sign up complete. Go back to the primary page and try to find the mission. A little buried but not too deep.
* Notice I'm not signed in. Ok, let's do that. Now I'm back on the main page and navigate back. Find the first document and open it. Really interesting to scan through the doc and to read. People back then generally had really nice handwriting.
* Ok, what next, how do I transcribe? ... ? Oh it says I'm not logged in again. Fine, click the link and...
* I'm logged in and directed back to the main page, again.
Look, this is an interesting project and I'd love to spend my spare cycles to help out. But they really need to clean up this process.
Volunteers shouldn't have to jump through kinda poorly designed interfaces to help out.
Re: Can you read this cursive handwriting? The National Archives wants your help
#86Earlier quoted context omitted.
> Go ahead and find something hard, and relate back the steps you took to find it. This is a strawman[0] argument. You proclaimed: A lot of what they want transcribed is totally straightforward to OCR And I replied: If it's that easy, then do it and be the hero they want. So do it or do not. Nowhere does my finding "something hard" have any relevance to your proclamation. 0 - https://en.wikipedia.org/wiki/Straw_man
I did in fact do it, and what I got was much, much easier than the samples in the article, which 4o did fine with. I'm sorry, but I declare the burden of proof here to be switched. Can you find a hard one? (I don't think you need to Wikipedia-cite "straw man" on HN).
Awesome.
Can you guarantee its results are completely accurate every time, with every document, and need no human review?
> I'm sorry, but I declare the burden of proof here to be switched.
If you are referencing my stating:
If it's that easy, then do it and be the hero they want.
Then I don't really know how to respond. Otherwise, if you are referencing my statement:> Perhaps "random humans" can perform tasks which could reshape your belief:
>> OCR is VERY good
To which I again ask, can you guarantee the correctness of OCR results will exceed what "random humans" can generally provide? What about "non-random motivated humans"?
My point is that automated approaches to tasks such as what the National Archives have outlined here almost always require human review/approval, as accuracy is paramount.
> (I don't think you need to Wikipedia-cite "straw man" on HN).
I do so for two purposes. First, if I misuse a cited term someone here will quickly correct me. Second, there is always a probability of someone new here which is unaware of the cited term(s).
Re: Can you read this cursive handwriting? The National Archives wants your help
#87Earlier quoted context omitted.
I did in fact do it, and what I got was much, much easier than the samples in the article, which 4o did fine with. I'm sorry, but I declare the burden of proof here to be switched. Can you find a hard one? (I don't think you need to Wikipedia-cite "straw man" on HN).
> I did in fact do it, and what I got was much, much easier than the samples in the article, which 4o did fine with. Awesome. Can you guarantee its results are completely accurate every time, with every document, and need no human review? > I'm sorry, but I declare the burden of proof here to be switched. If you are referencing my stating: If it's that easy, then do it and be the hero they want. Then I don't really k…
> > If it's that easy, then do it and be the hero they want.
> Then I don't really know how to respond.
If someone says a thing is easy, and you respond by demanding they do it a million times to prove that it's easy, you are the one that has screwed up the burden of proof.
Re: Can you read this cursive handwriting? The National Archives wants your help
#88Earlier quoted context omitted.
Oh, ok then.
I mean, all you have to do is feed the image to ChatGPT, and it will read it basically as well as you can. Denying/downvoting reality is always an option, of course.
And BugsJustFindMe can't downvote you, because it was a reply to him. So don't bite his head off over it. You got downvoted because you were a jerk, plain and simple.
Re: Can you read this cursive handwriting? The National Archives wants your help
#89Earlier quoted context omitted.
I did in fact do it, and what I got was much, much easier than the samples in the article, which 4o did fine with. I'm sorry, but I declare the burden of proof here to be switched. Can you find a hard one? (I don't think you need to Wikipedia-cite "straw man" on HN).
> I did in fact do it, and what I got was much, much easier than the samples in the article, which 4o did fine with. Awesome. Can you guarantee its results are completely accurate every time, with every document, and need no human review? > I'm sorry, but I declare the burden of proof here to be switched. If you are referencing my stating: If it's that easy, then do it and be the hero they want. Then I don't really k…
Re: Can you read this cursive handwriting? The National Archives wants your help
#90Earlier quoted context omitted.
> Go ahead and find something hard, and relate back the steps you took to find it. This is a strawman[0] argument. You proclaimed: A lot of what they want transcribed is totally straightforward to OCR And I replied: If it's that easy, then do it and be the hero they want. So do it or do not. Nowhere does my finding "something hard" have any relevance to your proclamation. 0 - https://en.wikipedia.org/wiki/Straw_man
There are two claims. The main one is that all of these documents are easy to individually transcribe by machine. The other is that a whole lot can be OCR'd, which is pretty simple to check. That's not a claim that processing the entire archive would be trivial. And even if it was, whether that would make someone the "hero they want" is part of what's being called into question. So your silly demand going unmet prove…
I made demands of no one.
> Also, "give me an example please" is not a strawman!
My identification of the strawman was that it referenced "find something hard" when I had said "be the hero they want" and that what is needed in this specific problem domain may be more difficult than what a generalization addresses.
> If you actually want to prove something, you need to show at least one document in the set that a human can do but not a machine, or to really make a good point you need to show that a non-neglibile fraction fit that description.
Maybe this is the proof you demand.
LLM's are statistical prediction algorithms. As such, they are nondeterministic and, therefore, provide no guarantees as to the correctness of their output.
The National Archives have specific artifacts requiring precise textual data extraction.
Use of nondeterministic tools known to produce provably incorrect results eliminate their applicability in this workflow due to all of their output requiring human review. This is an unnecessary step and can be eliminated by the human reading the original text themself.
Does that satisfy your demand?