Live data from Hacker News

A new Google model is nearly perfect on automated handwriting recognition

generativehistory.substack.com

211–220 of 328 posts

Re: A new Google model is nearly perfect on automated handwriting recognition

#211
post #88

Earlier quoted context omitted.

This is quite a narrow view of how the generation works. AI can extrapolate from the training set and explore new directions. It's not just cutting pieces and gluing together.

Calling it “exploring” is anthropomorphising. The machine has weights that yield meaningful programs given specification-like language. It’s a useful phenomenon but it may be nothing like what we do.

Or it may be remarkably similar to what we do

Re: A new Google model is nearly perfect on automated handwriting recognition

#212

Earlier quoted context omitted.

You are right to be skeptical. There are plenty of so called windows(or other) web 'os' clones. There were a couple of these posted on HN actually this very year. Here is one example I google dthat was also on HN : https://news.ycombinator.com/item?id=44088777 This is not an OS as in emulating a kernel in javascript or wasm, this is making a web app that looks like the desktop of an OS. I have seen plenty such projec…

Every time a model is about to be released, there are a bunch of these hype accounts that spin up. I don't know they get paid or they spring up organically to farm engagement. Last time there was such hype for a model was "strawberry" (o1) then gpt-5, and both turned out to be meaningful improvements but nowhere near the hype. I don't doubt though that new models will be very good at frontend webdev. In fact this is…

My guess is that there are insiders who know about the models and can’t keep their mouths shut. They like being on the inside and leaking.

Re: A new Google model is nearly perfect on automated handwriting recognition

#213

Earlier quoted context omitted.

I'm skeptical that they're actually capable of making something novel. There are thousands of hobby operating systems and video game emulators on github for it to train off of so it's not particularly surprising that it can copy somebody else's homework.

I remain confused but still somewhat interested as to a definition of "novel", given how often this idea is wielded in the AI context. How is everyone so good at identifying "novel"? For example, I can't wrap my head around how a) a human could come up with a piece of writing that inarguably reads "novel" writing, while b) an AI could be guaranteed to not be able to do the same, under the same standard.

[deleted]

Re: A new Google model is nearly perfect on automated handwriting recognition

#214

Earlier quoted context omitted.

Positively not. It is pure interpolation and not extrapolation. The training set is vast and supports an even vaster set of possible traversal paths; but they are all interpolative. Same with diffusion and everything else. It is not extrapolation that you can transfer the style of Van Gogh onto a photographl it is interpolation. Extrapolation might be something like inventing a style: how did Van Gogh do that? And, s…

This is knock against you at all , but in a naive attempt to spare someone else some time: remember that based on this definition it is impossible for an LLM to do novel things and more importantly , you're not going to change how this person defines a concept as integral to one's being as novelty. I personally think this is a bit tautological of a definition, but if you hold it, then yes LLMs are not capable of anyt…

I think you should reverse the question, why would we expect LLMs to even have the ability to do novel things?

It is like expecting a DJ remixing tracks to output original music. Confusing that the DJ is not actually playing the instruments on the recorded music so they can't do something new beyond the interpolation. I love DJ sets but it wouldn't be fair to the DJ to expect them to know how to play the sitar because they open the set with a sitar sample interpolated with a kick drum.

Re: A new Google model is nearly perfect on automated handwriting recognition

#215

Earlier quoted context omitted.

You have to realize AI is trained the same way one would train an auto-completer. Theres no cognition. It’s not taught language, grammar, etc. none of that! It’s only seen a huge amount of text that allows it to recognize answers to questions. Unfortunately, it appears to work so people see it as the equivalent to sci-fi movie AI. It’s really just a search engine.

no it's not I work on AI and what these things do are much much more then a search engine or an autocomplete. If an autocomplete passed the turing test you'd dismiss it because it's still an autocomplete. The characterization you are regurgitating here is from laymen who do not understand AI. You are not just mildly wrong but wildly uninformed.

To be fair, it's not clear human intelligence is much more than search or autocomplete. The only thing that's clear here is that LLMs can't reproduce it.

Re: A new Google model is nearly perfect on automated handwriting recognition

#216

Earlier quoted context omitted.

no it's not I work on AI and what these things do are much much more then a search engine or an autocomplete. If an autocomplete passed the turing test you'd dismiss it because it's still an autocomplete. The characterization you are regurgitating here is from laymen who do not understand AI. You are not just mildly wrong but wildly uninformed.

To be fair, it's not clear human intelligence is much more than search or autocomplete. The only thing that's clear here is that LLMs can't reproduce it.

Yes but colloquially this characterization you see used by laymen is deliberately used to deride AI and dismiss it. It is not honest about the on the ground progress AI has made and it’s not intellectual honest about the capabilities and weaknesses of Ai.

Re: A new Google model is nearly perfect on automated handwriting recognition

#217

> In tabulating the “errors” I saw the most astounding result I have ever seen from an LLM, one that made the hair stand up on the back of my neck. Reading through the text, I saw that Gemini had transcribed a line as “To 1 loff Sugar 14 lb 5 oz @ 1/4 0 19 1”. If you look at the actual document, you’ll see that what is actually written on that line is the following: “To 1 loff Sugar 145 @ 1/4 0 19 1”. For those unawa…

[flagged]

The comment above seems to violate several HN guidelines. Curious, I asked GPT and Gemini which ones stood out. Both replied with the same top three:

https://news.ycombinator.com/newsguidelines.html

They are:

1. “Be kind. Don't be snarky. … Edit out swipes.”

2. “Please don't sneer, including at the rest of the community.”

3. “Please don't post shallow dismissals, especially of other people's work. A good critical comment teaches us something.”

Re: A new Google model is nearly perfect on automated handwriting recognition

#218

Earlier quoted context omitted.

You have to realize AI is trained the same way one would train an auto-completer. Theres no cognition. It’s not taught language, grammar, etc. none of that! It’s only seen a huge amount of text that allows it to recognize answers to questions. Unfortunately, it appears to work so people see it as the equivalent to sci-fi movie AI. It’s really just a search engine.

no it's not I work on AI and what these things do are much much more then a search engine or an autocomplete. If an autocomplete passed the turing test you'd dismiss it because it's still an autocomplete. The characterization you are regurgitating here is from laymen who do not understand AI. You are not just mildly wrong but wildly uninformed.

Well, I also work on AI, and I completely agree with you. But I've reached the point of thinking it's hopeless to argue with people about this: It seems that as LLMs become ever better people aren't going to change their opinions, as I had expected. If you don't have good awareness of how human cognition actually works, then it's not evidently contradictory to think that even a superintelligent LLM trained on all human knowledge is just pattern matching and that humans are not. Creativity, understanding, originality, intent, etc, can all be placed into a largely self-consistent framework of human specialness.

Re: A new Google model is nearly perfect on automated handwriting recognition

#219

I really hope they have because I’ve also been experimenting with LLMs to automate searching through old archival handwritten documents. I’m interested in the Conquistadors and their extensive accounts of their expeditions, but holy cow reading 16th century handwritten Spanish and translating it at the same time is a nightmare, requiring a ton of expertise and inside field knowledge. It doesn’t help that they were of…

Where can I find these Conquistador documents? Sounds like something I might like to read and explore.

Re: A new Google model is nearly perfect on automated handwriting recognition

#220

Earlier quoted context omitted.

To be fair, it's not clear human intelligence is much more than search or autocomplete. The only thing that's clear here is that LLMs can't reproduce it.

Yes but colloquially this characterization you see used by laymen is deliberately used to deride AI and dismiss it. It is not honest about the on the ground progress AI has made and it’s not intellectual honest about the capabilities and weaknesses of Ai.

I disagree. The actual capabilities of LLMs remain unclear, and there's a great deal of reasons to be suspicious of anyone whose paycheck relies on pimping them.
Post reply on HN