It would be interesting to actively track how far long each progressive model gets...
The case for zero-error horizons in trustworthy LLMs
11–20 of 121 posts
Re: The case for zero-error horizons in trustworthy LLMs
#12Re: The case for zero-error horizons in trustworthy LLMs
#13People are going to misinterpret this and overgeneralize the claim. This does not say that AI isn't reliable for things. It provides a method for quantifying the reliability for specific tasks. You wouldn't say that a human who doesn't know how to read isn't reliable in everything, just in reading. Counting is something that even humans need to learn how to do. Toddlers also don't understand quantity. If a 2 year old…
I completely agree with you. LLMs are regurgitation machines with less intellect than a toddler, you nailed it.
AI is here!
Re: The case for zero-error horizons in trustworthy LLMs
#14People are going to misinterpret this and overgeneralize the claim. This does not say that AI isn't reliable for things. It provides a method for quantifying the reliability for specific tasks. You wouldn't say that a human who doesn't know how to read isn't reliable in everything, just in reading. Counting is something that even humans need to learn how to do. Toddlers also don't understand quantity. If a 2 year old…
No human who can program, solve advanced math problems, or can talk about advanced problem domains at expert level, however, would fail to count to 5.
This is not a mere "LLMs, like humans, also need to be taught this" but points to a fundamental mismatch about how humans and LLMs learn.
(And even if they merely needed to be taught, why would their huge corpus fail to cover that "teaching", but cover way more advanced topics in math solving and other domains?)
Re: The case for zero-error horizons in trustworthy LLMs
#15Re: The case for zero-error horizons in trustworthy LLMs
#16Whenveer I see these papers and try them, they always work. This paper is two months old, which in LLM years is like 10 years of progress. It would be interesting to actively track how far long each progressive model gets...
Re: The case for zero-error horizons in trustworthy LLMs
#17> This is surprising given the excellent capabilities of GPT-5.2 The real surprise is that someone writing a paper on LLMs doesn't understand the baseline capabilities of a hallucinatory text generator (with tool use disabled).
Re: The case for zero-error horizons in trustworthy LLMs
#18Whenveer I see these papers and try them, they always work. This paper is two months old, which in LLM years is like 10 years of progress. It would be interesting to actively track how far long each progressive model gets...
Is tokenization extremely efficient? Yes. Does it fundamentally break character-level understanding? Also yes. The only fix is endless memorization.
Re: The case for zero-error horizons in trustworthy LLMs
#19Whenveer I see these papers and try them, they always work. This paper is two months old, which in LLM years is like 10 years of progress. It would be interesting to actively track how far long each progressive model gets...
So yes.
And the valuations. Trillion dollar grifter industry.
Re: The case for zero-error horizons in trustworthy LLMs
#20This is well known and not that interesting to me - ask the model to use python to solve any of these questions and it will get it right every time.