Live data from Hacker News

Chat-based Large Language Models replicate the mechanisms of a psychic’s con

softwarecrisis.dev

1–10 of 14 posts

Re: Chat-based Large Language Models replicate the mechanisms of a psychic’s con

#2
As someone who admittedly belongs more to the "AI believer" side, I find the vagueness of the training data increasingly frustrating.

The thing that impressed me most about LLMs so far is less the factual correctness or incorrectness of its output but the fact that it appears (!) to understand the instructions that are given. I.e., even if you give it an improbable and outlandish task ("write a poem about kernel debugging in the style of Edgar Allan Poe", "write a script for a Star Trek TNG episode in which Picard only talks in curse words "), it always gives a response which is a valid fulfillment of the task.

Of course it could be that the tasks weren't really as outlandish as they seemed and somewhere in the vast amounts of training data there was already a matching TNG fanfic which just needed some slight adjustments or something.

But those kinds of arguments essentially shift the black box from the model to the training data: Instead of claiming the model has magical powers of intelligence, now the training data magically already contains anything you could possibly ask for. I personally don't find that approach that much more rational that believing in some kind of AI consciousness (or fragments of it).

...but of course it could be. This is why I'd wish for foundation models with more controlled training data, so we can make more certain statements about which responses could be reasonably be pulled from the training data and which would be truly novel.

Re: Chat-based Large Language Models replicate the mechanisms of a psychic’s con

#3
Oops, another blogger falling into the trap of not specifying how they define "intelligence" and then making a "no true scotsman" argument against their loose pre-existing beliefs.

If you're thinking about writing an article like this, please just define what you think intelligence is right at the top. That's the entirety of the discussion, the rest is fluff.

Also, as a society we need to minimize the amount of attention we give to debates over definitions. Once a discussion or political debate is reduced to a definitional issue, everyone starts talking past each other and forgets what the argument even is. (See discussions about the definitions of "life", "woman", "socialism", "capitalism", etc.) Words are lossy proxies to ideas, and they only matter insofar as they allow us to understand one another.

Re: Chat-based Large Language Models replicate the mechanisms of a psychic’s con

#4
In a parallel thread hardcore scientists struggle to understand a 100 neuron worm, while here less hardcore scientists proclaim they've understood nuances of a human brain.

Note, that there is a rapid rise of the "mechanical consciousness" dogma. Some very smart individuals are so impressed by it that rather than doubting the existence of intelligence in LLM AI, they've started thinking that they themselves might be machines! From there it's one step to giving machines rights on par with humans. The dogma is very powerful.

Re: Chat-based Large Language Models replicate the mechanisms of a psychic’s con

#5
To be honest, there have been occasions where I've been failing miserably for some time to come up with solutions to some sysadmin nature problems with my own systems and I've taken my problem straight to GPT-4 with a clear explanation of the issue alongside with any diagnostics I thought would make sense.

To my very surprise, GPT-4 did an astonishing job reasoning on the specifics of my system (aarch64 exotic setup, alpinelinux on asahi) and came back with a very specific on point list of suggestions which included the very solution as #1.

I've had it hold my hand many times while navigating relatively complex and niche systems like android smartphones with custom partitioning schemes booting linux and what have you and yes, it still was still, very reasonable, to say the least.

So to conclude, it has the ability to reason properly for systems and situations that it's not necessary trained in and displays the ability for coherent reasoning on specifics over things which at least for September 2021 were relatively unknown. I'm really wondering how far this thing needs to get in order for some people to admit its more than a model spitting a token next to another, or some type of mentalist doing an excellent job in hypnotising most of us into thinking it already displays incredible intelligence but its smoke and mirrors.

Re: Chat-based Large Language Models replicate the mechanisms of a psychic’s con

#6
post #2

As someone who admittedly belongs more to the "AI believer" side, I find the vagueness of the training data increasingly frustrating. The thing that impressed me most about LLMs so far is less the factual correctness or incorrectness of its output but the fact that it appears (!) to understand the instructions that are given. I.e., even if you give it an improbable and outlandish task ("write a poem about kernel debu…

Intelligence need not be magical.

I suspect when it comes to training data, it may need to be general enough to allow the architecture a chance to learn the meta concept of "learning". Ie identify the latent gestalt within a text corpus that we most identify as "reasoning ability".

If the training data is not rich enough, then these more refined emergent abilities will not be discovered through our current algorithms/architecture. Maybe in the future when more efficient algorithms are found (we know the lower bound must be at least as efficient as our human brains for example) then we won't need as much/as rich data. Or use Multi modal data.

From what we're seeing I believe we can already discount the tainted training data as likely hypothesis, and trend to the suspicion there is something deeper at play.

For instance, what if LLMs through pattern recognition of text alone may have built a coherent enough world model that it yields answers indistinguishable from human intelligence?

Nothing about that seems improbable from current neuroscience theories https://en.m.wikipedia.org/wiki/Predictive_coding

It may also suggest there to be nothing special functionally about the human brain; the ability for a system to recursively identify, model, and remix concepts may be sufficient to give rise to the phenomenology we know as intelligence.

Qualia, goals, "feelings", that sounds more nebulous and complicated to define and assess though.

Re: Chat-based Large Language Models replicate the mechanisms of a psychic’s con

#7
post #3

Oops, another blogger falling into the trap of not specifying how they define "intelligence" and then making a "no true scotsman" argument against their loose pre-existing beliefs. If you're thinking about writing an article like this, please just define what you think intelligence is right at the top. That's the entirety of the discussion, the rest is fluff. Also, as a society we need to minimize the amount of atten…

They need to go beyond defining intelligence. They need to operationalize intelligence.

The problems still are many. First, we've already operationalized intelligence, namely through IQ tests see Stanford-Binet Intelligence Scale, Universal Nonverbal Intelligence, Differential Ability Scales, Peabody Individual Achievement Test, Wechsler Individual Achievement Test, Wechsler Adult Intelligence Scale, &c. Even so, HN users will both point out the (in)validity of these but also the (in)validity with respect to applying them to AI.

The real problem is that intelligence is a socially defined phenomenon, as opposed to an essential metaphysical property. If we admit that, many of the definition foibles we have become irrelevant.

Re: Chat-based Large Language Models replicate the mechanisms of a psychic’s con

#8
post #4

In a parallel thread hardcore scientists struggle to understand a 100 neuron worm, while here less hardcore scientists proclaim they've understood nuances of a human brain. Note, that there is a rapid rise of the "mechanical consciousness" dogma. Some very smart individuals are so impressed by it that rather than doubting the existence of intelligence in LLM AI, they've started thinking that they themselves might be…

Scientism is a hell of a drug.

Re: Chat-based Large Language Models replicate the mechanisms of a psychic’s con

#9
post #5

To be honest, there have been occasions where I've been failing miserably for some time to come up with solutions to some sysadmin nature problems with my own systems and I've taken my problem straight to GPT-4 with a clear explanation of the issue alongside with any diagnostics I thought would make sense. To my very surprise, GPT-4 did an astonishing job reasoning on the specifics of my system (aarch64 exotic setup,…

Can you share the log?

Re: Chat-based Large Language Models replicate the mechanisms of a psychic’s con

#10
> There are two possible explanations for this effect:

> 1. The tech industry has accidentally invented the initial stages a completely new kind of mind, based on completely unknown principles, using completely unknown processes that have no parallel in the biological world.

> 2. The intelligence illusion is in the mind of the user and not in the LLM itself.

Great, now write 10k words more, but this time about the psychology of your unwillingness to change from (2) to (1) when the facts changed.

Post reply on HN