Live data from Hacker News

What is ChatGPT doing and why does it work?

writings.stephenwolfram.com

121–130 of 518 posts

Re: What is ChatGPT doing and why does it work?

#121

Earlier quoted context omitted.

ChatGPT is version 1, super rough, very broad training, more of a shotgun approach. They already made some quick math improvements. I'm sure if they focused on chess training for example it would be better. I myself am pretty crap at chess and could use some training as well.

> I'm sure if they focused on chess training for example it would be better. That misses the point. Intelligent beings (humans) can learn the rules of any board game given enough time. We don't need special training. What your parent comment says is it can't even learn the rules, let alone be good at it.

Like I said ChatGP has taken the shotgun brute force approach to training. So yea future versions can be trained to play games, just like people are, and future future versions will need less training overall.

The rules of English are ridiculously more complicated than chess, and it has that figured out just fine.

Re: What is ChatGPT doing and why does it work?

#122
It's worth keeping in mind that Stephen Wolfram likely didn't write this himself.

I know people who work at the company, and they sign agreements that any intellectual property (including mathematical proofs) they generate are owned by Stephen Wolfram. Anything Wolfram puts out, like blog posts, scientific articles, and books, are likely to be partly or wholly ghost-written.

Re: What is ChatGPT doing and why does it work?

#123

It's worth keeping in mind that Stephen Wolfram likely didn't write this himself. I know people who work at the company, and they sign agreements that any intellectual property (including mathematical proofs) they generate are owned by Stephen Wolfram. Anything Wolfram puts out, like blog posts, scientific articles, and books, are likely to be partly or wholly ghost-written.

[deleted]

Re: What is ChatGPT doing and why does it work?

#124
post #69

Earlier quoted context omitted.

Realness has little to do with the quality of the output. Or else my TI-83 would be the realest mathematician on the planet, it can multiply numbers that even Terence Tao can't in his head. Judging these systems solely by their output is to repeat the msitakes of behaviorism, even a dumb markov chain or a parrot converses better than an infant, but unlike the infant does not acquire an understanding or representation…

> has little to do with the quality of the output. Non-snarky question: What else can you judge by? Isn't any alternative just putting more precise conditions on the output? With ChatGPT, it's still easy enough to see it's mistakes, and its attempts at fiction and poetry, impressive as they are, are still clumsy to a trained eye, relative to expert human work. But what if they weren't? What happens when they're indis…

Very hard question to really answer. If Einstein said the sentences “are you sure? think about it”, you might spend years pondering what he meant. If GPT does you’ll ignore it. Same output but the context plus content will change what you do.

Same with a painting. If an old master draws a wireframe of a dog, people would bid it up at auction and wonder what he meant. If your kid or AI do, no money might change hands. Same output, different context.

So you can’t just use the output, surely?

Re: What is ChatGPT doing and why does it work?

#125

I am a game programmer but in my spare time I like to learn and get experience in random interesting areas. Eg. recently I learned electronics and Arduinos. Would ye recommend any projects I could do in order to get experience with and learn about this new AI stuff like ChatGPT?

Andrej Karpathy videos

Re: What is ChatGPT doing and why does it work?

#126
post #53

Earlier quoted context omitted.

> 'with enough data its easier to model than we expected' > a self-similar process is ultimately responsible for granting us our own intelligence In my view, intelligence essentially resides within language, specifically in the corpus of language. Both humans and AIs can be effectively colonized by language, as there are innumerable concepts and observations that are transmitted from one mind to another, and now even…

Ludwig Wittgenstein had same idea. but finally he found human experience is more then our language.

Could you elaborate please? For those of us not familiar with Wittgenstein's work, could you link to sources depicting before and after his views changed, preferably with summaries.

Re: What is ChatGPT doing and why does it work?

#127
post #69

Earlier quoted context omitted.

> has little to do with the quality of the output. Non-snarky question: What else can you judge by? Isn't any alternative just putting more precise conditions on the output? With ChatGPT, it's still easy enough to see it's mistakes, and its attempts at fiction and poetry, impressive as they are, are still clumsy to a trained eye, relative to expert human work. But what if they weren't? What happens when they're indis…

Very hard question to really answer. If Einstein said the sentences “are you sure? think about it”, you might spend years pondering what he meant. If GPT does you’ll ignore it. Same output but the context plus content will change what you do. Same with a painting. If an old master draws a wireframe of a dog, people would bid it up at auction and wonder what he meant. If your kid or AI do, no money might change hands.…

Maybe the answer is a repeated game over time. We might learn the deep wisdom/artistry inherent in the model, or we might learn its limitations. Einstein didn’t arrive and emit one sentence. Dali didn’t just draw once picture. It’s hard to judge an AI from one output.

Re: What is ChatGPT doing and why does it work?

#128

It's worth keeping in mind that Stephen Wolfram likely didn't write this himself. I know people who work at the company, and they sign agreements that any intellectual property (including mathematical proofs) they generate are owned by Stephen Wolfram. Anything Wolfram puts out, like blog posts, scientific articles, and books, are likely to be partly or wholly ghost-written.

It must be Wolfram's own. Every other sentence starts with "But", which is his signature feature :-)

Re: What is ChatGPT doing and why does it work?

#129
post #15

The answer to this is: "we don't really know as its a very complex function automatically discovered by means of slow gradient descent, and we're still finding out" Here are some of the fun things we've found out so far: - GPT style language models try to build a model of the world: https://arxiv.org/abs/2210.13382 - GPT style language models end up internally implementing a mini "neural network training algorithm" (…

Finally I'm tired of people saying it's just a probabilistic word generator and downplaying everything as if they know. If you said something along these lines before... then these papers show that you're not fully grasping the situation here.

There are clearly different angles of interpreting what these models are actually doing but people are stubbornly refusing to believe it's anything more then just statistical word jumbles.

I think part of it is a subconscious fear. chatGPT/LLMs represent a turning point in the story of humanity. The capabilities of AI can only expand from here. What comes after this point is unknown, and we fear the unknown.

I realize what I'm saying is rather dramatic but if you think about it carefully the change chatGPT represents is indeed dramatic... my reaction is extremely appropriate. It's our biases and our tendencies that are making a lot of us down play the whole thing. We'd rather keep doing what we do as if it's business as usual rather then acknowledge reality.

Last week a friend told me it's all just statistical word predictors and that I should look up how neural networks work and what LLMs are as if I didn't already know. Literally I showed him examples of chatGPT doing things that indicate deep understanding of self and awareness of complexity beyond just some predictive word generation. But he stubbornly refused to believe it was anything more. Now, I have a actual research paper to shove in his face.

Man.. People nowadays can't even believe that the earth is round without a research paper stating the obvious.

Re: What is ChatGPT doing and why does it work?

#130
post #87

Earlier quoted context omitted.

I saw a great comment here, and I will repeat it without the attribution it deserves: We may have realized it's easier to build a brain than to understand one

Define understand, and does an analog to Godel's incompleteness apply.

> does an analog to Godel's incompleteness apply

not GP but this seems like quite an attractive idea that many people have reached: a brain of a given "complexity" cannot comprehend the activity of another brain of equal or higher complexity. I'm positive I'm cribbing this from scifi somewhere, maybe Clarke Or Asimov, but, it's the same idea as the Chomsky hierarchy, and the Godel theorems seem like a generalization of that to general sets of rules rather than mere "automata".

For example, you can generalize a state automata to have N possible actors transitioning state at discrete clock intervals, but each actor can keep transitioning and perhaps even spawn additional ones. The machine never terminates until all actors have reached a termination state. That machine is probably impossible to model on any kind of a Turing machine in polynominal time. And a machine that operates at continuous intervals is of course impossible to model on a Discrete Neural Machine in polynomial time (integers vs reals categorization). There are perhaps a lot of complexity categories here, similar to alephs of infinity or problems in P/NP, and when you generalize the complexity categorization to infinity, you get godel incompleteness, just an abstract set of rules governing this categorization of rule sets and what amounts to their computability/decidability.

Everyone is fishing at this same idea, a human has no chance of slicing open a brain (or even imaging it) and having any idea what any of those electrical sparkles mean. At most you could perhaps model some tiny fraction for a tiny quantum, with great effort. We have to rely on machines to assist us for that - probably neural nets, a machine of equal or greater complexity. And we will probably have to rely on machine analysis to be like "ok this ganglion is the geographic center of the AI, and this flash here is the concept of Italy", as far as that even has any meaning at all in a brain. Mere line by line analysis of a Large Language Model or other deep neural network by a human is essentially impossible in any sort of realtime fashion, yeah you can probably model a quantum or two of it statistically and be like "aha this region lights up when we ask about the location of the alps" but the best you are going to do is observational analysis of a small quantum of it during a certain controlled known sequences of events. Unless you build a machine of similar complexity to interpret it. Just like a brain, and just like a state machine emulating a machine of higher complexity-category. They're all the same thing, categories of computability/power.

This is not in any way rigorous, just some casual observations of similarities and parallels between these concepts. It seems like everyone is brushing at that same concept, maybe that helps to get it out on paper.

For an actual hot take: it seems quite clear that our computability as a consciousness depends on the computing power of a higher complexity machine, the brain. Our consciousnesses are really emulated, we totally do live in a simulation and the simulator is your brain, a machine of higher complexity.

Isn't it such a disturbing thought that all your conscious impulses are reduced to a biological machine? Or at least it's of equivalent complexity to one. And the idea that our own conscious and unconscious desires are shaped by this biological machine that may not even be fully explicable. That has been a science fiction theme for a very long time, or the Phineas Gage case, the idea that we are all monsters but for circumstance and we are captives of this biological machine and its unpredictable impulses. We are the neural systems we've trained, and implacable biology they're running on - you change the machine and you also change the person, Phineas Gage was no less conscious and self-cognizent than any of us. He just was a completely different person minus that bit, his conscious being's thought-stream was different because of the biological machine behind it. It's the literal plato's cave, our conscious thoughts are the shadow played out by our biological machine and its program (not to say it's a simple one!).

It's not inherently a bad thing - we incorporate distributed linear/biological systems all over the body in addition to consciousness. reflexes fire before nerve impulses are processed by the conscious center, your eyes are chemical photosensors and can respond to extremely quick instantaneous (high shutter speed) "flash" exposures like silhouettes. And the brain is a highly parallel processor that responds to them. But logical consciousness is a very discrete and monodirectional thing compared to these peripheral biological systems and its computational category is fairly low compared to the massively-parallel brain it runs on. but, we've also mastered these other AI/computational-neural systems now to be a force multiplier for us, we can build systems that we direct in logical thought for us (Frank Herbert would like to remind us that this is a sin ;). Tool-making has always been one of the greatest signifiers of intelligence, it may be quintessentially the sign of intelligence in terms of evolution of consciousness between certain tiers of computation.

And humanity is about to build really good artificial brains on a working scale in the next 25 years, and probably interface with brains (in good and bad ways) before too many more decades after. But it doesn't make any logical sense to try and explain how the model works on a line by line level, any more than it does with the brain model we based it on. Completely pointless to try, it only makes sense if you look at the whole thing and what's going on, it's about the brainwaves, neurons firing in waves and clusters.

/not an AI, just fun at parties, condolences if you read all that shit ;)

Post reply on HN