Live data from Hacker News

Adding a feature because ChatGPT incorrectly thinks it exists

holovaty.com

351–360 of 451 posts

Re: Adding a feature because ChatGPT incorrectly thinks it exists

#351
post #256

Earlier quoted context omitted.

I'll play with your tact in this argument, although I certain do not agree it is accurate. You're asserting that creativity is a meld of past experience, both personal and the creative output of others. Yet this really doesn't jive, as an LLM does not "experience" anything. I would argue that raw knowledge is not "experience" at all. We might compare this to the university graduate, head full of books and data jammed…

Experiences are not materially different from knowledge once they are both encoded as memories. They're both just encoded in neurons as weights in their network of connections. But let's assume there is some ineffable difference between firsthand and secondhand experience, which fundamentally distinguishes the two in the brain in the present. The core question here, then, is why you are so certain that "creativity" r…

Experiences are not materially different from knowledge once they are both encoded as memories.

The storage medium is not relevant here, and actual experience has sensorially laden additions which are not currently possible from reading. Memories are laid down differently as well.

In terms of what top neuroscientists know or do not know, you pulled such out of your back pocket, and threw them into onto the floor, perhaps you should describe precisely how they negate what I am saying?

Is there consensus? LLM is indeed creative?

What you seem to be missing here, is I am not railing against machine intelligence, nor creativity. It is merely that an LLM is not it, and will never become it. This is no different than an argument over whether to use sysvinit or systemd, it is a discussion of technical capabilities of a technology.

LLMs may become a backing store, a "library" of sorts for any future AGI to use as data, knowledge, an exceptionally effective method to provide a wikipedia-ish like, non-sql backed data source.

But they provide no means for cognition.

And creativity requires cognition. Creativity is a conscious process, for it requires imagination, which is an offshoot of a conscious process. Redefining "creativity" to exclude the conscious process negates its very meaning.

You can say "Wow, this appears to be creative", and it may appear to be creative, yet without cognition the act is simply not possible. None would dare say that a large Goldberg machine, which spits out random answers dependent upon air currents was producing creative ideas.

Some may say "What a creative creation this machine is!", but none would attribute creativity to the output of any algorithmic production by that machine, and this is what we have here.

Should we derive a method of actual conscious cognition in a mind not flesh, so be it. Creativity may occur. But as things stand now, a mouse provides more creativity than an LLM, the technology is simply not providing the underlying requirements. There is no process for consciousness.

There are ways to provide for this, and I have pondered them (again, I'm validating here that it's not "oh no, machines are NOT thinking!!", but instead "LLMs aren't that").

One exceptionally rough and barely back-of-napkin concept, would be sleep.

No I am not trying to mimic the human mind here, but when the concept is examined front to end, the caboose seems to be 'sleep'. Right now, the problem is how we bring each LLM onto a problem. We simply throw massive context at it, and then allow it to proceed.

Instead, we need to have a better context window. Maybe we should call this 'short term' memory. An LLM is booted, and responds to questions, but has a floating context window which never shrinks. Its "context window" is not cleared. Perhaps we use a symbolic database, or just normal SQL with fluff modz, but we allow this ongoing context window to exist and grow.

After a point, this short term memory will grow too large to actively swap in/out of memory. Primarily, even RAM has bandwidth limits and is a detractor on response speed and energy requirements per query.

So -- the LLM "goes to sleep". During that time, the backend is converted to LLM, or I suppose in this case a small language model. We now have a fuzzy, RRD almost conversion of short-term memory to long-term memory, yet one which enables some very important things.

That being, an actual capacity to learn from interaction.

The next step is to expand that capability, and the capabilities of an LLM with senses. I frankly think the best here is real, non-emulated robotic control. Give the LLM something to manipulate, as well as senses.

At that point, we should inject agency. A reason to exist. With current life, the sole primary reason is "reproduce". Everything else has derived from that premise. I spoke of the mating urge, we should recreate this here.

(Note how non-creative this all actually is, yet it seems valid to me... we're just trying to provide what we know works as a starting base. It does not mean that we cannot expand the creation of conscious minds into other methods, once we have a better understanding and success.)

There are several other steps here, which are essential. The mind must, for example, reload its "long-term memory" backing store SLM, and when "sleep" comes, overlay new short-term thoughts over long-term. This is another fuzzy process, and it would be best to think of it(though technically not accurate) as unpacking its SLM, overlaying new thoughts, and creating an entirely new SLM. As its short-term memory would have output derived from the LLM, plus overlaid SLM, its short-term memory would be providing output derived from its prior SLM.

So there is a form of continuity here.

So we have:

* A mind which can retain information, and is not merely popping into creation with no stored knowledge, and killed at each session end

* That same mind has a longer term memory, which allows for ongoing concept modification and integration, eg new knowledge affecting the "perception" of old knowledge

* That SLM will be "overlaid" on top of its LLM, meaning experiences derived during waking moments will provide more context to an LLM (that moment when you ride a bike, and comprehend how all the literature you read, isn't the same as doing? That moment where you link the two? That's in the SLM, and the SLM has higher priority)

* A body (simple as it may be), which allows access to the environment

* Senses fed into the process

* Agency, as in, "perform to mate", with aspects of "perform" being "examine, discover, be impressive" that sort of thing

I think, this overlaid SLM, along with actual empirical data would provide a more apt method to simulate some form of consciousness. It would at least allow a stream of consciousness, regardless of whatever debates we might have about humans dying when they sleep (which makes no sense, as the brain is incredibly active during sleep, and constantly monitoring the surroundings for danger).

I'd speak more along the sensory aspects of this, but it's actually what I'm working on right now.

But what's key here, is independent data and sensory acquisition. I see this is a best-able way to kickstart a mind.

Re: Adding a feature because ChatGPT incorrectly thinks it exists

#352

In addition, we might consider writing the scientific papers ChatGPT hallucinates!

That's agentic AI, right? Run the LLM in a loop and give it a tool to publish to arxiv. If it cites a paper that doesn't exist, make it write and upload that one too, recursively. Should work for lawyers, too.

Re: Adding a feature because ChatGPT incorrectly thinks it exists

#353

Here's the thing: I don't think ChatGPT per se was the impetus to develop this new feature. The impetus was learning that your customers desire it. ChatGPT is operating as the kind of "market research" tool here, albeit it in a really unusual, inverted way. That said, if someone could develop a market research tool that worked this way, i.e. users went to it instead of you have to use it to go to users, I can see it…

They only want ASCII tablature parsing because that's what ChatGPT produces. If ChatGPT produced standard music notation, users would not care about ASCII tablature. ChatGPT has created this "market".

Re: Adding a feature because ChatGPT incorrectly thinks it exists

#354
post #10

I find it amusing that it's easier to ship a new feature than to get OpenAI to patch ChatGPT to stop pretending that feature exists (not sure how they would even do that, beyond blocking all mentions of SoundSlice entirely.)

If you gave a junior level developer just one or two files of your code, without any ability to look at other code, and asked them to implement a feature, none of them would make ANY reasonable assumptions about what is available?

This seems similar, and like a decent indicator that most people (aka the average developer) would expect X to exist in your API.

Re: Adding a feature because ChatGPT incorrectly thinks it exists

#355
post #129

Earlier quoted context omitted.

> Sometimes it comes up with a better approach than I had thought of. IMO this has always been the killer use case for AI—from Google Maps to Grammarly. I discovered Grammarly at the very last phase of writing my book. I accepted maybe 1/3 of its suggestions, which is pretty damn good considering my book had already been edited by me dozens of times AND professionally copy-edited. But if I'd have accepted all of Gram…

And that’s how everything gets flattened to same style/voice/etc. That’s like getting rid of all languages and accents and switch to the same language

The Esperanto utopia we were denied.

Re: Adding a feature because ChatGPT incorrectly thinks it exists

#356
post #174

Earlier quoted context omitted.

It has its place. We were told to use passive voice when writing scientific document (lab reports, papers etc).

To be fair, current scientific papers are full of utterly terrible writing. If you read scientific papers from a century and a half ago, a century ago, half a century ago, and today, you'll see a continuous and disastrous decline in readability, and I think some of that is driven by pressure to strictly follow genre writing conventions. One of those conventions is using the passive voice even when the active voice wo…

Are we talking about survivorship bias or are you comparing comparably important levels of papers?

Re: Adding a feature because ChatGPT incorrectly thinks it exists

#357

Earlier quoted context omitted.

The problem here is societal, not technological. An end state where people do less work than they do today but society is more productive is desirable, and we shouldn't be trying to force companies/governments/etc to employ people to do an unnecessary job. The problem is that people who are laid off often experience significant life disruption. And people who work in a field that is largely or entirely replaced by te…

> the fact people having their jobs replace by technology are completely screwed over is a result of the society we have all created together, it's not a rule of nature. How did the handloom weavers and spinners handle the rise of the machines?

> How did the handloom weavers and spinners handle the rise of the machines?

In the past, new jobs appeared that the workers could migrate to.

Today, it seems that AI may replace jobs much quicker than before and it's not clear to me which new jobs will be "invented" to balance the loss.

Optimists will say that we have always managed to invent new types of work fast enough to reduce the impact to society, but in my opinion it is unlikely to happen this time. Unless the politicians figure out a way to keep the unemployment content (basic income etc.),

I fear we may end up in a dystopia within our lifetimes. I may be wrong and we could end up in a post scarcity (star trek) world, but if the current ambitions of the top 1% is an indicator, it won't happen unless the politicians create a better tax system to compensate the loss of jobs. I doubt they will give up wealth and influence voluntarily.

Re: Adding a feature because ChatGPT incorrectly thinks it exists

#358
post #169

Earlier quoted context omitted.

> The problem is executives want to completely remove humans from the loop, which almost universally leads to disastrous results Thanks for your words of wisdom, which touch on a very important other point I want to raise: often, we (i.e., developers, researchers) construct a technology that would be helpful and "net benign" if deployed as a tool for humans to use, instead of deploying it in order to replace humans.…

The problem here is societal, not technological. An end state where people do less work than they do today but society is more productive is desirable, and we shouldn't be trying to force companies/governments/etc to employ people to do an unnecessary job. The problem is that people who are laid off often experience significant life disruption. And people who work in a field that is largely or entirely replaced by te…

> The problem here is societal, not technological.

I disagree. I think it's both. Yes, we need good frameworks and incentivizes on a economic/political level. But also, saying that it's not a tech problem is the same as saying "guns don't kill people". The truth is, if there was no AI tech developed, we would not need to regulate it so that greed does not take over. Same with guns.

Re: Adding a feature because ChatGPT incorrectly thinks it exists

#359
post #169

Earlier quoted context omitted.

> Sometimes it comes up with a better approach than I had thought of. IMO this has always been the killer use case for AI—from Google Maps to Grammarly. I discovered Grammarly at the very last phase of writing my book. I accepted maybe 1/3 of its suggestions, which is pretty damn good considering my book had already been edited by me dozens of times AND professionally copy-edited. But if I'd have accepted all of Gram…

> The problem is executives want to completely remove humans from the loop, which almost universally leads to disastrous results Thanks for your words of wisdom, which touch on a very important other point I want to raise: often, we (i.e., developers, researchers) construct a technology that would be helpful and "net benign" if deployed as a tool for humans to use, instead of deploying it in order to replace humans.…

The problem of those greedy business managers you speak of is that, they don't care how the company does 10 year down the line and I almost feel as if everybody is just doing things which work short term ignoring the long term consequences.

As the comment above said that we need a human in the loop for better results, Well firstly it also depends on human to human.

A senior can be way more productive in the loop than a junior.

So Everybody has just stopped hiring juniors because they cost money and they will deal with the AI almost-slop later/ someone else will deal with it.

Now the current seniors will one day retire but we won't have a new generation of seniors because nobody is giving juniors a chance or that's what I've heard about the job market being brutal.

Re: Adding a feature because ChatGPT incorrectly thinks it exists

#360
post #30

I've found this to be one of the most useful ways to use (at least) GPT-4 for programming. Instead of telling it how an API works, I make it guess, maybe starting with some example code to which a feature needs to be added. Sometimes it comes up with a better approach than I had thought of. Then I change the API so that its code works. Conversely, I sometimes present it with some existing code and ask it what it does…

When I see comments like yours I can't help but decry how bad was the "stochastic parrots" framing. A parrot does not hallucinate a better API.
Post reply on HN