Live data from Hacker News

Stop Anthropomorphizing Intermediate Tokens as Reasoning/Thinking Traces (2025)

arxiv.org

271–280 of 296 posts

Re: Stop Anthropomorphizing Intermediate Tokens as Reasoning/Thinking Traces (2025)

#271

Earlier quoted context omitted.

I’m not leaving consciousness undefined. I provided a definition and asserted LLMs were not conscious by that definition. I solicited an alternate definition and was given none. If you’d like to claim the LLM is conscious, you need to define what you mean, for both “the LLM” and “is conscious,” because it’s not passing any of our existing bars for consciousness, and the only thing you can point to are characteristics…

Philosophically - sure. But philosophy is a feedback loop. Humans trying to build a model of their brain within their own brain. There is by definition not enough oomph there, and that model will necessarily be approximate at best. Think that's what an out of touch techie would say? May I remind you that just a few hundred years ago the best philosophers were debating whether the world would descend into anarchy if m…

I'm still waiting for a definition of "consciousness," as well as a definition of the entity which you're claiming might possess it.

You've now added two additional entities that require definition for your statements to be meaningful in any sense - "philosophy," the entire field of which you're dismissing as useless, which is going to be extremely fun if you ever actually decide to dig into the neuroscience of consciousness and how the brain works to make sense of the world, and creativity, which has come into the conversation for some reason I'm not entirely sure of but also warrants a definition that isn't trivially satisfiable by either an I Ching or a double pendulum, which, if you're arguing those are conscious, sure, I guess we could throw anything in that bucket then.

You're arguing with all the rigor of a stoned college student, and fine, that's a register you can stay in, no problem with someone having hobbies, but at least have the decency to recognize what you're doing and acknowledge that other people have actually put in the work to be able to discuss and evaluate some of the questions you're positing as unknowable conundrums.

Re: Stop Anthropomorphizing Intermediate Tokens as Reasoning/Thinking Traces (2025)

#272

Earlier quoted context omitted.

> I've gently tried to explain it's like arguing with your compiler for giving you an incoherent error message - it's pointless. Unless doing so changed the compiler output, which is what happens when you say different things to an LLM.

code changes change the compiler output.

And arguing changes LLM responses, especially when they are wrong

Re: Stop Anthropomorphizing Intermediate Tokens as Reasoning/Thinking Traces (2025)

#273
post #172

Earlier quoted context omitted.

Ability to feel?

> Ability to feel Interesting, possibly productive, but still not clear: that can be interpreted as just "reacting to input".

> that can be interpreted as just "reacting to input"

How? That's not the same.

Re: Stop Anthropomorphizing Intermediate Tokens as Reasoning/Thinking Traces (2025)

#274
post #28

Is anthropomorphizing a real problem? From what I know, none of the serious LLM researchers believe it has anything to do with human reasoning, apart from Anthropic with their click-baity terminology like "LLM biology". It's just a metaphor. "Reasoning tokens" is simpler to say than "learned prompt augmentation tokens". I used to (and still do) anthropomorphize things long before LLMs, and I've seen my colleagues do…

Because plenty of people, even ones that should know better, really believe it's a conscious, thinking entity, not just some turn of phrase. I have a coworker that spends at least 10 hours a week arguing with his like you would with a conscious person. I've gently tried to explain it's like arguing with your compiler for giving you an incoherent error message - it's pointless. It doesn't understand, it can't understa…

No. You’re the one who is ignorant.

The LLM is a black box. We do not, in fact know how it works. We know the learning algorithm and we know the scaffold of the transformer network, but the end result of all the weights interacting with each other is something we do not understand. We do not understand this anymore than we understand the human brain.

Now from this the best technical answer we can give is that we don’t know whether the LLM understands or is conscious. But you have to realize that same lack of understanding applies to humans. From a technical standpoint, You cannot say whether your best friend is conscious or not for the same technical reasons as to why you cannot say the same for the LLM. You don’t in actuality know anything.

So, when we have a machine that produces output identical and indistinguishable from an intelligent entity it is actually reasonable to call it conscious, because we already do this for humans. There is no other factor involved. What is clear is that the LLM isn’t human… there is enough evidence to show that its nature is extremely alien. But to say it doesn’t understand or it isn’t self aware is not something anyone can definitively make a statement about other then the fact that it BEHAVES and communicates in a virtually indistinguishable way from something that is self aware.

HN is full of arm chair experts who think they know what they are talking about. But you guys actually don’t. HN was wrong about AI and self driving cars, now we have Waymo. 10 years of research produced self driving cars that are 10x safer than humans. HN was wrong about LLMs. In the beginning HN was sure all it could do was write slop bootstrapped code… now it writes code for all of us. More than the general public HN has been making wrong predictions and wrong statements about AI with an authority that is outright ludicrous. We need to stop. It’s embarrassing how wrong we’ve been.

Re: Stop Anthropomorphizing Intermediate Tokens as Reasoning/Thinking Traces (2025)

#275

Earlier quoted context omitted.

> if it could, you arguing with it isn't going to make it "learn" or act differently. Are you talking about a specific harness that doesn't have context retention mechanisms? For example, ChatGPT with disabled memory feature? Or in general where "it" is a fixed-weights network? The latter is trivially true, of course.

Even claude with “memory” enabled isn’t really “remembering” anything. It just injects it into the context and you hope it happens to find it relevant in its attention mechanisms, and then remembers to actually act on it. Anthropic’s own documentation states claude can and will ignore/truncate these. It’s a context trick, nothing approaching actual “memory,” and in fact, arguing with it will make a bunch of memory fi…

Reading this is like watching someone call cars moving a “trick” because it uses gasoline as fuel while humans don’t move with gasoline.

Bro the LLM is a token machine, it is reasonable to have its short term memory be represented as tokens in context because the LLM is a token machine. Call it a trick if you want but it does fit the actual definition of what memory in actuality is.

Re: Stop Anthropomorphizing Intermediate Tokens as Reasoning/Thinking Traces (2025)

#276

Earlier quoted context omitted.

Because plenty of people, even ones that should know better, really believe it's a conscious, thinking entity, not just some turn of phrase. I have a coworker that spends at least 10 hours a week arguing with his like you would with a conscious person. I've gently tried to explain it's like arguing with your compiler for giving you an incoherent error message - it's pointless. It doesn't understand, it can't understa…

Some tools like code rabbit (PR review bot) encourage you to do this. I couldn’t believe I found myself replying to code review comments to explain to an AI why we would rather let an exception crash the app than to catch and hide it several times so that it would stick in its memory. Having to interact with bots as if they are humans, especially when they are gate keeping, is degrading.

You shouldn’t be interacting like it is a human. It is a LLM. And if the interaction happens to be in a form of prose similar to how you talk to other humans, that is coincidental.

Re: Stop Anthropomorphizing Intermediate Tokens as Reasoning/Thinking Traces (2025)

#277

Earlier quoted context omitted.

And biological machine is? Don’t get me wrong. Biology I is full of molecules that we call machines. But you’re making a broader claim, saying that biology is only this. This needs you to answer some questions: 1. Why do the machine parts in biology show such flexible application? A gear cog won’t ever moonlight as a signaling chip, but in biology you often have molecules doing double and triple duty. 2. How is the b…

> And biological machine is? "Biological," too, is well defined. It relates to living things and their processes. > Why do the machine parts in biology show such flexible application? Evolution. > A gear cog won’t ever moonlight as a signaling chip A gear cog was purpose built for that purpose, but you will find that people often recycle parts into other systems, often in completely different roles. > How is the biol…

>"Biological," too, is well defined. It relates to living things and their processes.

And biological clearly exceeds the definition of “machine”. So once again, what the hell does “just a biological machine” mean?

> Evolution.

Not just any evolution. Evolution in biology follows a specific set of rules driven by structural and functional properties of its component molecules. Those rules do not hold for machines. Yet another reason calling life “just” biological machines is bizarre. The rulesets for change over time do not overlap between machines and biology.

> A gear cog was purpose built for that purpose, but you will find that people often recycle parts into other systems, often in completely different roles.

That’s the thing, you need people. And even with people intervening, our manufactured machines show nothing like the flexibility of function of biological molecules, showing again how these are different classes of things in the real world.

> Protein synthesis.

Lack of knowledge showing. Where do nucleotides and lipids come from then? But the deeper question is: why is there protein synthesis, lipid synthesis and nucleotide synthesis, but no natural silicon synthesis or chip assembly? Why does one arise naturally and sustain itself whereas the other is very reliant on human intervention?

> The way that a machine is built has no bearing on how the machine functions. I could build the same machine using a 3d printer or a CNC router.

Yeah this doesn’t hold for biology.

> Evolution selected for organisms that survive long enough to reproduce. Different biological systems handle this differently.

This isn’t an explanation. All of biology reproduces. All of biology doesn’t share an inner drive and agentic behavior. Once again, you show a 6th grade level understanding of biology while making sweeping claims about it.

> If you give an agent a goal, it will perform actions to achieve that goal. This is just as true for artificial agents as it is for biological agents.

Yes IF you give it a goal. This isn’t true for biology. You don’t need to give bacteria a goal. A newly formed bacterial cell interacts with its environment and then sets its goals.

I did specifically say not LLM has been found that works unprompted. You just moved the prompts to an agents goal document. It still needs a human prompt up the chain. Even if you had an LLM give the goal to another LLM, the first one still needed human prompting. This causal chain can’t be wished away just for you to ignore how biology is different.

> Many do. Even robotic vacuum cleaners will charge themselves without human prompting.

Seriously? Do you not understand that the robot vacuum runs on deterministic code?

Re: Stop Anthropomorphizing Intermediate Tokens as Reasoning/Thinking Traces (2025)

#278

Earlier quoted context omitted.

We do have entry and exit conditions. We can disrupt it and study the dynamics of the disruption. We have a good sense of the many mechanisms at play that undergird it. What we lack is a full map of the exact process from sensation to consciousness across all modalities. And I’m afraid when it comes to interception, we can’t unless we observe every one of the 36 trillion or so cells in a human body, as well as the 36…

None of this is an explanation of what consciousness is, either physically, logically or philosophically. (Maybe some physically) Fundamentally it comes down to an objective decision about what that is. If you say it is “feelings” based on inputs and feedback mechanisms from the brain, then we can do the philosophical discussion around that. “If the claim is that consciousness is limited to those with a specific type…

I’m baffled: how does the fact that each cell has an internal clock that continues to tick even when you take away all external signals not impact what consciousness is logically?

If the time order driving behavior is driven by an internal timekeeper, that is logically relevant to the behavior you’re interrogating. If, on the other hand, all component systems depended on an external clock, that logically points to a completely different dynamic process.

As for the philosophy of it all… it’s true I make no comment on it. I’d rather look at the physical substrate and see what it does, and compare it to behavior, than try to fit a philosophy that originated before we had such high resolution knowledge of the system. They are blind to these facts, just because they were written up too early.

I utterly reject this doesn’t show you anything logically about consciousness. If I can show you that the molecular dynamics of each cell organize to anticipate the dawn, and do so even in constant darkness (and all this is well established.. there was a Nobel for the field in 2017) how is that not logically distinct from a system that is outside of regular time, needs to check a clock to locate itself in time, and then do whatever dynamics it does to solve the problem at hand?

As for calling this a coping mechanism… that’s a bit rich coming from someone who seems to have very little idea of the molecular and cellular biology but seems to want to hold to the belief that we’ve solved consciousness with LLMs. We can both fling that accusation about. Seems more productive to compare the physical dynamics and see what’s different and what’s similar, no?

Re: Stop Anthropomorphizing Intermediate Tokens as Reasoning/Thinking Traces (2025)

#279

Earlier quoted context omitted.

Odd. Why don't the details of how consciousness arises in one system inform your judgment of whether it can exist in a different system that only has partial structural overlap? Seems wildly convenient. Where else in science can you show me such a comparable situation in how you define properties?

> Why don't the details of how consciousness arises in one system inform your judgment of whether it can exist in a different system that only has partial structural overlap? A light bulb doesn't need to do fusion to make light. An airplane doesn't need to flap its wings to fly. An ANN doesn't need to use a brain's structure to think.

You haven’t proven an LLM has thought. I can measure the spectral properties of light from both a bulb and the sun.

I can measure the dynamics that go with thought in a human.

I can measure some dynamics in an LLM. From all we know, there are huge differences, not least that you can literally turn off an LLM, whereas every process but death in biology shows continued activity even when the organism “looks” off,

What you haven’t shown me is that what the LLM does when it is on is “thinking”. Given that the dynamics are different, the substrate is different, and one can be fully turned off and the other cannot, why are these two things the same?

Re: Stop Anthropomorphizing Intermediate Tokens as Reasoning/Thinking Traces (2025)

#280
post #195

Earlier quoted context omitted.

Yes we can say it has conscious experience. The content and depth of it, we cannot yet fully grok, and of course, what it feels like from inside the ant is something we never will know.

Okay, what about a virus.

Nope. It’s got a genuine off state, when it’s out of a cell, where it lacks agency. It’s dormant and can be taken up by a cell only by chance, not through its own active efforts. It bears none of the descriptive hallmarks of consciousness, nor is there any internal physical dynamic inherent to a virus. By definition it needs a host to start having a dynamic, and ant that point, it anctivates its drive, which is singular: reproduction and transmission, which it must balance.

And this is where you see viruses show the beginning of agency, in the making of this choice. They do show some signs of sensing their environment , even communicating with each other, through a process called quorum sensing, to decide if they should follow a lysogeny (integrate and stay quiet) or lysis (rapidly replicate and destroy the cell).

Note, this requires a host, so in so far as we can talk of a viral consciousness, it’s restricted to finding a suitable host.

There’s also recent work on viruses upon host entry showing some sensing. Usually a single molecule for which they already have a receptor. So you can begin to see the thin levels of awareness a virus has. But it still can’t do anything on its own beyond deciding to not infect.

I’d argue viruses are right at the edge of consciousness, just as they are right at the edge of life. They can hijack life and demonstrate much more life like properties, including a proto consciousness.

But they have no internal clock. They are always sensing the environment through multiple rich channels, nor do they have an active biochemistry they need to maintain when they’re dormant.

Consciousness adjacent? Consciousness-compatible? Both seem to fit viruses. And also LLMs, funnily enough.

Post reply on HN