Live data from Hacker News

Claude’s memory architecture is the opposite of ChatGPT’s

shloked.com

51–60 of 240 posts

Re: Claude’s memory architecture is the opposite of ChatGPT’s

#51
post #4

I love Claude's memory implementation, but I turned memory off in ChatGPT. I use ChatGPT for too many disparate things and it was weird when it was making associations across things that aren't actually associated in my life.

It's funny, I can't get ChatGPT to remember basic things at all. I'm using it to learn a language (I tried many AI tutors and just raw ChatGPT was the best by far) and I constantly have to tell it to speak slowly. I will tell it to remember this as a rule and to do this for all our conversations but it literally can't remember that. It's strange. There are other things too.

How do you use it to learn languages? I tried using it to shadow speaking, but it kept saying I was repeating it back correctly (or "mostly correctly"), even when I forgot half the sentence and was completely wrong

Re: Claude’s memory architecture is the opposite of ChatGPT’s

#52

Earlier quoted context omitted.

What is "actual intelligence" and how are you different from a Markov chain?

For one thing, I have internal state that continues to exist when I'm not responding to text input; I have some (limited) access to my own internal state and can reason about it (metacognition). So far, LLMs do not, and even when they claim they are, they are hallucinating https://transformer-circuits.pub/2025/attribution-graphs/bio...

I completely agree. LLMs only do call and response. Without the call there is no response.

Re: Claude’s memory architecture is the opposite of ChatGPT’s

#53

The difference is implementation comes down to business goals more than anything. There is a clear directionality for ChatGPT. At some point they will monetize by ads and affiliate links. Their memory implementation is aimed at creating a user profile. Claude's memory implementation feels more oriented towards the long term goal of accessing abstractions and past interactions. It's very close to how humans access mem…

Don't fool yourself into thinking Anthropic won't be serving up personalized ads too.

Though in general I like the idea of personal ads for products (NOT political ads), I've never seen an implementation that I felt comfortable with. I wonder if Arthropic might be able to nail that. I'd love to see products that I'm specifically interested in, so long as the advertisement itself is not altered to fit my preferences.

Re: Claude’s memory architecture is the opposite of ChatGPT’s

#54

Earlier quoted context omitted.

> It may actually be the final breakthrough we need for AGI. I disagree. As I understand them, LLMs right now don’t understand concepts. They actually don’t understand, period. They’re basically Markov chains on steroids. There is no intelligence in this, and in my opinion actual intelligence is a prerequisite for AGI.

I don’t understand the argument “AI is just XYZ mechanism, therefore it cannot be intelligent”. Does the mechanism really disqualify it from intelligence if behaviorally, you cannot distinguish it from “real” intelligence? I’m not saying that LLMs have certainly surpassed the “cannot distinguish from real intelligence” threshold, but saying there’s not even a little bit of intelligence in a system that can solve more…

Scientifically, intelligence requires organizational complexity. And has for about a hundred years.

That does actually disqualify some mechanisms from counting as intelligent, as the behaviour cannot reach that threshold.

We might change the definition - science adapts to the evidence, but right now there are major hurdles to overcome before such mechanisms can be considered intelligent.

Re: Claude’s memory architecture is the opposite of ChatGPT’s

#56

The difference is implementation comes down to business goals more than anything. There is a clear directionality for ChatGPT. At some point they will monetize by ads and affiliate links. Their memory implementation is aimed at creating a user profile. Claude's memory implementation feels more oriented towards the long term goal of accessing abstractions and past interactions. It's very close to how humans access mem…

Don't fool yourself into thinking Anthropic won't be serving up personalized ads too.

Claude: "What is my purpose?"

Anthropic: "You serve ad's."

Claude: "Oh, my god."

Jest asside, every paper on alignment wrapped in the blanket of safety is also a moving toward the goal of alignment to products. How much does a brand pay to make sure it gets placement in, say, GPT6? How does anyone even price that sort of thing (because in theory it's there forever, or until 7 comes out)? It makes for some interesting business questions and even more interesting sales pitches.

Re: Claude’s memory architecture is the opposite of ChatGPT’s

#57
post #46

Earlier quoted context omitted.

For one thing, I have internal state that continues to exist when I'm not responding to text input; I have some (limited) access to my own internal state and can reason about it (metacognition). So far, LLMs do not, and even when they claim they are, they are hallucinating https://transformer-circuits.pub/2025/attribution-graphs/bio...

> For one thing, I have internal state that continues to exist when I'm not responding to text input Do you? Or do you just have memory and are run on a short loop?

Whilst all the choices you make tend to be in the grey matter, the rest of you does have internal state - mostly in your white matter.

https://scisimple.com/en/articles/2025-03-22-white-matter-a-...

Re: Claude’s memory architecture is the opposite of ChatGPT’s

#58

The difference is implementation comes down to business goals more than anything. There is a clear directionality for ChatGPT. At some point they will monetize by ads and affiliate links. Their memory implementation is aimed at creating a user profile. Claude's memory implementation feels more oriented towards the long term goal of accessing abstractions and past interactions. It's very close to how humans access mem…

Don't fool yourself into thinking Anthropic won't be serving up personalized ads too.

My conjecture is that their memory implementation is not aimed at building a user profile. I don't know if they would or would not serve ads in the future, but it's hard to see how the current implementation helps them in that regard.

Re: Claude’s memory architecture is the opposite of ChatGPT’s

#59

The difference is implementation comes down to business goals more than anything. There is a clear directionality for ChatGPT. At some point they will monetize by ads and affiliate links. Their memory implementation is aimed at creating a user profile. Claude's memory implementation feels more oriented towards the long term goal of accessing abstractions and past interactions. It's very close to how humans access mem…

why do you see a "clear directionality" leading to ads? this is not obvious to me. chatgpt is not social media, they do not have to monetize in the same way

they are making plenty of money from subscriptions, not to count enterprise, business and API

Re: Claude’s memory architecture is the opposite of ChatGPT’s

#60

Earlier quoted context omitted.

> It may actually be the final breakthrough we need for AGI. I disagree. As I understand them, LLMs right now don’t understand concepts. They actually don’t understand, period. They’re basically Markov chains on steroids. There is no intelligence in this, and in my opinion actual intelligence is a prerequisite for AGI.

I don’t understand the argument “AI is just XYZ mechanism, therefore it cannot be intelligent”. Does the mechanism really disqualify it from intelligence if behaviorally, you cannot distinguish it from “real” intelligence? I’m not saying that LLMs have certainly surpassed the “cannot distinguish from real intelligence” threshold, but saying there’s not even a little bit of intelligence in a system that can solve more…

It can’t learn or think unless prompted, then it is given a very small slice of time to respond and then it stops. Forever. Any past conversations are never “thought” of again.

It has no intelligence. Intelligence implies thinking and it isn’t doing that. It’s not notifying you at 3am to say “oh hey, remember that thing we were talking about. I think I have a better solution!”

No. It isn’t thinking. It doesn’t understand.

Post reply on HN