Claude’s memory architecture is the opposite of ChatGPT’s
1–10 of 240 posts
Re: Claude’s memory architecture is the opposite of ChatGPT’s
#2Re: Claude’s memory architecture is the opposite of ChatGPT’s
#3Re: Claude’s memory architecture is the opposite of ChatGPT’s
#4Re: Claude’s memory architecture is the opposite of ChatGPT’s
#5good (if superficial) post in general, but on this point specifically, emphatically: no, they do not -- no shade, nobody does, at least not in any meaningful sense
Re: Claude’s memory architecture is the opposite of ChatGPT’s
#6> Anthropic's more technical users inherently understand how LLMs work. good (if superficial) post in general, but on this point specifically, emphatically: no, they do not -- no shade, nobody does, at least not in any meaningful sense
Re: Claude’s memory architecture is the opposite of ChatGPT’s
#7ChatGPT is quickly approaching (perhaps bypassing?) the same concerns that parents, teachers, psychologists had with traditional social media. It's only going to get worse, but trying to stop the technological process will never work. I'm not sure what the answer is. That they're clearly optimizing for people's attention is more worrisome.
Re: Claude’s memory architecture is the opposite of ChatGPT’s
#8> Anthropic's more technical users inherently understand how LLMs work. good (if superficial) post in general, but on this point specifically, emphatically: no, they do not -- no shade, nobody does, at least not in any meaningful sense
There is a lot left to learn about the behaviour of LLMs, higher-level conceptual models to be formed to help us predict specific outcomes and design improved systems, but this meme that "nobody knows how LLMs work" is out of control.
Re: Claude’s memory architecture is the opposite of ChatGPT’s
#9This is really cool, I was wondering how memory had been implemented in ChatGPT. Very interesting to see the completely different approaches. It seems to me like Claude's is better suited for solving technical tasks while ChatGPT's is more suited to improving casual conversation (and, as pointed out, future ads integration).
I think it probably won't be too long before these language-based memories look antiquated. Someone is going to figure out how to store and retrieve memories in an encoded form that skips the language representation. It may actually be the final breakthrough we need for AGI.
Re: Claude’s memory architecture is the opposite of ChatGPT’s
#10> Anthropic's more technical users inherently understand how LLMs work. good (if superficial) post in general, but on this point specifically, emphatically: no, they do not -- no shade, nobody does, at least not in any meaningful sense
This is likely (certainly?) impossible. So not a useful definition.
Meanwhile, I have observed a very clear binary among people I know who use LLMs; those who treat it like a magic AI oracle, vs those who understand the autoregressive model, the need for context engineering, the fact that outputs are somewhat random (hallucinations exist), setting the temperature correctly...