Live data from Hacker News

Claude’s memory architecture is the opposite of ChatGPT’s

shloked.com

61–70 of 240 posts

Re: Claude’s memory architecture is the opposite of ChatGPT’s

#61

The difference is implementation comes down to business goals more than anything. There is a clear directionality for ChatGPT. At some point they will monetize by ads and affiliate links. Their memory implementation is aimed at creating a user profile. Claude's memory implementation feels more oriented towards the long term goal of accessing abstractions and past interactions. It's very close to how humans access mem…

why do you see a "clear directionality" leading to ads? this is not obvious to me. chatgpt is not social media, they do not have to monetize in the same way they are making plenty of money from subscriptions, not to count enterprise, business and API

One has a more obvious route to building a profile directly off that already collected data.

And while they are making lots of revenue even they have admitted on recent interviews that ChatGPT on it's own is still not (yet) breakeven. With the kind of money invested, in AI companies in general, introducing very targeted Ads is an obvious way to monetize the service more.

Re: Claude’s memory architecture is the opposite of ChatGPT’s

#62

Earlier quoted context omitted.

Don't fool yourself into thinking Anthropic won't be serving up personalized ads too.

Though in general I like the idea of personal ads for products (NOT political ads), I've never seen an implementation that I felt comfortable with. I wonder if Arthropic might be able to nail that. I'd love to see products that I'm specifically interested in, so long as the advertisement itself is not altered to fit my preferences.

There is no such thing as a good flow for showing sponsored items in an LLM workflow.

The point of using an LLM is to find the thing that matches your preferences the best. As soon as the amount of money the LLM company makes plays into what's shown, the LLM is no longer aligned with the user, and no longer a good tool.

Re: Claude’s memory architecture is the opposite of ChatGPT’s

#64

Earlier quoted context omitted.

I don’t understand the argument “AI is just XYZ mechanism, therefore it cannot be intelligent”. Does the mechanism really disqualify it from intelligence if behaviorally, you cannot distinguish it from “real” intelligence? I’m not saying that LLMs have certainly surpassed the “cannot distinguish from real intelligence” threshold, but saying there’s not even a little bit of intelligence in a system that can solve more…

It can’t learn or think unless prompted, then it is given a very small slice of time to respond and then it stops. Forever. Any past conversations are never “thought” of again. It has no intelligence. Intelligence implies thinking and it isn’t doing that. It’s not notifying you at 3am to say “oh hey, remember that thing we were talking about. I think I have a better solution!” No. It isn’t thinking. It doesn’t unders…

Just because it's not independent and autonomous does not mean it could not be intelligent.

If existing humans minds could be stopped/started without damage, copied perfectly, and had their memory state modified at-will would that make us not intelligent?

Re: Claude’s memory architecture is the opposite of ChatGPT’s

#65
post #16

Earlier quoted context omitted.

What is "actual intelligence" and how are you different from a Markov chain?

Roughly, actual intelligence needs to maintain a world model in its internal representation, not merely an embedding of language, which is a very different data structure and probably will be learned in a very different way. This includes things like: - a map of the world, or concept space, or a codebase, etc - causality - "factoring" which breaks down systems or interactions into predictable parts Language alone is…

[deleted]

Re: Claude’s memory architecture is the opposite of ChatGPT’s

#66

The difference is implementation comes down to business goals more than anything. There is a clear directionality for ChatGPT. At some point they will monetize by ads and affiliate links. Their memory implementation is aimed at creating a user profile. Claude's memory implementation feels more oriented towards the long term goal of accessing abstractions and past interactions. It's very close to how humans access mem…

why do you see a "clear directionality" leading to ads? this is not obvious to me. chatgpt is not social media, they do not have to monetize in the same way they are making plenty of money from subscriptions, not to count enterprise, business and API

> they are making plenty of money from subscriptions, not to count enterprise, business and API

...except that they aren't? They are not in the black and all that investor money comes with strings

Re: Claude’s memory architecture is the opposite of ChatGPT’s

#67

Earlier quoted context omitted.

Don't fool yourself into thinking Anthropic won't be serving up personalized ads too.

Claude: "What is my purpose?" Anthropic: "You serve ad's." Claude: "Oh, my god." Jest asside, every paper on alignment wrapped in the blanket of safety is also a moving toward the goal of alignment to products. How much does a brand pay to make sure it gets placement in, say, GPT6? How does anyone even price that sort of thing (because in theory it's there forever, or until 7 comes out)? It makes for some interesting…

Could be part of a LORA or some other kind of plug-in refinement.

Re: Claude’s memory architecture is the opposite of ChatGPT’s

#68

Earlier quoted context omitted.

Don't fool yourself into thinking Anthropic won't be serving up personalized ads too.

Though in general I like the idea of personal ads for products (NOT political ads), I've never seen an implementation that I felt comfortable with. I wonder if Arthropic might be able to nail that. I'd love to see products that I'm specifically interested in, so long as the advertisement itself is not altered to fit my preferences.

> Though in general I like the idea of personal ads for products (NOT political ads), I've never seen an implementation that I felt comfortable with.

No implementation will work for very long when the incentives behind it are misaligned.

The most important part of the architecture is that the user controls it for the user's best interests.

Re: Claude’s memory architecture is the opposite of ChatGPT’s

#69
post #4

I love Claude's memory implementation, but I turned memory off in ChatGPT. I use ChatGPT for too many disparate things and it was weird when it was making associations across things that aren't actually associated in my life.

I’m the opposite. ChatGPT’s ability to automatically pull from its memory is way better than remembering to to ask.

Re: Claude’s memory architecture is the opposite of ChatGPT’s

#70
post #9

The link to the breakdown of ChatGPT's memory implementation is broken, the correct link is: https://www.shloked.com/writing/chatgpt-memory-bitter-lesson This is really cool, I was wondering how memory had been implemented in ChatGPT. Very interesting to see the completely different approaches. It seems to me like Claude's is better suited for solving technical tasks while ChatGPT's is more suited to improving casual…

> It may actually be the final breakthrough we need for AGI. I disagree. As I understand them, LLMs right now don’t understand concepts. They actually don’t understand, period. They’re basically Markov chains on steroids. There is no intelligence in this, and in my opinion actual intelligence is a prerequisite for AGI.

Human thinking is also Markov chains on ultra steroids. I wonder if there are any studies out there which have shown the difference between people who can think with a language and people who don't have that language base to frame their thinking process in, based on some of those kids who were kept in isolation from society.

"Superhuman" thinking involves building models of the world in various forms using heuristics. And that comes with an education. Without an education (or a poor one), even humans are incapable of logical thought.

Post reply on HN