Claude’s memory architecture is the opposite of ChatGPT’s
111–120 of 240 posts
Re: Claude’s memory architecture is the opposite of ChatGPT’s
#112Is the result reliable and not just hallucination? Why would ChatGPT know how itself works and why would it be fed with these kind of learning material?
Re: Claude’s memory architecture is the opposite of ChatGPT’s
#113> Most of this was uncovered by simply asking ChatGPT directly. Is the result reliable and not just hallucination? Why would ChatGPT know how itself works and why would it be fed with these kind of learning material?
Re: Claude’s memory architecture is the opposite of ChatGPT’s
#114Earlier quoted context omitted.
What I mean is that the current generation of LLMs don’t understand how concepts relate to one another. Which is why they’re so bad at maths for instance. Markov chains can’t deduce anything logically. I can.
You and Chomsky are probably the last 2 persons on earth to believe that.
Re: Claude’s memory architecture is the opposite of ChatGPT’s
#115Earlier quoted context omitted.
> It may actually be the final breakthrough we need for AGI. I disagree. As I understand them, LLMs right now don’t understand concepts. They actually don’t understand, period. They’re basically Markov chains on steroids. There is no intelligence in this, and in my opinion actual intelligence is a prerequisite for AGI.
I don’t understand the argument “AI is just XYZ mechanism, therefore it cannot be intelligent”. Does the mechanism really disqualify it from intelligence if behaviorally, you cannot distinguish it from “real” intelligence? I’m not saying that LLMs have certainly surpassed the “cannot distinguish from real intelligence” threshold, but saying there’s not even a little bit of intelligence in a system that can solve more…
Current LLMs are a long way from there.
You may think "sure seems like it passes the Turing test to me!" but they all fail if you carry on a conversation long enough. AIs need some equivalent of neuroplasticity and as of yet they do not have it.
Re: Claude’s memory architecture is the opposite of ChatGPT’s
#116Earlier quoted context omitted.
Don't fool yourself into thinking Anthropic won't be serving up personalized ads too.
Claude: "What is my purpose?" Anthropic: "You serve ad's." Claude: "Oh, my god." Jest asside, every paper on alignment wrapped in the blanket of safety is also a moving toward the goal of alignment to products. How much does a brand pay to make sure it gets placement in, say, GPT6? How does anyone even price that sort of thing (because in theory it's there forever, or until 7 comes out)? It makes for some interesting…
Re: Claude’s memory architecture is the opposite of ChatGPT’s
#117Earlier quoted context omitted.
> It may actually be the final breakthrough we need for AGI. I disagree. As I understand them, LLMs right now don’t understand concepts. They actually don’t understand, period. They’re basically Markov chains on steroids. There is no intelligence in this, and in my opinion actual intelligence is a prerequisite for AGI.
> As I understand them, LLMs right now don’t understand concepts. In my uninformed opinion it feels like there's probably some meaningful learned representation of at least common or basic concepts. It just seems like the easiest way for LLMs to perform as well as they do.
Re: Claude’s memory architecture is the opposite of ChatGPT’s
#118Earlier quoted context omitted.
why do you see a "clear directionality" leading to ads? this is not obvious to me. chatgpt is not social media, they do not have to monetize in the same way they are making plenty of money from subscriptions, not to count enterprise, business and API
The router introduced in gpt-5 is probably the biggest signal. A router, while determining which model to route query, can determine how much $$ a query is worth. (Query here is conversation). This helps decide the amount of compute openai should spend on it. High value queries -> more chances of affiliate links + in context ads. Then, the way memory profile is stored is a clear way to mirror personalization. Ads wor…
Our goal for the router (whether you think we achieved it or not) was purely to make the experience smoother and spare people from having to manually select thinking models for tasks that benefit from extra thinking. Without the router, lots of people just defaulted to 4o and never bothered using o3. With the router, people are getting to use the more powerful thinking models more often. The router isn't perfect by any means - we're always trying to improve things - but any paid user who doesn't like it can still manually select the model they want. Our goal was always a smoother experience, not ad injection or cost optimization.
Re: Claude’s memory architecture is the opposite of ChatGPT’s
#119Earlier quoted context omitted.
But aren't we only worth something like $300/year each to Meta in terms of ads? I remember someone arguing something like that when the TikTok ban was being passed into law... essentially the argument was that TikTok was "dumping" engagement at far below market value (at something like $60/year) to damage American companies. That was something the argument I remember anyway.
If that’s the case, we have an even bigger problem on our hands. How will these companies ever be profitable? If we’re already paying $20/mo and they’re operating at a loss, what’s the next move (assuming we’re only worth an extra $300/yr with ads?) The math doesn’t add up, unless we stop training new models and degrade the ones currently in production, or have some compute breakthrough that makes hardware + operatin…
ChatGPT isn't going to capture all the engagement. And even then I don't know whether $300 is much particularly after subtracting operating overhead. I'm just saying I have trouble believing there's gold to be had at the end of this LLM ad rainbow. People just seem to throw out ideas like "ads!" as if it's a sure fire winning lottery ticket or something.
Re: Claude’s memory architecture is the opposite of ChatGPT’s
#120This post was great, very clear and well illustrated with examples.