The vocabulary truly is load-bearing, without these words the model is less able to think. Where a human can understand a concept without words, an LLM plainly cannot. This is based both on the technological limitations and based on the evidence we see: as these models get better at working they get worse at communication.
> as these models get better at working they get worse at communication.
There is no working other than communication. They are just text generators.
The vocabulary truly is load-bearing, without these words the model is less able to think. Where a human can understand a concept without words, an LLM plainly cannot. This is based both on the technological limitations and based on the evidence we see: as these models get better at working they get worse at communication.
> Where a human can understand a concept without words, an LLM plainly cannot. If we accept that the LLM can "understand" at all, why would we reject the possibility of this understanding living in "latent space", before tokens are output?
Because these modeks have no latent space. They work by tokens and nothing else.
I've recently seen this mentioned more and more, both on HN and on reddit. It seems these output patterns are getting worse. It's not just Claude, my impression is that all of the current models have this style issue. Their writing can get borderline incomprehensible. Is there some feedback loop or compounding happening with each model generation? Maybe newer models are ingesting too much AI content? If the ratio of…
Besides vocabulary I'd also be curious about how it structures sentences. The "X, not Y" is well known, but another thing that bothers me is "It s no " instead of "It doesn't ". For example: "the list contains no string" or "it changes no behavior" or "it holds no directory".
I'm not even sure a PhD helps. It just overuses jargon that has NO meaning. Sometimes, it actually hand waves too much as well while trying to dumb down stuff for you. I am not sure whether it's a consequence of learning to reason from its traces or some RLHF that trips it into using weird terms to sound smarter to the humans who rate it.
I have a PhD and can confirm. Oftentimes, the stuff which comes out of Claude is just impenetrable because it invents jargon on the fly, and uses verbs in the most atrocious ways. "The fibred side folded its capstone into the existing name, so the kinds are asymmetric." What on earth does it mean to fold a capstone into a name‽
Do remember Claude is not try to mean anything. It is extruding a mash-up of the web slop on which it was fed.
Congratulations! I was going to comment on the scroll field in particular when I saw this. I didn't even realize you had to hand-craft the component, but it's such a nice UI idea in general, the way the scrolling works and how the content above changes.
Thanks! The real and probably unnecessary pain was having the words in different sizes.
Frankly, that's the one part I would have done differently. Mixing font sizes rarely looks good, but that's just my personal opinion. If you want to visually indicate intensity, I'd have used color, or rather saturation. Not the text itself, maybe a small pill next to it. But I don't think it's necessary, the order already conveys some sort of ranking, I don't think cardinal information adds much here.
Besides vocabulary I'd also be curious about how it structures sentences. The "X, not Y" is well known, but another thing that bothers me is "It s no " instead of "It doesn't ". For example: "the list contains no string" or "it changes no behavior" or "it holds no directory".
Very nice!
I’d be curious to see if there are words that didn’t show a spike in utilization.
So basically all those words that were used before and are still being used today by LLMs and humans alike!
Interesting response from Claude, when I attempted to reduce the "load-bearing" in its responses: I added to my global prompt: - Orwell's first rule: never use a metaphor you're used to seeing in print. "Load-bearing", "the crux", "first-class citizen" signal insight instead of showing it. Name the specific mechanism when I asked what it thought of the change, its reply was: The Orwell bullet fights my own system pro…
So all it'd take for anthropic to fix this is to just adjust their system prompt? lmao
Well, they can reword it but then Claude would just have a new favourite word.