Earlier quoted context omitted.
but you don't see "load bearing" nearly as often in prose written by people, so it's not some irreplaceable phrase. It's just a token with a weirdly high likelihood in a lot of cases (given how Claude works, this kind of thing is bound to happen)
You don't think it's possible that an LLM's internal machinery could decide that an underused-by-humans word should be used more frequently in output than it sees in input because it maps cleanly onto a frequently needed semantic? I think that's possible
A more parsimonious explanation is that this term got more-or-less randomly boosted by the reinforcement learning loop because there was nothing in the training data to discourage its use.