Anybody have an explanation as to why repeating a token would cause it to regurgitate memorized text?
I'd guess it's a result of punishing repetition at the RLHF stage to stop it getting into the loops that copilot etc used to so easily fall into.
It’s one thing to be able to mimic human text, but to be able to ‘know’ what it means to repeat in general seems to be a slightly higher level of abstraction than I’d expect would just emerge.
…but maybe LLMs have developed more sophisticated models of language than I think.