I love the idea of using this for LLM output watermarking. It hits the sweet spot - will catch 99% of slop generators with no fuss, since they only copy and paste anyway, almost no impact on other core use cases. I wonder how much you’d embed with each letter or token that’s output - userid, prompt ref, date, token number? I also wonder how this is interpreted in a terminal. Really cool!
Just you wait until AI starts calling human output to be slop.
The difference is humans are responsible for what they write, whereas the human user who used an AI to generate text is responsible for what the computer wrote.