Earlier quoted context omitted.
What if I overfit my LLM so it spits out copyrighted work with special prompting? Where to draw the line in training?
I mean the human brain can memorize things as well and it’s not illegal. It’s only illegal if said memorized thing is distributed.
Even if LLMs were actual human-level AI (they are not - by far), a small bunch of rich people could use them to make enormous amounts of money without putting in the enormous amounts of work humans would have to.
All the while "training" (= precomputing transformations which among other things make plagiarism detection difficult) on work which took enormous amounts of human labor without compensating those workers.