Earlier quoted context omitted.
> a few copyright holders By which you mean every copyright holder. > AGI in the near future Something that is purely speculative, undefined, and has been promised in the near future for 50+ years. I don't see copyright holders lying down for someone else's benefit and I don't see governments gutting copyright, contract law, and several other avenues of protection that copyright holders can deploy in the name of some…
If a child is instructed to read a copyrighted work at school, which later becomes a factor in his own derivative works, he won't be in breach of copyright. Why should other intelligent entities be prevented from reading copyrighted works and gaining whatever there is to gain from those works the way any human might?
A more practical way of looking at this is: who is making money off of these models? How did they get their training data?
I’m not a fan of copyright in general, but we have serious outstanding issues with companies and organizations stealing or plastering work without compensating the original creators of said works. Thusfar, LLMs are becoming another method to concentrate wealth to whoever has the resources to train and sell these models at scale.