Earlier quoted context omitted.
If you study copyrighted material for four years at a university and then go on to earn money based on your education, do you owe something to the authors of your text books? I'm not sure how we should treat LLMs with respect to publicly accessible but copyrighted material, but it seems clear to me that "profiting" from copyrighted material isn't a sufficient criteria to cause me to "owe something to the owner".
We don’t, and shouldn’t, give LLMs the same rights as people.
The New York Times is suing OpenAI and Microsoft for copyright infringement
311–320 of 912 posts
Re: The New York Times is suing OpenAI and Microsoft for copyright infringement
#312Re: The New York Times is suing OpenAI and Microsoft for copyright infringement
#313Re: The New York Times is suing OpenAI and Microsoft for copyright infringement
#314Solidly rooting for NYT on this - it’s felt like many creative organizations have been asleep at the wheel while their lunch gets eaten for a second time (the first being at the birth of modern search engines.) I don’t necessarily fault OpenAI’s decision to initially train their models without entering into licensing agreements - they probably wouldn’t exist and the generative AI revolution may never have happened if…
It’s likely fair use.
This is also related to earlier studies about OpenAI where their models have a bad habit of just regurgitating training data verbatim. If your trained data is protected IP you didn’t secure the rights for then that’s a real big problem. Hence this lawsuit. If successful, the floodgates will open.
Re: The New York Times is suing OpenAI and Microsoft for copyright infringement
#315[dead]
OpenAI is separated from MS because they can claim openAI is "research" and thus claim "research" exemption for fair use in copyright law.
Re: The New York Times is suing OpenAI and Microsoft for copyright infringement
#316And here I am thinking it'd be amazing to have an AI that can on-demand read me every novel ever written. It'd be even cooler to jump into a text adventure game of any novel and have it actually follow the original text. I guess that clashes with our copyright world. (Is there hope of some kind of Netflix/Spotify model, with fractional royalties?)
Re: The New York Times is suing OpenAI and Microsoft for copyright infringement
#317Solidly rooting for NYT on this - it’s felt like many creative organizations have been asleep at the wheel while their lunch gets eaten for a second time (the first being at the birth of modern search engines.) I don’t necessarily fault OpenAI’s decision to initially train their models without entering into licensing agreements - they probably wouldn’t exist and the generative AI revolution may never have happened if…
Re: The New York Times is suing OpenAI and Microsoft for copyright infringement
#318The arguments about being able to mimic New York Times “style” are weak, but the fact that they got it to emit verbatim NY Times content seems bad for OpenAI: > As outlined in the lawsuit, the Times alleges OpenAI and Microsoft’s large language models (LLMs), which power ChatGPT and Copilot, “can generate output that recites Times content verbatim
Sarah Silverman is claiming the same thing about her book. But I've tried really hard to get ChatGPT to output sentences verbatim from her book and just can't get it to. In fact, I can't even get it to answer simple questions about facts that are in her book but nowhere else -- it just says it doesn't know. Similarly I haven't been able to reproduce any text in the NYT verbatim unless it's part of a common quote or p…
Re: The New York Times is suing OpenAI and Microsoft for copyright infringement
#319Earlier quoted context omitted.
> Critically the question is, did the developers put reasonable guardrails in place to prevent it? Why? If I steal a bunch of unique works of art and store them in my house for only me to see, am I still committing a crime?
violating copyright is not stealing - it's a government granted monopoly...
Re: The New York Times is suing OpenAI and Microsoft for copyright infringement
#320I hope this results in Fair Use being expanded to cover AI training. This is way more important to humanity's future than any single media outlet. If the NYT goes under, a dozen similar outlets can replace them overnight. If we lose AI to stupid IP battles in its infancy, we end up handicapping probably the single most important development in human history just to protect some ancient newspaper. Then another country…