Companies that have content all see dollar signs. NYT won't mind if you use their content to train LLMs - as long as they get a commission. Reddit will shut down their free API and make you pay to get training content. Discord is going to be selling content for AI training too - if they haven't already done so. Twitter is doing it. They didn't care before because LLMs were just experiments. Now we're talking trillion…
> They didn't care before because LLMs were just experiments. Now we're talking trillions of dollars of value. Can you make the argument this was their fault for not having forward vision/being asleep at the wheel and "accidentally, in hindsight" letting OpenAI/others have free, open, unlimited access to their content?
NY Times copyright suit wants OpenAI to delete all GPT instances
51–60 of 921 posts
Re: NY Times copyright suit wants OpenAI to delete all GPT instances
#52I read a NYT article and publish a summary of facts that I learned: totally legit. Train a model on NYT text that outputs a summary of facts that it learned: OMG literally murder.
I read a NYT article and publish an exact copy of that article on my website: copyright infringement.
Train a model on NYT text and it outputs an exact copy of that text: also copyright infringement.
Re: NY Times copyright suit wants OpenAI to delete all GPT instances
#53It's obviously a frivolous suit that will only net at best a ceremonial victory for NYTimes: 8 figure max payout and a promise to not use NYtimes material in the future. The trajectory and value to society of OpenAI vs NYtimes could not be greater. They have won no favors in the court of public opinion with their frequent misinformation. It's all just a big waste of time, the last of the old guard flailing against th…
Re: NY Times copyright suit wants OpenAI to delete all GPT instances
#54I've been arguing since ChatGPT came out that LLMs should fall under fair use as a "transformative work". I'm not a lawyer and this is just my non-expert opinion, but it will be interesting to see what the legal system has to say about this.
It gets harder to stand behind a blanket claim that LLMs or any AI we’ve got falls under fair use when they keep repeatedly reproducing complete and identifiable individual works and clearly violating copyright laws in specific instances. The models might be remixing and/or transformative most of the time, but we have proof that they don’t do that every time nor all the time… yet. Maybe the lawsuits will be the impetus we need to fix the AIs so they don’t reproduce specific works, and thus make the fair use claim solid and actually defensible?
Re: NY Times copyright suit wants OpenAI to delete all GPT instances
#55Won't hold in court. GPT is a platform mainly providing answer to private individuals asking. Is like you ask a professor a question and he answered verbatim what copyrighted materials available (due to photographic memory) word for word back to you. Now if you take this answer and write a book or publish enmass on blogs for example, then you are the one should be sued by NYT. If GPT use the exact same wordings and p…
Re: NY Times copyright suit wants OpenAI to delete all GPT instances
#56I read a NYT article and publish a summary of facts that I learned: totally legit. Train a model on NYT text that outputs a summary of facts that it learned: OMG literally murder.
Sounds like you didn't read the article. Here's a better synoposis: I read a NYT article and publish an exact copy of that article on my website: copyright infringement. Train a model on NYT text and it outputs an exact copy of that text: also copyright infringement.
Re: NY Times copyright suit wants OpenAI to delete all GPT instances
#57Companies that have content all see dollar signs. NYT won't mind if you use their content to train LLMs - as long as they get a commission. Reddit will shut down their free API and make you pay to get training content. Discord is going to be selling content for AI training too - if they haven't already done so. Twitter is doing it. They didn't care before because LLMs were just experiments. Now we're talking trillion…
"They" also include the people working there. Why someone work with full time writing articles should give the work for free just let someone to train it and make money out of it as a consequence?
OpenSource developers did that ;)
Re: NY Times copyright suit wants OpenAI to delete all GPT instances
#58Discussion here: https://news.ycombinator.com/item?id=38781941
Re: NY Times copyright suit wants OpenAI to delete all GPT instances
#59The suit demonstrates instances where ChatGTP / Bing Copilot copy from the NYT verbatim. I think it is hard to argue that such copying constitutes "fair use". However, OAI/MS should be able to fix this within the current paradigm: Just learn to recognize and punish plagiarism via RLHF. However, the suit goes far beyond claiming that such copying violates their copyright: "Unauthorized copying of Times Works without p…
Re: NY Times copyright suit wants OpenAI to delete all GPT instances
#60Earlier quoted context omitted.
Precisely. This tired 'fair use' excuses from AI bros whilst the GPT has reproduced the article text verbatim, word for word and it being monetized without the permission from the copyright holder and source (NYT) is an obvious copyright violation 101. Full stop. Again, just like Getty v. Stability, this copyright lawsuit will end in a licensing deal. Apple played it smart with their choice with licensing deals to tr…
> AI bros What (or whom) do you consider to be an "AI bro?" This sort of ad hominem generalization usually accompanies a weak argument.