Live data from Hacker News

OpenAI now tries to hide that ChatGPT was trained on copyrighted books

businessinsider.com

11–20 of 94 posts

Re: OpenAI now tries to hide that ChatGPT was trained on copyrighted books

#11
post #6

I don't get it. Isn't that what people wanted to happen? Don't produce copyrighted work in the output. Sure you can learn from it, much like a director might learn from hundreds of movies he's watched. He obviously can't copy the plot from Die Hard but he can use elements he's picked up from it to make a Christmas movie.

What do you mean he can't copy the plot? Any commercially successful movie is using one of about five basic plots.

Could you explain the five basic plots?

While the number of plot structures are fairly countable, a plot can still be considered different by changing part of its contents (character, setting, etc.). The decision on whether a plot outright infringes on an existing plot varies case-by-case. An example of outright copying is "Fistful of Dollars" directed by Sergio Leone, which lifts the plot from "Yojinbo" by Akira Kurosawa.

Re: OpenAI now tries to hide that ChatGPT was trained on copyrighted books

#13
People keep comparing it to a human ingesting content throughout their life and then being influenced in their own works. I'm sorry but that is not the same thing - at all.

The concept of training a model with the explicit intent of selling the output of that model is inherently different.

Not saying that it should be illegal. But it is clearly in violation of the spirit of existing copyright law, in my opinion.

They set out with the intent to make money, using copyrighted input. Seems pretty simple.

See dragonwriter's comment for the articulate version of what I'm saying

Re: OpenAI now tries to hide that ChatGPT was trained on copyrighted books

#14
post #6

I don't get it. Isn't that what people wanted to happen? Don't produce copyrighted work in the output. Sure you can learn from it, much like a director might learn from hundreds of movies he's watched. He obviously can't copy the plot from Die Hard but he can use elements he's picked up from it to make a Christmas movie.

What do you mean he can't copy the plot? Any commercially successful movie is using one of about five basic plots.

Fair use doctrine has a lot of nuanced case law. I believe fair use has to be "transformative" upon the original, and it also cannot act as a market substitute for the original. E.g. if a new movie is substantially the same as an old movie, not transformative (e.g. like commentary or parody) and could reasonably be viewed as a market substitute for the old movie, then it'd probably be considered unfair use.

Re: OpenAI now tries to hide that ChatGPT was trained on copyrighted books

#16
post #13

People keep comparing it to a human ingesting content throughout their life and then being influenced in their own works. I'm sorry but that is not the same thing - at all. The concept of training a model with the explicit intent of selling the output of that model is inherently different. Not saying that it should be illegal. But it is clearly in violation of the spirit of existing copyright law, in my opinion. They…

It is not illegal to make money with copyrighted content. Nor should it be

Re: OpenAI now tries to hide that ChatGPT was trained on copyrighted books

#17
post #6

I don't get it. Isn't that what people wanted to happen? Don't produce copyrighted work in the output. Sure you can learn from it, much like a director might learn from hundreds of movies he's watched. He obviously can't copy the plot from Die Hard but he can use elements he's picked up from it to make a Christmas movie.

> He obviously can't copy the plot from Die Hard

I dunno, as long as you swap out genre tropes, make slight changes of geography, and/or change a few plot points, coolpying plot is routine in Hollywood.

Re: OpenAI now tries to hide that ChatGPT was trained on copyrighted books

#18
post #13

People keep comparing it to a human ingesting content throughout their life and then being influenced in their own works. I'm sorry but that is not the same thing - at all. The concept of training a model with the explicit intent of selling the output of that model is inherently different. Not saying that it should be illegal. But it is clearly in violation of the spirit of existing copyright law, in my opinion. They…

Yeah and people read copyrighted books so they can get jobs and sell the use of the knowledge they acquired from reading the books.

The only difference is a machine doing it at a larger scale.

Re: OpenAI now tries to hide that ChatGPT was trained on copyrighted books

#19
post #13

People keep comparing it to a human ingesting content throughout their life and then being influenced in their own works. I'm sorry but that is not the same thing - at all. The concept of training a model with the explicit intent of selling the output of that model is inherently different. Not saying that it should be illegal. But it is clearly in violation of the spirit of existing copyright law, in my opinion. They…

It is not illegal to make money with copyrighted content. Nor should it be

If you don't procure the appropriate license for that content then yes, it is (exceptions for things like parody notwithstanding).

For example, clearance of music samples.

Re: OpenAI now tries to hide that ChatGPT was trained on copyrighted books

#20
post #19

Earlier quoted context omitted.

It is not illegal to make money with copyrighted content. Nor should it be

If you don't procure the appropriate license for that content then yes, it is (exceptions for things like parody notwithstanding). For example, clearance of music samples.

So aspiring authors should avoid all reading, because they need licenses for all works that inspired them?
Post reply on HN