Live data from Hacker News

AI is in danger of being swallowed up by copyright law

heathermeeker.com

421–430 of 705 posts

Re: AI is in danger of being swallowed up by copyright law

#421

Earlier quoted context omitted.

> but what the customers of AI people want isn't available under those terms. What the customers of AI want is accurate predictions of the models, and they can get that even if everyone demanding to get removed from the training set would be removed. The makers of generative AI could remove every living artist who wants to from the dataset, the model would still develop a general solution of color theory, composition…

Then do it! I swear when I see this argument because it makes me angry. You’re right, but they didnt , because they were too lazy and cheap to do it that way. …and that’s why people are angry, and rightly so. Fully licensed models are the future, and it’s both irritating and disappointing that we are where we are right now because the people training these models were too lazy to assemble a training dataset that wasn…

No one outside some artists is going to give a fuck about the legality. Joe Schmuck out there is too busy either not knowing this exists or making funny pictures of Han Solo eating a banana on a toilet made from the skin of Yoda.

That reputation damage you think matters doesn’t exist.

Re: AI is in danger of being swallowed up by copyright law

#422

Earlier quoted context omitted.

> Do you ask for permission when you train your mind on copyrighted books yes, thats why I pay a fee to buy/borrow one (or someone pays the fee in the case of a library.) > listen to music again money is exchanged. > Humans are constantly ingesting gobs of "copyrighted" insights that they eventually remix into their own creations without necessarily reimbursing the original source(s) of their creativity. yes, and so…

You don't need to pay for or "borrow" anything to learn from copyrighted works. Nobody has had that expectation for years, and that is also not what copyright pertains to. It's not that AI breaks into libraries and isn't paying the fees. You can google an image of any great work of art and look at it for as long as you like, for free and take from it what you can and use all of that to create something else and get p…

> You can google an image of any great work of art and look at it for as long as you like, for free and take from it what you can and use all of that to create something else and get paid for that thing.

If you're referring to piracy, that is very much being kept in check. Otherwise, the vast majority of copyrighted art is only available for payment in various ways (streaming services, museum and theatre access fees, library cards, buying e-books etc).

Re: AI is in danger of being swallowed up by copyright law

#423

Earlier quoted context omitted.

Better title: The advancement of AI is being slowed by copyright. But "eating" is a fun word.

Even better title: tech giants don't want to pay copyright to small creators, but will easily be coherced to pay it to disney.

It’s up to the courts to decide if this is a copyright infringement. The EU at least already allows the use of copyrighted material for research purposes into text and data mining, so the main question will hinge on whether or not the result of such research can be commercially exploited.

Re: AI is in danger of being swallowed up by copyright law

#424
post #369

Earlier quoted context omitted.

I can't help but feel like you're slightly anthropomorphising an algorithm. It's a really damn cool and powerful algorithm, don't get me wrong, but it's still not a person. At the end of the day, it's also not the algorithm "benefiting" from it, but the corporation using the algorithm. It's also a bit hypocritical, because if you did the same thing to them as a human (in this example let's say be a Tesla copycat) you…

The anthropomorphisation of the ML baffles me. This is not AI but a large ML model trained on often proprietary data. The whole discussion is ridiculous. There should be a lineage of model data so we should know what's the original source (at least for the "core" of the answer) in form of citations. I assume that the model loose that information during the training and that would be quite hard as everything is mixed…

Is it really that baffeling? Take ChatGPT, a model specifically fine tuned on dialog interactions to seem more human like. Of course people are going to anthropomorphise it. And the whole thing about it not being AI but an ML model, I think we can let that one go. It didn't stop the term cloud computing ("it's just someone else's server") and it won't stop the term AI from being used for this tech.

Re: AI is in danger of being swallowed up by copyright law

#425

There's no part of AI that is being swallowed up by copyright. AI companies can ask for permission if they want to train their models on other people's works. It's not that hard, various image hosting sites have already added an opt-in/opt-out toggle to their services. Sites might even get away with using this stuff as compensation for free hosting. The fact of the matter is that the AI companies don't want to ask fo…

Regardless of whether one agrees or not with paying creators of the training data, I think the deeper issue here is about societal wealth distribution and who gets paid for X now that X is being done very well by AIs. A less equitable world has Google or billionaires getting paid. A more equitable world has the artists. But I want to argue here that for purposes of this latter question, your proposal of copyright enf…

> - Even if you get the system to work, what about future artists and writers? Are we just creating an entrenched historical group of creatives getting royalties forever?

The boat has long since sailed on this… ands it’s globally entrenched as a norm of international trade that we are all “ok with this” regime of 75 years or century plus copyright terms …

And arguably the entire copyright vs AI/ML training datasets debate is founded on the notion that the artists individual copyright will last long enough that it’s going to outlive the average artist. If we look at one of the old copyright regimes, for comparison… in a world where copyright is a short default/implicit/automatic term (14 or 28 years) and the copyright owner can elect to register and pay for extensions (for a more modern twist, preferably combined with increasing incentive to prevent perpetual renewal abuses by Disney, et al)… now imagine how much data from up to 28 years ago there is, the catalogue of art and photographs and text and books and academic writings… all public domain because the authors didn’t consider them of sufficient value… all free for the ML model training… this gets even larger with a 14 year term…

Suffice to say that we are seeing systemic impacts already, culturally we’re seeing more and more money put behind less and less content controlled by fewer and fewer people due to a slow death spiral off copyright stranglehold across multiple industries, written, visual, audio and video arts are all dominated by large corporations holding IP … yes individuals continue to create, but other than rare breakthrough chance successes and internet age viral success (which are often just completely arbitrarily/random and have no real quality) these companies decide what will be popular culture…

My prediction is that the AI/ML models will be allowed but heavily scrutinised, under the simple legal doctrine that the user is the one committing the infringement since the primary purpose of these models is not infringement but unique creation, but suspicion will linger by artists and it will become a normal part of contracts in the art word…effectively an artist equivalent of the way police in many places view spray cans… just as the primary purpose of spray paint is not to create illegal graffiti, which is the justification many places used to overturn poorly justified civic bans on possession of spray paint.

I’d like to see any more draconian spread of derivative work rights (style rights etc) to be accompanied by drastic reductions in the automatic copyright term, as the ability to churn out lots of automatic content drastically lowers the value of long long terms, and the counter argument that it makes the existing rights more valuable is fucking insane as we do not need to pass copyright down to the great-great-great-great-grandchildren… the terms are already too long.

Re: AI is in danger of being swallowed up by copyright law

#426

Earlier quoted context omitted.

You don't need to pay for or "borrow" anything to learn from copyrighted works. Nobody has had that expectation for years, and that is also not what copyright pertains to. It's not that AI breaks into libraries and isn't paying the fees. You can google an image of any great work of art and look at it for as long as you like, for free and take from it what you can and use all of that to create something else and get p…

> What if humans from here on will only be paid to create stuff that an AI can't? I look forward to a life of horrific poverty

That's either a very pessimistic take on human creativity, or a very optimistic take on the ability of AI to mimic human emotion and experience.

Re: AI is in danger of being swallowed up by copyright law

#427

I think the underlying question is one of "degrees of derivation". There's a famous Carl Sagan quote: “If you wish to make an apple pie from scratch, you must first invent the universe” which hints at the problem: Nothing is created in a vacuum. Let's compare what Stable Diffusion does with what Franz von Holzhausen, head of design at Tesla, does. Franz didn't come into existence out of nothing and knew how to design…

I can't help but feel like you're slightly anthropomorphising an algorithm. It's a really damn cool and powerful algorithm, don't get me wrong, but it's still not a person. At the end of the day, it's also not the algorithm "benefiting" from it, but the corporation using the algorithm. It's also a bit hypocritical, because if you did the same thing to them as a human (in this example let's say be a Tesla copycat) you…

I think the parent poster has the right idea. This is not just anthropomorphization; it's an analogy.

They take away is contained in the first and last sentences.

Derivation is key.

Copyright protects original expressions, and copying means to reproduce (read and write) something. The analogy OP made is focused on the reading and writing done by humans and the reading and writing done by an algorithm.

Algos like stable diffusion are reading, and their user is controlling what is written.

If the user produces a work that is unique, but uses the style of a particular artist, that seems like it should be valid, since style is simply a process. It's how to create art, but it is not art, and processes are not subject to copyright based on copyright.gov.

With all that said legality isn't morality, and I sympathize with the artist.

Re: AI is in danger of being swallowed up by copyright law

#428

Earlier quoted context omitted.

You don't need to pay for or "borrow" anything to learn from copyrighted works. Nobody has had that expectation for years, and that is also not what copyright pertains to. It's not that AI breaks into libraries and isn't paying the fees. You can google an image of any great work of art and look at it for as long as you like, for free and take from it what you can and use all of that to create something else and get p…

> You can google an image of any great work of art and look at it for as long as you like, for free and take from it what you can and use all of that to create something else and get paid for that thing. If you're referring to piracy, that is very much being kept in check. Otherwise, the vast majority of copyrighted art is only available for payment in various ways (streaming services, museum and theatre access fees,…

They're talking about googling any copyrighted image and looking at it, as a human or an AI.

Re: AI is in danger of being swallowed up by copyright law

#430

Earlier quoted context omitted.

That's what I am saying. Stop doing that. Train co-pilot, or whatever, to follow the terms of the license.

That's tough when Jimmy is using copilot to generate code for his proprietary company codebase. AI is just a tool, do you also sue the company that sold the paintbrush with which an infringing painting was made?

The thing is, we aren't really talking about AI here. We are talking about datasets being used to train AI and the companies offering the services of a trained AI. If a company trained their own AI on their own code, the discussion would be very different (likely centred on moral issues rather than legal ones). We probably won't ever have a discussion about companies submitting proprietary code to train an AI created by a third-party since there would be a contract between the two parties (and it is unlikely a company will submit proprietary code to train an AI that will be used by others in the first place).

What we are looking at is a kin to a company that makes paint brushes, trains graphics artists, and contracts out their graphics artists to use those paint brushes. If it turned out that the graphics artists turned out derivative works without the rights holder's permission, you can be assured that people would want to "sue the company that sold the paintbrush".

I'm not going to claim that my example is equivalent to what is happening with these AI services. And while you may be right about AI fundamentally being a tool, like a paint brush, I would suggest that is only true if you ignore the data that is fed into it.

Post reply on HN