Earlier quoted context omitted.
Exactly. A lot of the difficulty here is how they skip is the hugely important issue: An entirely reasonable, if not fully tested, statement is the following: Every single one of these AI weight things itself is a result of unencumbered, massive, law-breaking, right-violating copyright infringement -- accordingly, it's extremely difficult to say anything morally justifiable or authoritative about anyone elses "rights…
> is a result of unencumbered, massive, law-breaking, right-violating copyright infringement Why? Copyright covers expression not information, AIs can learn information from any source regardless of copyright. They should just not regurgitate copyrighted content, that's all. And much of what organic content is online is common knowledge, thus can't be copyright-controlled.
AI weights are not open “source”
91–100 of 274 posts
Re: AI weights are not open “source”
#92Earlier quoted context omitted.
IP itself violates a natural right (Yes the idea of rights is also unnatural and absent from visions such as anarchy)
Yeah I didn't even think that was controversial. I'd always been taught that copyright and patents exist to explicitly restrict what people can do by granting a monopoly to the owners in order to encourage invention and creative work. Edit to add I'm not saying I agree with the justification or am trying to argue for it, only that the point above is commonly raised as the justification, implying that the intrusion on…
Re: AI weights are not open “source”
#93Earlier quoted context omitted.
Exactly. A lot of the difficulty here is how they skip is the hugely important issue: An entirely reasonable, if not fully tested, statement is the following: Every single one of these AI weight things itself is a result of unencumbered, massive, law-breaking, right-violating copyright infringement -- accordingly, it's extremely difficult to say anything morally justifiable or authoritative about anyone elses "rights…
If I collect a set of copyright free data or public domain data would we conclude that the weights are also public domain?
Re: AI weights are not open “source”
#94Earlier quoted context omitted.
Why is it unestablished? Is a document not copyrightable based on its contents? Weights are just a different kind of a document.
Copyright is not for "documents", it is for works that have creativity in them. The legal bar for that level of creativity is low, so low that it is easy to come away thinking that anything that can be cast as a "document" must be copyrightable, but the bar is in fact not zero. In particular, taking other documents and shoving them through a process that generates a lot of other numbers with no human or creative inte…
Re: AI weights are not open “source”
#95Earlier quoted context omitted.
IP itself violates a natural right (Yes the idea of rights is also unnatural and absent from visions such as anarchy)
Natural rights are a fiction to pretend that someone’s moral code is a privileged aspect of physical reality in a way every competing moral code is not.
Re: AI weights are not open “source”
#96Earlier quoted context omitted.
> Every single one of these AI weight things itself is a result of unencumbered, massive, law-breaking, right-violating copyright infringement Maybe the popular and free ones. Adobe has a product in beta that uses "ethical training data" as a selling point.
Interesting. I wonder what they mean by "Ethical" -- instead of e.g. saying "definitely free and open." I'm willing to bet "stuff they gathered from likely unwitting Adobe users."
Re: AI weights are not open “source”
#97Earlier quoted context omitted.
These are not bad arguments, but I don't think they're conclusive. I am a lawyer, and I could absolutely see this going the other way. "You can't make these machine things without literally feeding this copyrighted information into them, therefore they do contain a copy. You can see this by when they reproduce, e.g. the "getty images" deal." *this is not legal advice, dangit commenter person below
It's just a race for which test case gets to the supreme court first really...
They could be completely biased, could completely ignore everyone and everything else and rule however they want.
I'm almost surprised they still bother to write any kind of "legal reasoning" in their ruling and don't simply focus on what the ruling is rather than why they ruled that way. But I guess such "reasoning" still serves a propaganda purpose and still provides a fig leaf for those who still believe in the quaint absurdity that "we are a nation of laws, not men."
Re: AI weights are not open “source”
#98Earlier quoted context omitted.
That's also my understanding, either the weights are copyrightable and then all the models need explicit agreements for any work they include in it because models become derivatives or they are not copyrightable being just machine data (the most likely scenario in my opinion), they can't have it both ways.
I think there could be an argument that it's copyrightable but not a derivative work. If I read a few books about a subject as research, and then I write an article about the subject, it's my own copyright. The fact that I did research doesn't make it derivative of those books (correct me if I'm wrong, IANAL). Perhaps a model created from copyrighted material be treated in the same way?
Re: AI weights are not open “source”
#99Earlier quoted context omitted.
No, that doesn’t follow at all. The argument is that either the training or the expression violated existing cooyrights through the making of unlicensed copies. It’s not based on open source licensing. Although OSS viral licensing may well apply if fair use is not a successful defense.
What copyright is violated by training on public domain data?
Re: AI weights are not open “source”
#100Earlier quoted context omitted.
Exactly. A lot of the difficulty here is how they skip is the hugely important issue: An entirely reasonable, if not fully tested, statement is the following: Every single one of these AI weight things itself is a result of unencumbered, massive, law-breaking, right-violating copyright infringement -- accordingly, it's extremely difficult to say anything morally justifiable or authoritative about anyone elses "rights…
> is a result of unencumbered, massive, law-breaking, right-violating copyright infringement Why? Copyright covers expression not information, AIs can learn information from any source regardless of copyright. They should just not regurgitate copyrighted content, that's all. And much of what organic content is online is common knowledge, thus can't be copyright-controlled.
I suspect we're going to see the same kind of rethink about intellectual property in the age of AI.