Live data from Hacker News

AI Data Laundering

waxy.org

1–10 of 120 posts

Re: AI Data Laundering

#2
This reminds me of the Jedi Mind trick of Uber of waving a smartphone to argue that labor & other laws all of a sudden don't apply to them, to the detriment of the public that'll now shoulder the costs.

Re: AI Data Laundering

#3
The Authors Guild v Google decision about Google Books seems relevant:

> In late 2013, after the class action status was challenged, the District Court granted summary judgement in favor of Google, dismissing the lawsuit and affirming the Google Books project met all legal requirements for fair use. The Second Circuit Court of Appeal upheld the District Court's summary judgement in October 2015, ruling Google's "project provides a public service without violating intellectual property law." The U.S. Supreme Court subsequently denied a petition to hear the case.

[...]

> The court's summary of its opinion is:

[...]

> Google’s unauthorized digitizing of copyright-protected works, creation of a search functionality, and display of snippets from those works are non-infringing fair uses. The purpose of the copying is highly transformative, the public display of text is limited, and the revelations do not provide a significant market substitute for the protected aspects of the originals. Google’s commercial nature and profit motivation do not justify denial of fair use.

https://en.wikipedia.org/wiki/Authors_Guild,_Inc._v._Google,....

This doesn't touch on the ethics of course – at minimum I think allowing people to exclude themselves or their work from a dataset is necessary.

Re: AI Data Laundering

#4
post #3

The Authors Guild v Google decision about Google Books seems relevant: > In late 2013, after the class action status was challenged, the District Court granted summary judgement in favor of Google, dismissing the lawsuit and affirming the Google Books project met all legal requirements for fair use. The Second Circuit Court of Appeal upheld the District Court's summary judgement in October 2015, ruling Google's "proj…

> I think allowing people to exclude themselves or their work from a dataset is necessary.

or they could open it all up for everybody and stop protecting the rights of death people (authors dead less then 70 years ago)

then again, that will make the publishers starve... but why pretend publishing corporations need food?

Re: AI Data Laundering

#5
post #4
post #3

The Authors Guild v Google decision about Google Books seems relevant: > In late 2013, after the class action status was challenged, the District Court granted summary judgement in favor of Google, dismissing the lawsuit and affirming the Google Books project met all legal requirements for fair use. The Second Circuit Court of Appeal upheld the District Court's summary judgement in October 2015, ruling Google's "proj…

> I think allowing people to exclude themselves or their work from a dataset is necessary. or they could open it all up for everybody and stop protecting the rights of death people (authors dead less then 70 years ago) then again, that will make the publishers starve... but why pretend publishing corporations need food?

This is larger than publishers, this is every artist, film-maker, photographer, every writer, every engineer, anybody who has ever written or created something and shared it publicly is liable to have their work assimilated and an infinite amount of derivatives produced with no control over how they're used and by whom.

Comment generated with gpt-neox prompt: Comment about AI and data collection and generation and its pitfalls, expressing concern, emphasis on professions, emphasis on automation, written by Stephen King, creative writing, award winning, trending on reddit, trending on hacker news, written by Greg Rutkowski, written by Zola, written by Voltaire, written by authpor, written by moyix.

(Just kidding, it wasn't AI generated but you see my point.)

Re: AI Data Laundering

#6
post #3

The Authors Guild v Google decision about Google Books seems relevant: > In late 2013, after the class action status was challenged, the District Court granted summary judgement in favor of Google, dismissing the lawsuit and affirming the Google Books project met all legal requirements for fair use. The Second Circuit Court of Appeal upheld the District Court's summary judgement in October 2015, ruling Google's "proj…

Do we allow artists to withhold their works from the minds of eager, learning children? [1]

Tell me how ML is different than the mind of a toddler ravenous for new information.

For every billion dollar start-up using data at scale, there are tens of thousands more researchers and hobbyists doing the exact same, producing wonderful results and advances.

If we stop this growth dead in the tracks, other countries more willing to look past the IP laws will jump ahead. And if Stability locks away their secret sauce, some new party will come and give away the keys to the kingdom yet again.

You can't block the signal. Except, of course, by legislating against it in some Luddite hope we can prevent the future from happening.

Instead of worrying careers will end, we should look at this as being the end of specialization. No longer do we need to pay 20,000 hours to learn one thing to the exclusion of all others we would like to try. Now we'll be able to clearly articulate ourselves with art, music, poetry. We'll become powerful beings of thought and expression.

Humans aren't the end or the peak of evolution. We should be excited to watch this unfold.

[1] Maybe Disney would like you to pay more for a premium learning plan for your child, but thankfully that's not (yet) possible.

Re: AI Data Laundering

#7
post #3

The Authors Guild v Google decision about Google Books seems relevant: > In late 2013, after the class action status was challenged, the District Court granted summary judgement in favor of Google, dismissing the lawsuit and affirming the Google Books project met all legal requirements for fair use. The Second Circuit Court of Appeal upheld the District Court's summary judgement in October 2015, ruling Google's "proj…

> the revelations do not provide a significant market substitute for the protected aspects of the originals

It does seem like generative AI systems provide a significant market substitute, so this ruling probably wouldn’t apply, in court.

edit: see https://news.ycombinator.com/item?id=33194623 for some initial thoughts on how this problem (and others) could be rectified.

For example, with a database of protected works and self-censorship algorithms for generative AI systems, conscientiously objecting creatives could have a mechanism for excluding their works.

Re: AI Data Laundering

#8
post #4

Earlier quoted context omitted.

> I think allowing people to exclude themselves or their work from a dataset is necessary. or they could open it all up for everybody and stop protecting the rights of death people (authors dead less then 70 years ago) then again, that will make the publishers starve... but why pretend publishing corporations need food?

This is larger than publishers, this is every artist, film-maker, photographer, every writer, every engineer, anybody who has ever written or created something and shared it publicly is liable to have their work assimilated and an infinite amount of derivatives produced with no control over how they're used and by whom. Comment generated with gpt-neox prompt: Comment about AI and data collection and generation and it…

> anybody who has ever written or created something and shared it publicly is liable to have their work assimilated and an infinite amount of derivatives produced with no control over how they're used and by whom.

This has been the case ever since people started putting their art on the Internet publicly. The only difference is that now it's algorithms creating the derivatives, not people.

Re: AI Data Laundering

#9
post #3

The Authors Guild v Google decision about Google Books seems relevant: > In late 2013, after the class action status was challenged, the District Court granted summary judgement in favor of Google, dismissing the lawsuit and affirming the Google Books project met all legal requirements for fair use. The Second Circuit Court of Appeal upheld the District Court's summary judgement in October 2015, ruling Google's "proj…

> the revelations do not provide a significant market substitute for the protected aspects of the originals It does seem like generative AI systems provide a significant market substitute, so this ruling probably wouldn’t apply, in court. edit: see https://news.ycombinator.com/item?id=33194623 for some initial thoughts on how this problem (and others) could be rectified. For example, with a database of protected work…

A substitute for what though? Copyright law is only concerned with substituting the work under copyright. That is to say, the consideration is whether the infringing aspects of the secondary work would alter the demand and market for the work being infringed.

In all the talk about AI data laundering there really hasn't been any indication that the AI generated item substitutes for the item it's alleged to infringe on. Substituting for a whole profession and its practitioners doesn't enter into the concerns of copyright law. There might be some argument that it should (to "promote the progress of science and useful arts" as it were), but copyright law to my knowledge hasn't been used to prevent new tech from putting professionals as a whole out of business.

Re: AI Data Laundering

#10
post #4

Earlier quoted context omitted.

> I think allowing people to exclude themselves or their work from a dataset is necessary. or they could open it all up for everybody and stop protecting the rights of death people (authors dead less then 70 years ago) then again, that will make the publishers starve... but why pretend publishing corporations need food?

This is larger than publishers, this is every artist, film-maker, photographer, every writer, every engineer, anybody who has ever written or created something and shared it publicly is liable to have their work assimilated and an infinite amount of derivatives produced with no control over how they're used and by whom. Comment generated with gpt-neox prompt: Comment about AI and data collection and generation and it…

this is larger than the arts. anybody has ever participated creatively in our culture understands that it's absolute bullshit to pretend we need money in order to want to contribute artistically.

we need money because food is for sale, because most of us do not own where we live hence we are forced (a priori) to come up with a whole lot of money every month or else you're out in the streets.

Post reply on HN