Live data from Hacker News

Japan’s government will not enforce copyrights on data used in AI training

technomancers.ai

151–160 of 426 posts

Re: Japan’s government will not enforce copyrights on data used in AI training

#151

If I compress an artist's painting into a jpeg and rehost part of it for individual t-shirt designs I am committing a crime. If I compress an artist's painting into a model & rehost what's essentially a highly flexible complete version of their painting for infinite, perpetual use of any kind ... I'm not committing a crime?

>compress an artist's painting into a model

That's not how image models work.

Re: Japan’s government will not enforce copyrights on data used in AI training

#152

Japan also ranks 3rd (behind the USA & India, with larger populations) in ChatGPT usage: https://www.demandsage.com/chatgpt-statistics/ There's also been discussion of their government using ChatGPT to reduce red tape: https://www.bloomberg.com/news/articles/2023-04-18/japan-gov... It's cool to see Japan and Japanese culture taking techno-optimist stances on AI.

Which I find bizarre given how backwards Japan is in the adoption of other technologies. Eg their continued reliance on paper records and fax machines.

Re: Japan’s government will not enforce copyrights on data used in AI training

#153

If I compress an artist's painting into a jpeg and rehost part of it for individual t-shirt designs I am committing a crime. If I compress an artist's painting into a model & rehost what's essentially a highly flexible complete version of their painting for infinite, perpetual use of any kind ... I'm not committing a crime?

>compress an artist's painting into a model That's not how image models work.

Painting features => back propagation => weights.

Yes it is.

Re: Japan’s government will not enforce copyrights on data used in AI training

#154
post #101

Earlier quoted context omitted.

> But you will find no copy of the logo in the model data. You wont find a copy of a plaintext in a cyphertext. But you can still extract the plaintext from the cyphertext.

If you search "Superman Logo" you find actual copies of the Superman logo which are served from Google's cache. If you ask a VFX artist to create the "Superman Logo" with Photoshop they'll do an excellent job. The first one isn't copyright violation because it is fair use. The second maybe if it is redistributed but we don't ban the use of photoshop by artists because they can choose to reproduce copyright things wit…

I agree, and I honestly think that a big part of the issue with AI image generation is people just really have a hard time conceiving of a technology that can make such accurate images from a relatively small model like this.

"It must have a copy" - but van Gogh didn't make paintings of hot rods or whatever, and you can't copyright style or technique.

Re: Japan’s government will not enforce copyrights on data used in AI training

#155
Not trying to express an opinion on the legal matter, but as a technical matter it's pretty obvious that LLMs create copies of (some of) their training data.

Here's GPT-3.5 reciting the Declaration of Independence: https://chat.openai.com/share/eb30c373-7fec-4280-892d-479567...

Unless you're claiming that GPT-3.5 is deriving the Declaration of Independence (from information about the founding fathers?) I don't see how there's room for debate about whether information has been "copied" into the model.

I have done this test in the past with copyrighted material (harry potter) but they have since added safeguards against it, but my understanding is that the model is still capable of it.

Re: Japan’s government will not enforce copyrights on data used in AI training

#156

What's surprising here is that Japan is usually crazy gung ho on copyright enforcement ... at least against individuals. So it's kind of disgusting to see this relaxation, when it suits some corporate or national interests. https://en.wikipedia.org/wiki/File_sharing_in_Japan "Unlike most other countries, filesharing copyrighted content is not just a civil offense, but a criminal one, with penalties of up to ten years…

[flagged]

Re: Japan’s government will not enforce copyrights on data used in AI training

#157

If I compress an artist's painting into a jpeg and rehost part of it for individual t-shirt designs I am committing a crime. If I compress an artist's painting into a model & rehost what's essentially a highly flexible complete version of their painting for infinite, perpetual use of any kind ... I'm not committing a crime?

>compress an artist's painting into a model That's not how image models work.

It has been shown that image models can produce originals, or at least extremely close to the originals. If the outcome is the same, what is the difference between compression/decompression vs training/generation regarding copyright?

Re: Japan’s government will not enforce copyrights on data used in AI training

#158

If I compress an artist's painting into a jpeg and rehost part of it for individual t-shirt designs I am committing a crime. If I compress an artist's painting into a model & rehost what's essentially a highly flexible complete version of their painting for infinite, perpetual use of any kind ... I'm not committing a crime?

What I wonder is:

Load the source code to all versions of unix, with all licenses.

"write me a version of unix"

Since there is no model copyright and the result was written by AI the software is now in the public domain.

Re: Japan’s government will not enforce copyrights on data used in AI training

#159
post #155

Not trying to express an opinion on the legal matter, but as a technical matter it's pretty obvious that LLMs create copies of (some of) their training data. Here's GPT-3.5 reciting the Declaration of Independence: https://chat.openai.com/share/eb30c373-7fec-4280-892d-479567... Unless you're claiming that GPT-3.5 is deriving the Declaration of Independence (from information about the founding fathers?) I don't see ho…

Pretty good argument but it has one fatal flaw. People can memorize the Declaration of Independence too. Or Harry Potter. If people mostly recite HP from memory but apply enough creative changes, it's not copyright infringement.

So proving a system can memorize and recite proves nothing.

Re: Japan’s government will not enforce copyrights on data used in AI training

#160

So if you train an audio model on say, Eminem's voice, then write some songs and have it perform them...Would this output be legal to publish?

What's the difference between that and someone else who just happens to sound like Eminem in terms out output? As long as you don't market yourself as Eminem that should be completely legal.

well, cover bands have to abide by copyright.
Post reply on HN