Live data from Hacker News

David Guetta uses ChatGPT and uberduck.ai to deepfake Eminem rap for DJ set

twitter.com

121–130 of 180 posts

Re: David Guetta uses ChatGPT and uberduck.ai to deepfake Eminem rap for DJ set

#121

Earlier quoted context omitted.

What song did Eminem write that he would be owed loyalties for?

His entire discography, the thing that ChatGPT's impersonation of him is based on.

The voice synthesis models can be done with as little as a few seconds of someone speaking. Not an entire discography.

Re: David Guetta uses ChatGPT and uberduck.ai to deepfake Eminem rap for DJ set

#122

Oh -- so when David Guetta does it on stage it's interesting now? :P https://reticulated.net/dailyai/gpt3-uberduck-linux-gangster...

I will also say -- UberDuck and AI TTS in general, when compared to the SURGE of development and tools that's happened on the image/video side of AI, is TERRIBLE.

UberDuck's community specifically seems geared towards kids making memes -- I suspect they just ended up there and didn't design it that way, but wading through the terrible user created models to find ones that work was tiresome.

I tried to get https://coqui.ai/ setup to do similar things, but have not been successful.

Surely this will all explode in the next 18 mo max

Re: David Guetta uses ChatGPT and uberduck.ai to deepfake Eminem rap for DJ set

#123

Earlier quoted context omitted.

While I think the case you're presenting is very likely, I don't think Sony v Universal applies too significantly. The VCR was demonstrably not in any way based on Universal's IP - it was only a tool that could be used to create copies of Universal's IP. In their capacity as tools, LLMs probably would fall under a similar model. However, the LLM itself is substantially based on the IP of all of the creators of the da…

Sony v Universal established a very important legal doctrine with regards to "commercially significant non-infringing use". You can take my word for it, you can go an do your own research, you can confirm with an IP lawyer, or you can wait for the the court's opinion. Or I guess you can give me a little bit of time to go and help you do some of your own research, which I will do right now, so just hold on a bit! I'm…

Edit: after the text you added, I believe we are essentially in agreement. IF it is accepted that the way GPT3 was created is a fair use of the works in the training data, THEN I fully agree that (1) OpenAI has every right to sell it even if (2) some uses of it would still constitute copyright infringement, since (3) only specific users would be liable for copyright infringement in their uses.

Where we differ is in how certain we are that the IF is true. I for one believe there is a good chance that the training of an LLM on copyrighted works does infringe on the copyright of those works (if no other exceptions apply, such as the LLM being trained only for academic research purposes, of course).

My original response:

Where Sony v Universal definitely applies though is when evaluating whether OpenAI's selling of GPT3 to others who then use it to create copyright-infringing works would make OpenAI liable for contributory infringement. Here, the similarities are crystal clear, and the conclusion is simple: since there clearly exist non-infringing uses of GPT3 (such as Supabase Clippy), OpenAI is fully in the clear to sell GPT3, just as much as Sony was for selling the VCR.

However, this assumes that OpenAI has the rights to the IP of GPT3 itself in the first place, which is a prerequisite to them being allowed to sell it at all. Sony certainly had the rights to the IP of the VCR - Universal never claimed that the VCR was a derivative work of their movies.

Essentially, in Sony v Universal, Universal was claiming (1) that Sony was liable for contributory infringement, since (2) all customers of Sony who used it to record and then playback a Universal show were guilty of copyright infringement. The court established that (2) was in fact fair use, and from there automatically (1) become false, since now there was an established legal way of using the Sony product.

But, in a hypothetical OpenAI v Universal, Universal could plausibly claim that (1) OpenAI is liable for copyright infringement directly, since they are distributing GPT3 , (2) which is a derived work of Universal's IP used in the training set of GPT3.

Re: David Guetta uses ChatGPT and uberduck.ai to deepfake Eminem rap for DJ set

#124

Earlier quoted context omitted.

In a way, it's similar to the actors selling their digital images in the movie The Congress. It's inevitable.

wow, not many people have seen this movie. 100%.

I'd never heard of it, but definitely going to check it out now. A lot has happened in the ten years since it was created.

Re: David Guetta uses ChatGPT and uberduck.ai to deepfake Eminem rap for DJ set

#128
I believe David Guetta that he was really just playing around with these cool tools here (who isn't blown away by them?). But I guess it won't go unnoticed by the teams and investors of uberduck and other AI startups that this is the perfect guerilla marketing stunt.

Gotta admit it, I am guilty too, never heard of uberduck before and caught myself creating an account and browsing their pricing site today.

Re: David Guetta uses ChatGPT and uberduck.ai to deepfake Eminem rap for DJ set

#129

*david guetta's ghost producer

What claim do you have to say that? Here is a 30-minute lecture from the guy recording himself producing a track: https://www.youtube.com/watch?v=LfEhLdITOac

> What claim do you have to say that? Here is a 30-minute lecture from the guy recording himself producing a track: https://www.youtube.com/watch?v=LfEhLdITOac

Joachim Garraud, among others, has been producing David Guetta's tracks for 20 years.

Re: David Guetta uses ChatGPT and uberduck.ai to deepfake Eminem rap for DJ set

#130

The video reference the common ChatGPT "in the style of" prompt. Is there a list of what "styles" ChatGPT's model has been trained on ? Or is this information not disclosed since it would be an admission that ChatGPT has been trained on copyrighted lyrics, etc. by which is produces these derivative results ?

Remember it is just a stochastic model. The prompt "in the style of" + any word just makes it more likely that certain words will be predicted next.

If it never heard the name Eminem or read any of the lyrics it couldn't predict anything similar.

Post reply on HN