Live data from Hacker News

Google denies training Bard on ChatGPT chats from ShareGPT

twitter.com

121–130 of 342 posts

Re: Google denies training Bard on ChatGPT chats from ShareGPT

#123

Earlier quoted context omitted.

Our writing, our code, our artwork... Furthermore, the U.S. Copyright Office (USCO) concluded that AI-generated works on their own cannot be copyright, so these ChatGPT logs are free game. It would be hypocritical to think that Google is wrong and OpenAI is not.

its not even that on their own those works cant be copywritten. its that even when you make changes to those works, your changes might qualify for copyright but they do not affect the copyright status of the ai generated portions of the work. if you used ai to design a new superhero and then added pink shoes, yellow hair, and a beard, only those three elements would possibly be able to be protected by copywrite. your…

How could that be ever really be enforceable?

If I use an AI tool to design my Superhero, can't I just submit it without disclosing the help I received from an AI.

I get that it would be very nice to prevent AI SPAM copyrighting of every possible superhero, but if I use the AI to come up with a concept, then quickly redraw it myself with pen and paper, I feel like it would never be provable that it came from an AI.

Re: Google denies training Bard on ChatGPT chats from ShareGPT

#126
post #88

1. Google denies doing it, so at the very least the title should have an "allegedly". 2. Even if they did – so what? The output from ChatGPT is not copyrightable by OpenAI. In fact it is OpenAI that is training its models on copyrighted data, pictures, code from all over the internet.

But remember many years back when it was news that Bing used Google search results to improve its results.

> Google catches Bing copying [search results], Microsoft says “so what?”

https://arstechnica.com/information-technology/2011/02/googl...

Re: Google denies training Bard on ChatGPT chats from ShareGPT

#128
post #125

I love that OpenAI uses a ton of other peoples work to train their model, yet when someone uses OpenAI to train their model, they get all up in arms. As far as I'm concerned, OpenAI has decided terms of use don't exist anymore.

Where are they up in arms?

Re: Google denies training Bard on ChatGPT chats from ShareGPT

#129
post #110

I don't care at all about this from a copyright or data ownership perspective, but I am a little skeptical that it's a good idea to be this incestuous with training data in the long run. It's one thing to do fine tuning or knowledge distillation for specialized domains or shrinking models. But if you're trying to train your own foundation model, is relying on output from other foundation models going to make them lea…

Where are any LLMs going to get data from as they become more ubiquitous and humans produce less publicly accessible original and thoughtful content? The whole thing is a plateaued feedback loop.

It'd be cool to have an LLM that's trained almost exclusively on books from good publishers, and other select sources. Working out licensing deals would be a challenge, of course.

Re: Google denies training Bard on ChatGPT chats from ShareGPT

#130

Earlier quoted context omitted.

They literally copied the Chatgpt UI, lol, only it looks like a dated Google UI. How do you prefer answers with less data?... that's crazy.

doing a visual diff will show you it's not a literal copy

I'm talking design, not code, lol...
Post reply on HN