Live data from Hacker News

Google denies training Bard on ChatGPT chats from ShareGPT

twitter.com

251–260 of 342 posts

Re: Google denies training Bard on ChatGPT chats from ShareGPT

#251

Earlier quoted context omitted.

Not to mention it's embarrassing. Google playing second banana to OpenAI.

I think Amazon was first in the (free) banana business

you joke, but first producy they changed on whole foods were the bananas.

before: organic (south america) and regular (central ou SEA) for 69, 59.

then: both chikita's brand with regular and organic stickers (clearly the same produce, always from SEA) for 49 and 39 cents.

thats was days after the announcement

Re: Google denies training Bard on ChatGPT chats from ShareGPT

#252
post #186

Earlier quoted context omitted.

Two wrongs don’t make a right.

forgive me if i have limited sympathy when a burglars house gets robbed

It's even less worthy of sympathy - like a counterfeit piece of art being counterfeited. And there isn't even an original, just like a made up counterfeit.

Re: Google denies training Bard on ChatGPT chats from ShareGPT

#253
post #88

1. Google denies doing it, so at the very least the title should have an "allegedly". 2. Even if they did – so what? The output from ChatGPT is not copyrightable by OpenAI. In fact it is OpenAI that is training its models on copyrighted data, pictures, code from all over the internet.

Regarding point 2, I think there's nothing "wrong" with it, mainly it's funny that they don't know how to do it themselves. Provides additional evidence that Google is outgunned in this fight.

Yup

The idea of doing this is embarrassing enough for Google.

Google index the whole web, some of the documents are due to be generated by ChatGPT, there is no way around it.

Re: Google denies training Bard on ChatGPT chats from ShareGPT

#255
post #88

1. Google denies doing it, so at the very least the title should have an "allegedly". 2. Even if they did – so what? The output from ChatGPT is not copyrightable by OpenAI. In fact it is OpenAI that is training its models on copyrighted data, pictures, code from all over the internet.

Ok, I've added that information to the title—thanks. There's also https://www.theverge.com/2023/3/29/23662621/google-bard-chat....

Unfortunately the original report (https://www.theinformation.com/articles/alphabets-google-and...) is hardwalled.

Re: Google denies training Bard on ChatGPT chats from ShareGPT

#256
It’s interesting when we say Google did this. It’s actually and likely some people that work for Google and are on this forum did this. Knowingly, not by accident while slurping up the rest of the internet, and got paid to do it. I wonder what the engineer view on this was/is. I have to assume they ballpark know the terms of the openai data (regardless if you disagree or not).

Anyone care to steel man the argument for why this was a good idea?

Re: Google denies training Bard on ChatGPT chats from ShareGPT

#257

It’s interesting when we say Google did this. It’s actually and likely some people that work for Google and are on this forum did this. Knowingly, not by accident while slurping up the rest of the internet, and got paid to do it. I wonder what the engineer view on this was/is. I have to assume they ballpark know the terms of the openai data (regardless if you disagree or not). Anyone care to steel man the argument fo…

I don't understand why it's a bad idea? Did openai ask for permission for using the data it uses (no)?

Re: Google denies training Bard on ChatGPT chats from ShareGPT

#259

Earlier quoted context omitted.

OpenAI Terms of service forbid training competitor models via their ML outputs (LoRa alpaca laundering is probably not allowed for commercial use).

Google has no contract with OpenAI though. They used a third party site to scrape conversations. If the outputs themselves are not copyrighted, and they never agreed to the terms of service, it should be fine, right? Albeit unethical and embarrassing.

No more unethical or embarrassing than scraping the web for millions of copyrighted works and selling access to unauthorized derivative works.
Post reply on HN