Live data from Hacker News

Tell HN: OpenAI keeps re-enabling the 'allow training' setting

news.ycombinator.com

151–160 of 172 posts

Re: Tell HN: OpenAI keeps re-enabling the 'allow training' setting

#152

It really is going to get to the point where mathematicians are going to start inserting obvious "tells" in their proofs - like map makers used to do with "trap streets" and similar non-existent features, to catch copies.-

Do you think the training process will retain these artifacts? I doubt it. If they were simply stealing the content - sure it would make sense - but I suspect they’re feeding it into training data and RL might distill these out.

I am thinking along these lines. You do raise a great point.-

https://www.anthropic.com/research/small-samples-poison?from...

Re: Tell HN: OpenAI keeps re-enabling the 'allow training' setting

#153
post #49

Weird, I'm the opposite. I don't recall ever setting mine and I just checked and it was set to "disallow training".

Is this Settings > Data Controls > "Improve the model for everyone" or is there a "disallow training" somewhere else? presumably its default behavior will vary depending on your user subscription or if your account belongs to an organization (Business, Enterprise, Edu?).

Yes that's what I was referring to. I'm a non paid user but I did briefly pay for a consumer subscription before. Not sure what region I'm assigned to as I'm in China and always access through a VPN.

Re: Tell HN: OpenAI keeps re-enabling the 'allow training' setting

#154
post #119

Earlier quoted context omitted.

There is nothing even close to a proof. A lot of accusations, a lot of people ready with pitchforks and torches (sadly, also here on HN), but not a lot of facts. Did the researches opt out from data sharing on subsidised subs? Did anyone prove that their methods enabled OpenAI models to produce the solution? For a discussion about science, there is almost no scientifical method applied to proving anyone stole anythin…

On one side, yes we don't have hard evidence that intentional plagiarism is exactly what happened. On the other side, the lack of evidence is pretty damning. Only OpenAI can try to prove that they came by these results legitimately, and the case they're making is quite weak. They could make public metadata about what their model was trained on and whether it did train on the conversations in question; they have not.…

Actually no, the lack of evidence can be readily fixed by the researchers simply disclosing the pertinent parts of their notes and/or chats. The discovery has been scooped, so I don't see any value in keeping them private anymore. Then everybody can see how related the models' and the researchers' works are.

Re: Tell HN: OpenAI keeps re-enabling the 'allow training' setting

#155
post #59
post #9

Based on their behavior over the past few years, why would you assume that checkbox even does anything at all?

If it didn't do anything, they wouldn't keep re-enabling it.

Did you encounter it too? (Trying to get a rough estimate of how many people are reporting it vs how many people aren't.)

Re: Tell HN: OpenAI keeps re-enabling the 'allow training' setting

#156
post #155
post #59

Earlier quoted context omitted.

If it didn't do anything, they wouldn't keep re-enabling it.

Did you encounter it too? (Trying to get a rough estimate of how many people are reporting it vs how many people aren't.)

No, mine stays disabled and there's no way to enable it. It's just a label that says "disabled". Not sure why, maybe some sort of company wide profile?

Re: Tell HN: OpenAI keeps re-enabling the 'allow training' setting

#157

Earlier quoted context omitted.

Yes, blame the users.

Up until ~1800, the way societies handled this kind of depravity was to hit the bad actors with sticks or rocks until their skull opened up so the evil spirits could leave their bodies. It's unfortunate that most societies have outlawed this practice. I'm certain if that were in place today, we wouldn't have this problem. I've yet to hear any alternative solution that's as effective.

They used to burn witches too. I don't know if it was really effective, though, they kept finding witches all over.

Re: Tell HN: OpenAI keeps re-enabling the 'allow training' setting

#158
post #155

Earlier quoted context omitted.

Did you encounter it too? (Trying to get a rough estimate of how many people are reporting it vs how many people aren't.)

No, mine stays disabled and there's no way to enable it. It's just a label that says "disabled". Not sure why, maybe some sort of company wide profile?

Interesting, yeah... also just found out about "Advanced Account Security" which sounds like something a company would enable: https://news.ycombinator.com/item?id=49643999

Re: Tell HN: OpenAI keeps re-enabling the 'allow training' setting

#159
post #158

Earlier quoted context omitted.

No, mine stays disabled and there's no way to enable it. It's just a label that says "disabled". Not sure why, maybe some sort of company wide profile?

Interesting, yeah... also just found out about "Advanced Account Security" which sounds like something a company would enable: https://news.ycombinator.com/item?id=49643999

Oh, I enabled that myself.

Re: Tell HN: OpenAI keeps re-enabling the 'allow training' setting

#160
post #35
post #2

If you document it properly then this basically destroys any legal claim they can make about that checkbox.

Out of curiosity, how one is supposed to "document it properly"?

You can obtain a cryptographic proof by recording the tls exchange, including the keys

You need to use a tls intercepting proxy for that.

I couldn't find any ready-made tool unfortunately, there's tlsnotary.org but it seems far from simple.

Post reply on HN