Tell HN: OpenAI keeps re-enabling the 'allow training' setting
151–160 of 175 posts
Re: Tell HN: OpenAI keeps re-enabling the 'allow training' setting
#152It really is going to get to the point where mathematicians are going to start inserting obvious "tells" in their proofs - like map makers used to do with "trap streets" and similar non-existent features, to catch copies.-
Do you think the training process will retain these artifacts? I doubt it. If they were simply stealing the content - sure it would make sense - but I suspect they’re feeding it into training data and RL might distill these out.
https://www.anthropic.com/research/small-samples-poison?from...
Re: Tell HN: OpenAI keeps re-enabling the 'allow training' setting
#153Weird, I'm the opposite. I don't recall ever setting mine and I just checked and it was set to "disallow training".
Is this Settings > Data Controls > "Improve the model for everyone" or is there a "disallow training" somewhere else? presumably its default behavior will vary depending on your user subscription or if your account belongs to an organization (Business, Enterprise, Edu?).
Re: Tell HN: OpenAI keeps re-enabling the 'allow training' setting
#154Earlier quoted context omitted.
There is nothing even close to a proof. A lot of accusations, a lot of people ready with pitchforks and torches (sadly, also here on HN), but not a lot of facts. Did the researches opt out from data sharing on subsidised subs? Did anyone prove that their methods enabled OpenAI models to produce the solution? For a discussion about science, there is almost no scientifical method applied to proving anyone stole anythin…
On one side, yes we don't have hard evidence that intentional plagiarism is exactly what happened. On the other side, the lack of evidence is pretty damning. Only OpenAI can try to prove that they came by these results legitimately, and the case they're making is quite weak. They could make public metadata about what their model was trained on and whether it did train on the conversations in question; they have not.…
Re: Tell HN: OpenAI keeps re-enabling the 'allow training' setting
#155Based on their behavior over the past few years, why would you assume that checkbox even does anything at all?
If it didn't do anything, they wouldn't keep re-enabling it.
Re: Tell HN: OpenAI keeps re-enabling the 'allow training' setting
#156Earlier quoted context omitted.
If it didn't do anything, they wouldn't keep re-enabling it.
Did you encounter it too? (Trying to get a rough estimate of how many people are reporting it vs how many people aren't.)
Re: Tell HN: OpenAI keeps re-enabling the 'allow training' setting
#157Earlier quoted context omitted.
Yes, blame the users.
Up until ~1800, the way societies handled this kind of depravity was to hit the bad actors with sticks or rocks until their skull opened up so the evil spirits could leave their bodies. It's unfortunate that most societies have outlawed this practice. I'm certain if that were in place today, we wouldn't have this problem. I've yet to hear any alternative solution that's as effective.
Re: Tell HN: OpenAI keeps re-enabling the 'allow training' setting
#158Earlier quoted context omitted.
Did you encounter it too? (Trying to get a rough estimate of how many people are reporting it vs how many people aren't.)
No, mine stays disabled and there's no way to enable it. It's just a label that says "disabled". Not sure why, maybe some sort of company wide profile?
Re: Tell HN: OpenAI keeps re-enabling the 'allow training' setting
#159Earlier quoted context omitted.
No, mine stays disabled and there's no way to enable it. It's just a label that says "disabled". Not sure why, maybe some sort of company wide profile?
Interesting, yeah... also just found out about "Advanced Account Security" which sounds like something a company would enable: https://news.ycombinator.com/item?id=49643999
Re: Tell HN: OpenAI keeps re-enabling the 'allow training' setting
#160If you document it properly then this basically destroys any legal claim they can make about that checkbox.
Out of curiosity, how one is supposed to "document it properly"?
You need to use a tls intercepting proxy for that.
I couldn't find any ready-made tool unfortunately, there's tlsnotary.org but it seems far from simple.