AI-Generated Data Can Poison Future AI Models
11–20 of 87 posts
Re: AI-Generated Data Can Poison Future AI Models
#12This reminds me of how fascinated I was as a kid of the artifacts you get from recursively photocopying a piece of paper.
They didn’t graduate to become computer scientists, but did indeed get admitted to the royal school of art the year after.
I found it strangely therapeutic.
Re: AI-Generated Data Can Poison Future AI Models
#13I think it's interesting that human minds generally (though not always!) improve when exposed to the output of other human minds. It seems to be the opposite for current LLMs.
While some persons can strive in these kind of environment (think Kant for example), many would become crazy.
Re: AI-Generated Data Can Poison Future AI Models
#14I think it's interesting that human minds generally (though not always!) improve when exposed to the output of other human minds. It seems to be the opposite for current LLMs.
Re: AI-Generated Data Can Poison Future AI Models
#15I think it's interesting that human minds generally (though not always!) improve when exposed to the output of other human minds. It seems to be the opposite for current LLMs.
Re: AI-Generated Data Can Poison Future AI Models
#16Re: AI-Generated Data Can Poison Future AI Models
#17Unless the internet is no longer useful because there is no way to find anything reliable, there would be enough signal to train and align models.
Re: AI-Generated Data Can Poison Future AI Models
#18I think it's interesting that human minds generally (though not always!) improve when exposed to the output of other human minds. It seems to be the opposite for current LLMs.
Maybe it's less about "Human VS Robot" and more about exposure to "Original thoughts VS mass-produced average thoughts". I don't think a human mind would be improving if they're in a echo-chamber with no new information. I think the reason the human mind is improving is because we're exposed to new, original and/or different thoughts, that we hadn't considered or come across before. Meanwhile, a LLM will just regurgi…
If this were true of humans, we would have never made it this far
Humans are very capable of looking around themselves and thinking "I can do better than this", and then trying to come up with ways how
LLMs are not
Re: AI-Generated Data Can Poison Future AI Models
#19Re: AI-Generated Data Can Poison Future AI Models
#20Some perspectives from someone working in the image space. These tests don't feel practical - That is, they seem intended to collapse the model, not demonstrate "in the wild" performance. The assumption is that all content is black or white - AI or not AI - and that you treat all content as equally worth retraining on. It offers no room for assumptions around data augmentation, human-guided quality discrimination, or…
That this happens doesn't surprise me, but I'd love to see a curve of how each organic vs machine content mixe ratio results in model collapse over N generations.