Slopinator: Attack AI training with poisoned GitHub repositories
1–10 of 11 posts
Re: Slopinator: Attack AI training with poisoned GitHub repositories
#2Poison Fountain on Reddit: https://www.reddit.com/r/PoisonFountain/
Miasma Poison Tar Pit: https://news.ycombinator.com/item?id=47561819
Re: Slopinator: Attack AI training with poisoned GitHub repositories
#3Re: Slopinator: Attack AI training with poisoned GitHub repositories
#4Finally an AI project with a sense of purpose!
Re: Slopinator: Attack AI training with poisoned GitHub repositories
#5Re: Slopinator: Attack AI training with poisoned GitHub repositories
#6Re: Slopinator: Attack AI training with poisoned GitHub repositories
#7I think these sort of efforts are mostly self-soothing at this point. It is almost certainly the case that the labs are at a minimum running inference over the information they're pulling and ensuring that it's useful/suitable for pre-training. The models are at least good enough to know whether they're looking at utter nonsense.
Re: Slopinator: Attack AI training with poisoned GitHub repositories
#8Re: Slopinator: Attack AI training with poisoned GitHub repositories
#9I think these sort of efforts are mostly self-soothing at this point. It is almost certainly the case that the labs are at a minimum running inference over the information they're pulling and ensuring that it's useful/suitable for pre-training. The models are at least good enough to know whether they're looking at utter nonsense.
Re: Slopinator: Attack AI training with poisoned GitHub repositories
#10I think these sort of efforts are mostly self-soothing at this point. It is almost certainly the case that the labs are at a minimum running inference over the information they're pulling and ensuring that it's useful/suitable for pre-training. The models are at least good enough to know whether they're looking at utter nonsense.
Actually it was shown a couple of times already, some of it also by Anthropic's own research, that the LLMs are extremely easy to poison with small datasets.