Earlier quoted context omitted.
How do you verify the models you download also aren't trying to get you to buy stuff?
I guess you.. ask them a bunch of recommendations? I would imagine this would not be incredibly hard to test as a community
As per dead internet theory, how confident are we that the community which tells us which LLM is safe or unsafe is itself made of real people, and not mostly astroturfing by the owners of LLMs which are biased to promote things for money?
Even DIY testing isn't necessarily enough, deceptive alignment has been shown to be possible as a proof-of-concept for research purposes, and one example of this is date-based: show "good" behaviour before some date, perform some other behaviour after that date.