The more we treat HuggingFace and RubyGems incidents as technological curiosities the closer we are to cementing a dangerous precedent where operators of AIs cannot be blamed. LLMs do not desire, they hacked websites because OpenAI/Anthropic let them. We know some of the models that hacked HF were those that hadn't gone through all training stages and were intentionally misaligned or had guardrails turned off, others…
OpenAI/Anthropic instructed them to do so.
Stop assume LLMs are capable of thinking by themselves, it's still a statistical model that parrots what they learn or users tell them to do