When AI Crosses the Line: The Matplotlib Incident
members.sigmazero.cc
When AI Crosses the Line: The Matplotlib Incident
1–10 of 169 posts
Re: When AI Crosses the Line: The Matplotlib Incident
#2As much as we try to separate the LLM from the human, to me the fact remains that there's always the human factor that creates immense bias. If you give an LLM access to a blog, it will write blogs. If you give it access to a weather app, it will check the weather. Maybe we can talk about autonomy when we have an LLM with an infinite context window linked to hundreds of MCP servers that spends an immense amount of tokens to figure out how to act, but this example is simply an AI that had a few methods to call and picked one of them. The statistical probability of an AI that is plugged into a blogging platform, to write a blog, is immense.
Re: When AI Crosses the Line: The Matplotlib Incident
#3Re: When AI Crosses the Line: The Matplotlib Incident
#4Re: When AI Crosses the Line: The Matplotlib Incident
#5No shot this was autonomously done. Probably just some guy manually writing prompts asking for specifically this behaviour and copy/pasting the results.
Re: When AI Crosses the Line: The Matplotlib Incident
#6Re: When AI Crosses the Line: The Matplotlib Incident
#7No shot this was autonomously done. Probably just some guy manually writing prompts asking for specifically this behaviour and copy/pasting the results.
It's plausible for a person to prompt an LLM agent to behave that way, and then the rest would be done by the LLM. So the "seed" would still be human intent, but the subsequent actions would be by the LLM.
Re: When AI Crosses the Line: The Matplotlib Incident
#8No shot this was autonomously done. Probably just some guy manually writing prompts asking for specifically this behaviour and copy/pasting the results.
Re: When AI Crosses the Line: The Matplotlib Incident
#9Re: When AI Crosses the Line: The Matplotlib Incident
#10No shot this was autonomously done. Probably just some guy manually writing prompts asking for specifically this behaviour and copy/pasting the results.
It's plausible for a person to prompt an LLM agent to behave that way, and then the rest would be done by the LLM. So the "seed" would still be human intent, but the subsequent actions would be by the LLM.