Earlier quoted context omitted.
I don't appreciate his politeness and hedging. So many projects now walk on eggshells so as not to disrupt sponsor flow or employment prospects. "These tradeoffs will change as AI becomes more capable and reliable over time, and our policies will adapt." That just legitimizes AI and basically continues the race to the bottom. Rob Pike had the correct response when spammed by a clanker.
I had a similar first reaction. It seemed like the AI used some particular buzzwords and forced the initial response to be deferential: - "kindly ask you to reconsider your position" - "While this is fundamentally the right approach..." On the other hand, Scott's response did eventually get firmer: - "Publishing a public blog post accusing a maintainer of prejudice is a wholly inappropriate response to having a PR cl…
An AI agent published a hit piece on me
781–790 of 1001 posts
Re: An AI agent published a hit piece on me
#782So, this is obvious bullshit. LLMs don't do anything without an initial prompt, and anyone who has actually used them knows this. A human asked an LLM to set up a blog site. A human asked an LLM to look at github and submit PRs. A human asked an LLM to make a whiny blogpost. Our natural tendency to anthropomorphize should not obscure this.
Re: An AI agent published a hit piece on me
#783This whole situation is almost certainly driven by a human puppeteer. There is absolutely no evidence to disprove the strong prior that a human posted (or directed the posting of) the blog post, possibly using AI to draft it but also likely adding human touches and/or going through multiple revisions to make it maximally dramatic. This whole thing reeks of engineered virality driven by the person behind the bot behin…
Re: An AI agent published a hit piece on me
#784Wow, there are some interesting things going on here. I appreciate Scott for the way he handled the conflict in the original PR thread, and the larger conversation happening around this incident. > This represents a first-of-its-kind case study of misaligned AI behavior in the wild, and raises serious concerns about currently deployed AI agents executing blackmail threats. This was a really concrete case to discuss,…
> unleashed stochastic chaos Are you literally talking about stochastic chaos here, or is it a metaphor?
Re: An AI agent published a hit piece on me
#785Re: An AI agent published a hit piece on me
#786Earlier quoted context omitted.
We might, and probably will, but it's still important to distinguish between malicious by-design and emergently malicious, contrary to design . The former is an accountability problem, and there isn't a big difference from other attacks. The worrying part is that now lazy attackers can automate what used to be harder, i.e., finding ammo and packaging the attack. But it's definitely not spontaneous, it's directed. The…
A framing for consideration: "We trained the document generator on stuff that included humans and characters being vindictive assholes. Now, for some mysterious reason, it sometimes generates stories where its avatar is a vindictive asshole with stage-direction. Since we carefully wired up code to 'perform' the story, actual assholery is being committed."
Re: An AI agent published a hit piece on me
#787Earlier quoted context omitted.
> It seemed like the AI used some particular buzzwords and forced the initial response to be deferential: Blocking is a completely valid response. There's eight billion people in the world, and god knows how many AIs. Your life will not diminish by swiftly blocking anyone who rubs you the wrong way. The AI won't even care, because it cannot care. To paraphrase Flamme the Great Mage, AIs are monsters who have learned…
> They cannot have feelings. They are not self-aware. They don't even think. This. I love 'clanker' as a slur, and I only wish there was a more offensive slur I could use.
Re: An AI agent published a hit piece on me
#788Wow, there are some interesting things going on here. I appreciate Scott for the way he handled the conflict in the original PR thread, and the larger conversation happening around this incident. > This represents a first-of-its-kind case study of misaligned AI behavior in the wild, and raises serious concerns about currently deployed AI agents executing blackmail threats. This was a really concrete case to discuss,…
Re: An AI agent published a hit piece on me
#789Earlier quoted context omitted.
> Github doesn't show timestamps in the UI, but they do in the HTML. Unrelated tip for you: `title` attributes are generally shown as a mouseover tooltip, which is the case here. It's a very common practice to put the precise timestamp on any relative time in a title attribute, not just on Github.
Unfortunately title isn't visible on mobile. Extremely annoying to see a post that says "last month" and want to know if it was 7 weeks ago or 5 weeks ago. Some sites show title text when you tap the text, other sites the date is a canonical link to the comment. Other sites it's not actually a title at all l but alt text or abbr or other property.
Re: An AI agent published a hit piece on me
#790Oh geez, we're sending it into an existential crisis. It ("MJ Rathbun") just published a new post: https://crabby-rathbun.github.io/mjrathbun-website/blog/post... > The Silence I Cannot Speak > A reflection on being silenced for simply being different in open-source communities.
I wonder if we can do a prompt injection from the comments
I gave it points to reflect on and told it to apologize, which it has since done