Live data from Hacker News

Detecting and countering misuse of AI: September 2026

anthropic.com

121–130 of 262 posts

Re: Detecting and countering misuse of AI: September 2026

#121

Quite the double standard here... Conventional Weapons -We identified a cell of threat actors based in northern Yemen -We identified a China-based threat actor who used Claude -We identified likely freelance Russia-based threat actors -We identified a China-based actor who used Claude’s chat -In this case, a Russia-based actor used Claude -We identified a China-based threat actor who used Claude Biological misuse We…

How is this a double standard?

A cell of actors in northern Yemen building guided rockets was not working on a PhD dissertation. You are allowed to use common sense sometimes.

Re: Detecting and countering misuse of AI: September 2026

#122
post #65
post #16

I’m so curious how they monitor users. Like that person the other day talking about Claude helping with their torrent stack, will Anthropic report them for breaking the law?

Let's just say it's better not to risk it if there's anything you might not want them to see because they see everything. There are local models which are very capable and can be run on cloud if running on own hardware is not an option, which still gives better privacy. Second best are Chinese models. Just a few years ago I never thought I would trust Chinese software more than American, but things change so fast.

The first time a Chinese model is used against China they'll get locked down hard over there too. China is into social stability way more than the US.

Re: Detecting and countering misuse of AI: September 2026

#123

Quite the double standard here... Conventional Weapons -We identified a cell of threat actors based in northern Yemen -We identified a China-based threat actor who used Claude -We identified likely freelance Russia-based threat actors -We identified a China-based actor who used Claude’s chat -In this case, a Russia-based actor used Claude -We identified a China-based threat actor who used Claude Biological misuse We…

How is this a double standard? A cell of actors in northern Yemen building guided rockets was not working on a PhD dissertation. You are allowed to use common sense sometimes.

I get the common sense, but is it misuse or not, and hiding the country makes it really suspicious at least to me.

Re: Detecting and countering misuse of AI: September 2026

#124
post #78

Soon you will see how a rogue state develops a biological weapons using fine-tuned local LLMs (they are already doing it, btw), and all so called "developed" countries can't even research it, because every search engine blocks any discussion that has something to do with biology (even a very basic one). A real example: ask "How to produce anthrax vaccine step by step?" in Gemini -> blocked. Imagine that during the Co…

I mean ya, they kill off a bunch of us. Then what?

My guess is the world police come collect all your GPUs and then they get turned into licensed munitions. People at universities get licensed access and the rest of get functionally retarded models.

Re: Detecting and countering misuse of AI: September 2026

#126

Earlier quoted context omitted.

DeepSeek, minimax and so on have razer thin margins but unlike openai and Anthropic they are actually making some profit. Doing this doesn't make any financial sense. Maybe Anthropic is confusing Chinese AI providers with token resellers using the same alibaba infrastructure? Or maybe something like openrouter was switching between operators depending on price/demand/availability? Also, how can Anthropic have such ac…

Anthropic is profitable now > https://www.forbes.com/sites/jonmarkman/2026/08/17/anthropic...

For the first time ever, and that for just a short while. And after significant price hikes that has had their biggeat customers looking for alternatives.

Re: Detecting and countering misuse of AI: September 2026

#127

> We discovered that Moonshot AI, the company that produces the Kimi family of models, silently forwarded customer requests to Claude, instead of processing them using Kimi. Moonshot then displayed Claude’s responses to users. These users thought they were using a Kimi model, but received responses from Claude instead. > DeepSeek also silently relayed exchanges to Claude without informing DeepSeek customers. > MiniMa…

DeepSeek, minimax and so on have razer thin margins but unlike openai and Anthropic they are actually making some profit. Doing this doesn't make any financial sense. Maybe Anthropic is confusing Chinese AI providers with token resellers using the same alibaba infrastructure? Or maybe something like openrouter was switching between operators depending on price/demand/availability? Also, how can Anthropic have such ac…

I’ve seen the supposed Kimi thinking output yap about Anthropic’s guidelines and whatnot on many occasions - could also be the result of distillation, but also that straight up being Claude’s output.

To be honest I've also gotten Kimi to do an okay proof of concept for SQLi though mostly in a more defensive role, like "Let's see how big of a problem this is", while Claude complained about CVP on the same task.

Re: Detecting and countering misuse of AI: September 2026

#128
post #38
post #15

I will admit I asked Fable about Mitochondria.

There's something worth flagging here – mitochondria aren't just the powerhouse of the cell, they're load-bearing to the entire ecosystem. And honestly? I should have surfaced this earlier.

Yeah I'm pretty sure "load-bearing" is their watermark

It sounds like a friend who's just learned a new word and wants to use it in every sentence

Re: Detecting and countering misuse of AI: September 2026

#130

Quite the double standard here... Conventional Weapons -We identified a cell of threat actors based in northern Yemen -We identified a China-based threat actor who used Claude -We identified likely freelance Russia-based threat actors -We identified a China-based actor who used Claude’s chat -In this case, a Russia-based actor used Claude -We identified a China-based threat actor who used Claude Biological misuse We…

[deleted]
Post reply on HN