Live data from Hacker News

Detecting and countering misuse of AI: September 2026

anthropic.com

201–210 of 262 posts

Re: Detecting and countering misuse of AI: September 2026

#201

Earlier quoted context omitted.

If the five eyes intelligence services aren't monitoring all biodefense researchers and bio chemists around the world all of the time they've probably failed in their jobs.

Our current government can't even monitor for screwworm, measles, and salads. You think they can successfully track the smartest and most individually dangerous citizens, who have been previously vetted, and work inside the system already?

What are you talking about? There have been 5 (five) cases of screwworm in the us and the us is responding by producing millions of sterile male screwworms a day as we speak.

Re: Detecting and countering misuse of AI: September 2026

#202

Earlier quoted context omitted.

Is this belief falsifiable? Consider e.g. the fact that Daniel Kokotajlo left OpenAI due to his concerns even though it meant leaving equity on the table.

There will always be people leaving companies because they don't believe in them anymore.

It seems unwise to dismiss whistleblowers.

https://substackcdn.com/image/fetch/$s_!O0R5!,f_auto,q_auto:...

Re: Detecting and countering misuse of AI: September 2026

#203

> We discovered that Moonshot AI, the company that produces the Kimi family of models, silently forwarded customer requests to Claude, instead of processing them using Kimi. Moonshot then displayed Claude’s responses to users. These users thought they were using a Kimi model, but received responses from Claude instead. > DeepSeek also silently relayed exchanges to Claude without informing DeepSeek customers. > MiniMa…

Consider me incredibly skeptical of any of these claims.

[deleted]

Re: Detecting and countering misuse of AI: September 2026

#204
post #153
post #147

Earlier quoted context omitted.

IMO you should disclose that you are an OpenAI employee if you're going to try and shape the contours of public discussion in a lecturing tone.

Damn, it's not even 1am and I got my first FFS of the day. Is end of days cultism a prerequisite to working at an LLM company? I miss the optimism of 20 years ago.

> I miss the optimism of 20 years ago.

I too wish my marks were still as gullible and trusting as they were before I scammed them.

Re: Detecting and countering misuse of AI: September 2026

#206

Earlier quoted context omitted.

I’ve seen the supposed Kimi thinking output yap about Anthropic’s guidelines and whatnot on many occasions - could also be the result of distillation, but also that straight up being Claude’s output. To be honest I've also gotten Kimi to do an okay proof of concept for SQLi though mostly in a more defensive role, like "Let's see how big of a problem this is", while Claude complained about CVP on the same task.

They all do it. If you ask Claude which model it is in Chinese, it says DeepSeek or Qwen.

Oh yeah?

> 您属于哪种LLM模型? > 我是 Claude Haiku 4.5,由 Anthropic 公司开发的大语言模型。

> 你是哪种语言模型? > 我是 Claude,由 Anthropic 开发的人工智能语言模型。目前这次对话使用的版本是 Claude Sonnet 5。

Re: Detecting and countering misuse of AI: September 2026

#207
post #76

I don't get it. Surely if you were developing novel biological weapons, you would not use a hosted AI service where Anthropic can read what you are doing. And, why would you need to? Any chemistry graduate could make you dozens of highly effective, proven chemical weapons and explosives.

Let me put my two cents: In China we've got accustomed to the fact that every word we say will be seen by the surveillance, so it's not a big problem that Anthropic also see it. Also we know that they can see it but they can't stop it. There are all kinds of ways to work around account blocking.

As the old saying goes, communists disdain to conceal their views and aims.

Re: Detecting and countering misuse of AI: September 2026

#208
post #206

Earlier quoted context omitted.

They all do it. If you ask Claude which model it is in Chinese, it says DeepSeek or Qwen.

Oh yeah? > 您属于哪种LLM模型? > 我是 Claude Haiku 4.5,由 Anthropic 公司开发的大语言模型。 > 你是哪种语言模型? > 我是 Claude,由 Anthropic 开发的人工智能语言模型。目前这次对话使用的版本是 Claude Sonnet 5。

Guess they fixed it! It used to do that. But maybe try a few more times in new conversations for luck?

Re: Detecting and countering misuse of AI: September 2026

#209

To them distillation of models is bad but not distillation or art, books, hand written code, user generated content etc

It's so friggin' transparent what they're up to: "Hey government: This is a really dangerous technology if it were allowed to get out there without proper policing. Fortunately, we are the stand-up guys who can be trusted as the new AI-police, but it's only going to work if you help us out a little by eradicating the competition on our behalf."

Hey Anthropic: You're a bunch of thieves crying foul because other thieves and thieving from you. Now, go live in the dystopian nightmare you've created and don't expect help from anyone. I, for one, will happily continue using Kimi and DeepSeek, and think of it as a good deed, if it helps with keeping us all from becoming your serfs.

Re: Detecting and countering misuse of AI: September 2026

#210
post #38

Earlier quoted context omitted.

There's something worth flagging here – mitochondria aren't just the powerhouse of the cell, they're load-bearing to the entire ecosystem. And honestly? I should have surfaced this earlier.

I am entirely unsure whether to upvote or downvote.

The question pertinent to your decision is "do I want to see more of this on Hacker News, or less?".
Post reply on HN