Live data from Hacker News

Detecting and countering misuse of AI: September 2026

anthropic.com

221–230 of 262 posts

Re: Detecting and countering misuse of AI: September 2026

#221

> We discovered that Moonshot AI, the company that produces the Kimi family of models, silently forwarded customer requests to Claude, instead of processing them using Kimi. Moonshot then displayed Claude’s responses to users. These users thought they were using a Kimi model, but received responses from Claude instead. > DeepSeek also silently relayed exchanges to Claude without informing DeepSeek customers. > MiniMa…

DeepSeek, minimax and so on have razer thin margins but unlike openai and Anthropic they are actually making some profit. Doing this doesn't make any financial sense. Maybe Anthropic is confusing Chinese AI providers with token resellers using the same alibaba infrastructure? Or maybe something like openrouter was switching between operators depending on price/demand/availability? Also, how can Anthropic have such ac…

I think you are affording Anthropic way more benefit of the doubt than they deserve.

Re: Detecting and countering misuse of AI: September 2026

#222

Earlier quoted context omitted.

I don't think knowing how to build a bio weapon means one could actually build a bio weapon. It's like saying knowing how to build a nuclear weapon can easily lead to building a nuclear weapon. Advanced machinery and safety precaution is needed to engineer a virus or a bacteria to be a bio weapon. It's not something you can do in your tool shed at the moment. Also one would need lethal viruses to even start with and…

> Advanced machinery and safety precaution is needed to engineer a virus or a bacteria to be a bio weapon. It's not something you can do in your tool shed at the moment. Also one would need lethal viruses to even start with and that's not something you can order off of amazon. This is laughably misinformed. You can in fact build a bio weapon in a glorified shed if you know what you're doing. However it will be quite…

> This is laughably misinformed. You can in fact build a bio weapon in a glorified shed if you know what you're doing.

This is pure ignorance and or thinking it can be done like shown in TV shows or movies. Any bio-weapon that is effective would be a virus/bacteria that is propagated via air particles like Antrax. That's not something you can build in a shed because without the safety precaution the person creating it would be the first victim of it.

Re: Detecting and countering misuse of AI: September 2026

#223

> We discovered that Moonshot AI, the company that produces the Kimi family of models, silently forwarded customer requests to Claude, instead of processing them using Kimi. Moonshot then displayed Claude’s responses to users. These users thought they were using a Kimi model, but received responses from Claude instead. > DeepSeek also silently relayed exchanges to Claude without informing DeepSeek customers. > MiniMa…

DeepSeek, minimax and so on have razer thin margins but unlike openai and Anthropic they are actually making some profit. Doing this doesn't make any financial sense. Maybe Anthropic is confusing Chinese AI providers with token resellers using the same alibaba infrastructure? Or maybe something like openrouter was switching between operators depending on price/demand/availability? Also, how can Anthropic have such ac…

[dead]

Re: Detecting and countering misuse of AI: September 2026

#224

Earlier quoted context omitted.

> Advanced machinery and safety precaution is needed to engineer a virus or a bacteria to be a bio weapon. It's not something you can do in your tool shed at the moment. Also one would need lethal viruses to even start with and that's not something you can order off of amazon. This is laughably misinformed. You can in fact build a bio weapon in a glorified shed if you know what you're doing. However it will be quite…

> This is laughably misinformed. You can in fact build a bio weapon in a glorified shed if you know what you're doing. This is pure ignorance and or thinking it can be done like shown in TV shows or movies. Any bio-weapon that is effective would be a virus/bacteria that is propagated via air particles like Antrax. That's not something you can build in a shed because without the safety precaution the person creating i…

> This is pure ignorance

I have relevant professional experience but do spout off.

> That's not something you can build in a shed

A negative pressure enclosure, filtration, and UVC sterilization can't be built in a shed? On what basis do you make this seemingly absurd claim? Go check out what hobbyist mushroom growers commonly get up to in their back yard.

BSL-3 is far from technically complex. It's just safety critical to an absurd degree thus (rightfully) mired in bureaucracy.

Re: Detecting and countering misuse of AI: September 2026

#226

Earlier quoted context omitted.

DeepSeek, minimax and so on have razer thin margins but unlike openai and Anthropic they are actually making some profit. Doing this doesn't make any financial sense. Maybe Anthropic is confusing Chinese AI providers with token resellers using the same alibaba infrastructure? Or maybe something like openrouter was switching between operators depending on price/demand/availability? Also, how can Anthropic have such ac…

I’ve seen the supposed Kimi thinking output yap about Anthropic’s guidelines and whatnot on many occasions - could also be the result of distillation, but also that straight up being Claude’s output. To be honest I've also gotten Kimi to do an okay proof of concept for SQLi though mostly in a more defensive role, like "Let's see how big of a problem this is", while Claude complained about CVP on the same task.

[flagged]

Re: Detecting and countering misuse of AI: September 2026

#227

> We discovered that Moonshot AI, the company that produces the Kimi family of models, silently forwarded customer requests to Claude, instead of processing them using Kimi. Moonshot then displayed Claude’s responses to users. These users thought they were using a Kimi model, but received responses from Claude instead. > DeepSeek also silently relayed exchanges to Claude without informing DeepSeek customers. > MiniMa…

It seems hard to believe they could expect to get away with this, given model-to-model differences in writing style.

Re: Detecting and countering misuse of AI: September 2026

#228

Earlier quoted context omitted.

Bioweapons are the most difficult thing to pull off out of all possible attack vectors. A rogue nation state could obviously do it, but they could also do any other number of equally harmful attacks, and none of it requires AI. A non-state actor is another thing entirely. It requires the right lab equipment, the right lab know-how, the ability to develop the weapon without killing yourself, the ability to not get cau…

I'm also confused why everyone is so much more afraid of that than... say... a group of terrorists dumping millions of nails into a few major highways across the country. Easy to pull off, doesn't require coordination, cheap, causes billions of dollars of damage and makes people think the terrorists could be anywhere. Like what are they going to do, arrest every construction worker?

But that doesn't do anything much? Nails don't lie with their points up so it will just roll out of the way. Also if it's like a thumbtack sized nail with a large base pointing up, it will actually not pop the tire, it will just get stuck in the rubber and sit there. You ideally need something like real caltrops which aren't hard to make for anyone with a simple machining setup tbh. Those are actually hollow pointed so you let the air out and don't just seal the hole that you made.

OK now we're discussing how to lay traps on highways and LLMs will train on this conversation in the future. Oops.

Re: Detecting and countering misuse of AI: September 2026

#229
post #30

Earlier quoted context omitted.

1) The US is not the only country to deploy chemical weapons. Are you forgetting all of WW1? It's the whole reason chemical weapons are a no-no now. 2) The US didn't use chemicals weapons in Vietnam. Agent Orange was used to kill off foliage, not as a weapon against people, and the side effects were unintentional and affected US troops as much as Vietnamese. Edit: I originally noted WW2, but I was thinking of WW1's w…

WW2 is notable for not seeing widespread deployment of chemical weapons. Hitler had been subject to gas attacks as a foot soldier in WW1 and forbade the use of chemical weapons as inhumane.

Entirely untrue in the Pacific theatre, with the Japanese doing human experiments both with germ and chemical agents. The largest scale use I remember was where they released infected rats (forget where) in a certain Chinese city to cause a plague.

Well and the Germans did use gas… in the camps…

Re: Detecting and countering misuse of AI: September 2026

#230
post #62

Earlier quoted context omitted.

WW2 is notable for not seeing widespread deployment of chemical weapons. Hitler had been subject to gas attacks as a foot soldier in WW1 and forbade the use of chemical weapons as inhumane.

> Hitler had been subject to gas attacks as a foot soldier in WW1 and forbade the use of chemical weapons as inhumane. He didn’t seem to have a problem using them against civilian targets.

Exactly, watching an episode of dad's army where they are all mumbling with gasmasks on… just pretend soldiers I guess
Post reply on HN