Live data from Hacker News

Detecting and countering misuse of AI: September 2026

anthropic.com

111–120 of 262 posts

Re: Detecting and countering misuse of AI: September 2026

#111
post #103

[flagged]

The previous discussion shows no obvious signs to me of being flagged, but perhaps it did at some earlier point. What is your evidence that they flagged it, please?

(And by "they" do you mean Anthropic? How would they have the ability to do that?)

Re: Detecting and countering misuse of AI: September 2026

#113

> We discovered that Moonshot AI, the company that produces the Kimi family of models, silently forwarded customer requests to Claude, instead of processing them using Kimi. Moonshot then displayed Claude’s responses to users. These users thought they were using a Kimi model, but received responses from Claude instead. > DeepSeek also silently relayed exchanges to Claude without informing DeepSeek customers. > MiniMa…

DeepSeek, minimax and so on have razer thin margins but unlike openai and Anthropic they are actually making some profit. Doing this doesn't make any financial sense. Maybe Anthropic is confusing Chinese AI providers with token resellers using the same alibaba infrastructure? Or maybe something like openrouter was switching between operators depending on price/demand/availability? Also, how can Anthropic have such ac…

You do this to distill a model.

You can submit your users' questions async too, but if you do it sync, then you can also RLHF on the users' behavior after the output.

Re: Detecting and countering misuse of AI: September 2026

#114

This is just an advertorial. "Our AI is so powerful that Bad Guys could actually use it for real Bad Guy Work!"

"Look daddy government, all these bad guys are doing bad things and we stopped them. You should regulate AI in the US so that companies can't use open source models or buy from China."

Re: Detecting and countering misuse of AI: September 2026

#116

This is just an advertorial. "Our AI is so powerful that Bad Guys could actually use it for real Bad Guy Work!"

I mean, yes any AI of sufficient intelligence and range of data will have this capability.

And yes, it makes the future really messy and all the nice little lines we've drawn on paper that make sense stop making sense.

Re: Detecting and countering misuse of AI: September 2026

#117

Meta: The potential proliferation of biological weapons is serious. Millions could die. It's easy to joke about before it happens, but try to imagine how this thread might look after the successful deployment of a biological weapon by a rogue state or non-state actor. I encourage you to take this topic seriously and contribute posts that add new information or perspectives to the discussion. (I myself think the odds…

I don't think knowing how to build a bio weapon means one could actually build a bio weapon. It's like saying knowing how to build a nuclear weapon can easily lead to building a nuclear weapon. Advanced machinery and safety precaution is needed to engineer a virus or a bacteria to be a bio weapon. It's not something you can do in your tool shed at the moment. Also one would need lethal viruses to even start with and…

>not something you can order off of amazon.

I mean there are services in which you can order things from wet labs so it's not completely hypothetical.

Re: Detecting and countering misuse of AI: September 2026

#118
post #116

This is just an advertorial. "Our AI is so powerful that Bad Guys could actually use it for real Bad Guy Work!"

I mean, yes any AI of sufficient intelligence and range of data will have this capability. And yes, it makes the future really messy and all the nice little lines we've drawn on paper that make sense stop making sense.

[deleted]

Re: Detecting and countering misuse of AI: September 2026

#119

[flagged]

This forum is fundamentally unserious and complete kneejerk cynicism these days. At one point it was prescient. Now these types of things were discussed years ago in other channels and when we reach a point where finally everyone comes around to the right conclusion based on tangible evidence in the news, a majority of people here go "huh i guess so"
Post reply on HN