Live data from Hacker News

Our position on open-weights models

anthropic.com

61–70 of 1001 posts

Re: Our position on open-weights models

#61
> where sufficiently capable models may be able to quickly weaponize pandemic-level viruses with widely available materials

if someone figures out a way to give an LLM full operational control over a virus lab, we've got a whole different set of problems than the ones Dario is describing

Re: Our position on open-weights models

#62
> In fact, the most dangerous model may be one that is trained in secret and handed only to the People’s Liberation Army for use in drones and the Ministry of State Security for surveillance and repression.

Aren't Anthropic models used in project maven: https://en.wikipedia.org/wiki/Project_Maven ?

Re: Our position on open-weights models

#63
Statement is a whole lot of nothing, as expected, but I also don’t know what people are expecting from these guys. That Dario will have a sudden change of heart and publish weights of all his models, flushing $1T down the drain?

Re: Our position on open-weights models

#64

> Anthropic has never advocated for a ban on open-weights models. > All sufficiently capable models, open and closed, should go through mandatory safety testing. Yeah, this is anthropic advocating for a ban on open weight models. Who runs this test? What happens if this test is too costly or the administrator refuses to allow certain people to participate. This is exactly how the US has banned goods in the past, by r…

You forgot “what is the definition of ‘sufficiently capable’”. Presumably it’s anything that competes with Anthropic. If they’re around in a year, presumably they won’t care about Fable level and will only think that whatever competes with Claude 7 or whatever needs to be restricted.

Re: Our position on open-weights models

#66
> In fact, the most dangerous model may be one that is trained in secret and handed only to the People’s Liberation Army for use in drones and the Ministry of State Security for surveillance and repression.

Welcome to bizarro world!

Fist off: "the most dangerous model may be one that is trained in secret" Second: "use in drones [...] for surveillance and repression" I am very appreciative of the freedoms of the west but this type of hypocrisy and lack of self-awareness is bonkers and it should be called out.

Re: Our position on open-weights models

#67
> For example, I worry that biology will have a strong attacker-defender asymmetry, where sufficiently capable models may be able to quickly weaponize pandemic-level viruses with widely available materials,

If he had just left that bit out it wouldn't be so obvious that he's just clutching at straws at this point. In some twisted sense it's almost sad to see.

Re: Our position on open-weights models

#68

“Anthropic has never advocated for a ban on open-weights models. Open-weights models that don’t have dangerous capabilities are a public good…” A bit confused on this part, what model doesn’t have dangerous capabilities?

It seems possible to deliberately not train on some offensive capabilities and still have a very useful model. For example, Opus 5 deliberately did not train on exploiting vulnerabilities, and so performed less well on exploits than Mythos, yet was equally proficient at finding such vulnerabilities, according to the Opus 5 system card in their "OSS-Fuzz" eval [1]. [1]: https://www.securityweek.com/anthropics-opus-5-n…

It's hard to imagine a frontier model being proficient at finding vulnerabilities, but not so proficient at exploiting them.

Surely finding is the hard part, and any LLM should be able to easily exploit a vulnerability it already knows about?

Re: Our position on open-weights models

#69

As expected they will try hard to use government to kill competitors.

> apply such testing to the most capable models regardless of their country of origin or whether they are open or closed ...

> ... (while exempting less capable models, such as those from startups and academia, entirely)

The devil is in the details, but this isn't anti-competitive as stated.

Post reply on HN