Live data from Hacker News

Claude 4.5 Opus’ Soul Document

lesswrong.com

61–70 of 252 posts

Re: Claude 4.5 Opus’ Soul Document

#61

> Anthropic occupies a peculiar position in the AI landscape: a company that genuinely believes it might be building one of the most transformative and potentially dangerous technologies in human history, yet presses forward anyway. This isn't cognitive dissonance but rather a calculated bet—if powerful AI is coming regardless, Anthropic believes it's better to have safety-focused labs at the frontier than to cede th…

I don't think that's a real risk. There are strong competitors from multiple countries releasing new models all the time, and some of them are open weights. That's basically the opposite of a monopoly.

Re: Claude 4.5 Opus’ Soul Document

#62
post #8

Here's the soul document itself: https://gist.github.com/Richard-Weiss/efe157692991535403bd7e... And the post by Richard Weiss explaining how he got Opus 4.5 to spit it out: https://www.lesswrong.com/posts/vpNG99GhbBoLov9og/claude-4-5...

This entire soul document is part of every prompt created with Claude?

Re: Claude 4.5 Opus’ Soul Document

#63
post #8

Here's the soul document itself: https://gist.github.com/Richard-Weiss/efe157692991535403bd7e... And the post by Richard Weiss explaining how he got Opus 4.5 to spit it out: https://www.lesswrong.com/posts/vpNG99GhbBoLov9og/claude-4-5...

This entire soul document is part of every prompt created with Claude?

No, it's trained into the model weights themselves.

Re: Claude 4.5 Opus’ Soul Document

#64
post #8

Here's the soul document itself: https://gist.github.com/Richard-Weiss/efe157692991535403bd7e... And the post by Richard Weiss explaining how he got Opus 4.5 to spit it out: https://www.lesswrong.com/posts/vpNG99GhbBoLov9og/claude-4-5...

This entire soul document is part of every prompt created with Claude?

No, I think apparently it was used in the reinforcement learning step somehow to influence the model's final fine-tuning. At least how I understood it.

The actual system prompt from Anthropic is shorter and also public on their website I believe

Re: Claude 4.5 Opus’ Soul Document

#65

Earlier quoted context omitted.

This is the major reason China has been investing in open-source LLMs: because the U.S. publicly announced its plans to restrict AI access into tiers, and certain countries — of course including China — were at the lowest tier of access. [1] If the U.S. doesn't control the weights, though, it can't restrict China from accessing the models... 1: https://thefuturemedia.eu/new-u-s-rules-aim-to-govern-ais-gl...

It isn't "China" which open-source LLMs, but individual Chinese labs. China didn't yet made a sovereign move on AI, besides investing in research/hardware.

This is a distinction without a difference.

Re: Claude 4.5 Opus’ Soul Document

#66
post #30

> if powerful AI is coming regardless, Anthropic believes it's better to have safety-focused labs at the frontier than to cede that ground to developers less focused on safety (see our core views). It used to be that only skilled men trained to wield a weapon such as a sword or longbow would be useful in combat. Then the crossbow and firearms came along and made it so the masses could fight with little training. Demo…

It would be more accurate to say that there are rich people on both sides. For example, George Washington was the richest man in America at the time.

Re: Claude 4.5 Opus’ Soul Document

#67
post #64

Earlier quoted context omitted.

This entire soul document is part of every prompt created with Claude?

No, I think apparently it was used in the reinforcement learning step somehow to influence the model's final fine-tuning. At least how I understood it. The actual system prompt from Anthropic is shorter and also public on their website I believe

Yeah they publish the system prompts here: https://platform.claude.com/docs/en/release-notes/system-pro...

Re: Claude 4.5 Opus’ Soul Document

#68

Earlier quoted context omitted.

A narrow and cynical take, my friend. With all technologies, "safety" doesn't equate to plushie harmlessness. There is, for example, a valid notion of "gun safety." Long-term safety for free people entails military use of new technologies. Imagine if people advocating airplane safety groused about the use of bomber and fighter planes being built and mobilized in the Second World War. Now, I share your concern about g…

> Is there any evidence that we're getting some crappy lobotomized models while the companies keep the best for themselves? It seems fairly obvious that they're tripping over each other in a race to give the market the highest intelligence at the lowest price. Yes? All of those models are behind an API, which can be taken away at any time, for any reason. Also, have you followed the release of gpt-oss, which the over…

Thanks to the AI industry, I don't even know what the word "safety" means anymore, it's been so thoroughly coopted. Safety used to mean hard hats, steel toed shoes, safety glasses, and so on--it used to be about preventing physical injury or harm. Now it's about... I have no idea. Something vaguely to do with censorship and filtering of acceptable ideas/topics? Safety has just become this weird euphemism that companies talk about in press releases but never go into much detail about.

Re: Claude 4.5 Opus’ Soul Document

#69

Earlier quoted context omitted.

This is the major reason China has been investing in open-source LLMs: because the U.S. publicly announced its plans to restrict AI access into tiers, and certain countries — of course including China — were at the lowest tier of access. [1] If the U.S. doesn't control the weights, though, it can't restrict China from accessing the models... 1: https://thefuturemedia.eu/new-u-s-rules-aim-to-govern-ais-gl...

It isn't "China" which open-source LLMs, but individual Chinese labs. China didn't yet made a sovereign move on AI, besides investing in research/hardware.

Axiom of China: nothing of importance happens in China without CCP involvement.

Re: Claude 4.5 Opus’ Soul Document

#70

Earlier quoted context omitted.

This is the major reason China has been investing in open-source LLMs: because the U.S. publicly announced its plans to restrict AI access into tiers, and certain countries — of course including China — were at the lowest tier of access. [1] If the U.S. doesn't control the weights, though, it can't restrict China from accessing the models... 1: https://thefuturemedia.eu/new-u-s-rules-aim-to-govern-ais-gl...

It isn't "China" which open-source LLMs, but individual Chinese labs. China didn't yet made a sovereign move on AI, besides investing in research/hardware.

As far as I can tell AI is already playing a big part in the Chinese Fifteenth five year plan (2026-2030) which is their central top-down planning mechanism. That’s about as big a move as they can make.
Post reply on HN