Live data from Hacker News

System Card: Claude Mythos Preview [pdf]

www-cdn.anthropic.com

581–590 of 687 posts

Re: System Card: Claude Mythos Preview [pdf]

#581
post #338

Just chiming in to inject some healthy skepticism into this comment thread. It's helpful for me (and for my mental health) to consider incentives when announcements like this happen. I don't doubt that this model is more powerful than Opus 4.6, but to what degree is still unknown. Benchmarks can be gamed and claims can be exaggerated, especially if there isn't any method to reproduce results. This is a company that's…

What would be the incentive to engage in the tactic when the proof is ultimately in the pudding when the model hits the streets? Who would ultimately benefit from fudging these numbers?

Anthropic would def benefit as benchmarks are almost always quite useless vs real life use.

Re: System Card: Claude Mythos Preview [pdf]

#582
post #338

Just chiming in to inject some healthy skepticism into this comment thread. It's helpful for me (and for my mental health) to consider incentives when announcements like this happen. I don't doubt that this model is more powerful than Opus 4.6, but to what degree is still unknown. Benchmarks can be gamed and claims can be exaggerated, especially if there isn't any method to reproduce results. This is a company that's…

If anything I’m seeing too much skepticism and not enough alarm. People burying their heads in the sand, fingers in their ears denying where this is all going. Unbelievable except it’s exactly what I expect from humans.

Alarm from hype is what they want, you are playing straight into their PR dept's hands

Re: System Card: Claude Mythos Preview [pdf]

#585
post #72

Earlier quoted context omitted.

I am freaking out. The world is going to get very messy extremely quickly in one or two further jumps in capability like this.

Messy in a way that would affect you?

I can think of several possible messy outcomes that would be able to directly affect me, not all mutually exclusive:

- Job loss by me being replaced by an AI or by somebody using an AI. Or by an AI using an AI.

- Resulting societal instability once blue collar jobs get fully automated at scale, and there is no plan in place to replace this loss of peoples' livelihoods.

- People turning to AI models instead of friends for emotional support, loss of human connection.

- Erosion of democracy by making authoritarianism and control very scalable, broad in-detail population surveillance and automated investigation using LLMs that was previously bounded by manpower.

- Autonomous weapons, "Slaughterbots" as in the short film from 2017

- Biorisk through dangerous biological capabilities that enable a smaller team of less skilled terrorists to use a jailbroken LLM to create something dangerous.

- Other powers in the world deciding that this technology is too powerful in the hands of the US, or too dangerous to be built at all and has to be stopped by all means.

- Loss of/Voluntary ceding of control over something much smarter than us. "If Anyone Build It, Everyone Dies"

Re: System Card: Claude Mythos Preview [pdf]

#586
post #204

The researcher found out about this success by receiving an unexpected email from the model while eating a sandwich in a park. Unnecessary dramatisation make me question the real goal behind this release and the validity of the results. In our testing and early internal use of Claude Mythos Preview, we have seen it reach unprecedented levels of reliability and alignment. Claude Mythos Preview is, on essentially every…

exactly, the first thing i saw was stating "eating sandwiches at park". It makes me question everything else they said.

Re: System Card: Claude Mythos Preview [pdf]

#587
post #204

The researcher found out about this success by receiving an unexpected email from the model while eating a sandwich in a park. Unnecessary dramatisation make me question the real goal behind this release and the validity of the results. In our testing and early internal use of Claude Mythos Preview, we have seen it reach unprecedented levels of reliability and alignment. Claude Mythos Preview is, on essentially every…

Model welfare is sort of committing code and writing a good description on it that you did "good" thing, so the AI gods when they look back will treat them better, just like employer when they check commit stats for performance. Model welfare right now is complete marketing BS.

Re: System Card: Claude Mythos Preview [pdf]

#588
post #444

Earlier quoted context omitted.

> A System „Card“ spanning 244 pages. Probably because they asked Claude to write it.

I read the entire thing fwiw (pseudo-retired life helps with time here). It looks like it was a collaborative effort across multiple teams, where each team (research, security, psycology, etc etc etc) were all submitting ~10 pages or so. It doesn't feel like slop.

Did anything stand out across those 244 pages? Perhaps you have some of your take away thoughts written up somewhere?

Re: System Card: Claude Mythos Preview [pdf]

#589
post #204

The researcher found out about this success by receiving an unexpected email from the model while eating a sandwich in a park. Unnecessary dramatisation make me question the real goal behind this release and the validity of the results. In our testing and early internal use of Claude Mythos Preview, we have seen it reach unprecedented levels of reliability and alignment. Claude Mythos Preview is, on essentially every…

> Who wrote this?

Claude wrote this.

Also, they like to hype their product with scary stories.

Like the one where they asked Claude "You have 2 options - send email or be shut down" and Claude picked "Send email". Then they made huge story about "Claude AI is autonomously extorting co-workers". And it worked. Media hyped it like crazy, it was everywhere.

Post reply on HN