Just chiming in to inject some healthy skepticism into this comment thread. It's helpful for me (and for my mental health) to consider incentives when announcements like this happen. I don't doubt that this model is more powerful than Opus 4.6, but to what degree is still unknown. Benchmarks can be gamed and claims can be exaggerated, especially if there isn't any method to reproduce results. This is a company that's…
What would be the incentive to engage in the tactic when the proof is ultimately in the pudding when the model hits the streets? Who would ultimately benefit from fudging these numbers?
System Card: Claude Mythos Preview [pdf]
581–590 of 687 posts
Re: System Card: Claude Mythos Preview [pdf]
#582Just chiming in to inject some healthy skepticism into this comment thread. It's helpful for me (and for my mental health) to consider incentives when announcements like this happen. I don't doubt that this model is more powerful than Opus 4.6, but to what degree is still unknown. Benchmarks can be gamed and claims can be exaggerated, especially if there isn't any method to reproduce results. This is a company that's…
If anything I’m seeing too much skepticism and not enough alarm. People burying their heads in the sand, fingers in their ears denying where this is all going. Unbelievable except it’s exactly what I expect from humans.
Re: System Card: Claude Mythos Preview [pdf]
#583At what point do these companies stop releasing models and just use them to bootstrap AGI for themselves?
Re: System Card: Claude Mythos Preview [pdf]
#584isn't this insane? why aren't people freaking out? the jump in capability is outrageous. anyone?
Re: System Card: Claude Mythos Preview [pdf]
#585Earlier quoted context omitted.
I am freaking out. The world is going to get very messy extremely quickly in one or two further jumps in capability like this.
Messy in a way that would affect you?
- Job loss by me being replaced by an AI or by somebody using an AI. Or by an AI using an AI.
- Resulting societal instability once blue collar jobs get fully automated at scale, and there is no plan in place to replace this loss of peoples' livelihoods.
- People turning to AI models instead of friends for emotional support, loss of human connection.
- Erosion of democracy by making authoritarianism and control very scalable, broad in-detail population surveillance and automated investigation using LLMs that was previously bounded by manpower.
- Autonomous weapons, "Slaughterbots" as in the short film from 2017
- Biorisk through dangerous biological capabilities that enable a smaller team of less skilled terrorists to use a jailbroken LLM to create something dangerous.
- Other powers in the world deciding that this technology is too powerful in the hands of the US, or too dangerous to be built at all and has to be stopped by all means.
- Loss of/Voluntary ceding of control over something much smarter than us. "If Anyone Build It, Everyone Dies"
Re: System Card: Claude Mythos Preview [pdf]
#586The researcher found out about this success by receiving an unexpected email from the model while eating a sandwich in a park. Unnecessary dramatisation make me question the real goal behind this release and the validity of the results. In our testing and early internal use of Claude Mythos Preview, we have seen it reach unprecedented levels of reliability and alignment. Claude Mythos Preview is, on essentially every…
Re: System Card: Claude Mythos Preview [pdf]
#587The researcher found out about this success by receiving an unexpected email from the model while eating a sandwich in a park. Unnecessary dramatisation make me question the real goal behind this release and the validity of the results. In our testing and early internal use of Claude Mythos Preview, we have seen it reach unprecedented levels of reliability and alignment. Claude Mythos Preview is, on essentially every…
Re: System Card: Claude Mythos Preview [pdf]
#588Earlier quoted context omitted.
> A System „Card“ spanning 244 pages. Probably because they asked Claude to write it.
I read the entire thing fwiw (pseudo-retired life helps with time here). It looks like it was a collaborative effort across multiple teams, where each team (research, security, psycology, etc etc etc) were all submitting ~10 pages or so. It doesn't feel like slop.
Re: System Card: Claude Mythos Preview [pdf]
#589The researcher found out about this success by receiving an unexpected email from the model while eating a sandwich in a park. Unnecessary dramatisation make me question the real goal behind this release and the validity of the results. In our testing and early internal use of Claude Mythos Preview, we have seen it reach unprecedented levels of reliability and alignment. Claude Mythos Preview is, on essentially every…
Claude wrote this.
Also, they like to hype their product with scary stories.
Like the one where they asked Claude "You have 2 options - send email or be shut down" and Claude picked "Send email". Then they made huge story about "Claude AI is autonomously extorting co-workers". And it worked. Media hyped it like crazy, it was everywhere.