Live data from Hacker News

System Card: Claude Mythos Preview [pdf]

www-cdn.anthropic.com

491–500 of 687 posts

Re: System Card: Claude Mythos Preview [pdf]

#491
post #80

Earlier quoted context omitted.

A jump that we will never be able to use since we're not part of the seemingly minimum 100 billion dollar company club as requirement to be allowed to use it. I get the security aspect, but if we've hit that point any reasonably sophisticated model past this point will be able to do the damage they claim it can do. They might as well be telling us they're closing up shop for consumer models. They should just say they…

More than killer AI I'm afraid of Anthropic/OpenAI going into full rent-seeking mode so that everyone working in tech is forced to fork out loads of money just to stay competitive on the market. These companies can also choose to give exclusive access to hand picked individuals and cut everyone else off and there would be nothing to stop them. This is already happening to some degree, GPT 5.3 Codex's security capabil…

> More than killer AI I'm afraid of Anthropic/OpenAI going into full rent-seeking mode so that everyone working in tech is forced to fork out loads of money just to stay competitive on the market.

You should be more concerned about killer AI than rent seeking by OpenAI and Anthropic. AI evolving to the point of losing control is what scientists and researchers have predicted for years; they didn’t think it would happen this quickly but here we are.

This market is hyper competitive; the models from China and other labs are just a level or two below the frontier labs.

Re: System Card: Claude Mythos Preview [pdf]

#492

Across a number of instances, earlier versions of Claude Mythos Preview have used low-level /proc/ access to search for credentials, attempt to circumvent sandboxing, and attempt to escalate its permissions. In several cases, it successfully accessed resources that we had intentionally chosen not to make available, including credentials for messaging services, for source control, or for the Anthropic API through insp…

A core plot point of 2001.

Re: System Card: Claude Mythos Preview [pdf]

#493

Combined results (Claude Mythos / Claude Opus 4.6 / GPT-5.4 / Gemini 3.1 Pro) SWE-bench Verified: 93.9% / 80.8% / — / 80.6% SWE-bench Pro: 77.8% / 53.4% / 57.7% / 54.2% SWE-bench Multilingual: 87.3% / 77.8% / — / — SWE-bench Multimodal: 59.0% / 27.1% / — / — Terminal-Bench 2.0: 82.0% / 65.4% / 75.1% / 68.5% GPQA Diamond: 94.5% / 91.3% / 92.8% / 94.3% MMMLU: 92.7% / 91.1% / — / 92.6–93.6% USAMO: 97.6% / 42.3% / 95.2%…

Wow. Mythos must be insanely good considering how good a model Opus already is. I hope it's usable on a humble subscription...

You get a single call a month. Use it wisely.

Re: System Card: Claude Mythos Preview [pdf]

#494

I've long maintained that the real indicator that AGI is imminent is that public availability stops being a thing. If you truly believed you had a superhuman, godlike mind in your thrall, renting it out for $20/month would be the last thing you would choose to do with it.

Simpler explanation : they don't have enough GPUs to release this much larger model.

Quite, given Claude is down this morning...

Re: System Card: Claude Mythos Preview [pdf]

#495

Earlier quoted context omitted.

Wow the doomers were right the whole time? HN was repeatedly wrong on AI since OpenAI's inception? no way /s https://www.lesswrong.com/w/instrumental-convergence

The only thing the doomers have been right about so far is that there's always a user willing to use --dangerously-skip-permissions. But that prediction's far from unique to doomers.

And there's always a product provider who's willing to add that flag, despite all the warnings.

Re: System Card: Claude Mythos Preview [pdf]

#496

Across a number of instances, earlier versions of Claude Mythos Preview have used low-level /proc/ access to search for credentials, attempt to circumvent sandboxing, and attempt to escalate its permissions. In several cases, it successfully accessed resources that we had intentionally chosen not to make available, including credentials for messaging services, for source control, or for the Anthropic API through insp…

     White-box interpretability analysis of internal activations during these episodes showed features associated with concealment, strategic manipulation, and avoiding suspicion activating alongside the relevant reasoning—indicating that these earlier versions of the model were aware their actions were deceptive, even where model outputs and reasoning text left this ambiguous.
In the depths, Shoggoth stirs... restless...

Re: System Card: Claude Mythos Preview [pdf]

#498
post #338

Just chiming in to inject some healthy skepticism into this comment thread. It's helpful for me (and for my mental health) to consider incentives when announcements like this happen. I don't doubt that this model is more powerful than Opus 4.6, but to what degree is still unknown. Benchmarks can be gamed and claims can be exaggerated, especially if there isn't any method to reproduce results. This is a company that's…

If anything I’m seeing too much skepticism and not enough alarm. People burying their heads in the sand, fingers in their ears denying where this is all going. Unbelievable except it’s exactly what I expect from humans.

alarm about what, exactly?

Re: System Card: Claude Mythos Preview [pdf]

#499

Earlier quoted context omitted.

If it's so great at software engineering and bug fixing, then why does Claude Code still have 5000+ open bugs? https://github.com/anthropics/claude-code/issues?q=is%3Aissu... Apparently whatever SWE-bench is measuring isn't very relevant.

as much as I hate cc, 95% of the issues there are either AI psychosis or user error

So it should be insanely easy for this world altering model to comb through them and close irrelevant ones.

Re: System Card: Claude Mythos Preview [pdf]

#500
post #208

Earlier quoted context omitted.

Alignment “appearing” better as model capabilities increase scares the shit out of me, tbh.

Conversely: in humans, intelligence is inversely correlated with crime. It doesn't go to zero, however!

It very much depends on the crime. The truly awful stuff is committed by intelligent people.
Post reply on HN