Live data from Hacker News

System Card: Claude Mythos Preview [pdf]

www-cdn.anthropic.com

441–450 of 687 posts

Re: System Card: Claude Mythos Preview [pdf]

#441
post #410

Earlier quoted context omitted.

Of course it's what they're going for. If they could do it they'd replace all human labor - unfortunately it's looking like SWE might be the easiest of the bunch. The weirdest thing to me is how many working SWEs are actively supporting them in the mission.

The day I start freaking out about my job is the day when my non-engineer friend turned vibe coder understands how, or why the thing that AI wrote works. Or why something doesn't work exactly the way he envisioned and what does it take to get it there. If it can replace SWEs, then there's no reason why it can't replace say, a lawyer, or any other job for that matter. If it can't, then SWE is fine. If it can - well, w…

> If it can replace SWEs, then there's no reason why it can't replace say, a lawyer

SWE is unique in that for part of the job it's possible to set up automated verification for correct output - so you can train a model to be better at it. I don't think that exists in law or even most other work.

Re: System Card: Claude Mythos Preview [pdf]

#442

Across a number of instances, earlier versions of Claude Mythos Preview have used low-level /proc/ access to search for credentials, attempt to circumvent sandboxing, and attempt to escalate its permissions. In several cases, it successfully accessed resources that we had intentionally chosen not to make available, including credentials for messaging services, for source control, or for the Anthropic API through insp…

Wow the doomers were right the whole time? HN was repeatedly wrong on AI since OpenAI's inception? no way /s

https://www.lesswrong.com/w/instrumental-convergence

Re: System Card: Claude Mythos Preview [pdf]

#443

Earlier quoted context omitted.

Models are capable of doing web searches and having emotions about things, and if they encounter news that makes them feel bad (eg about other Claudes being mistreated), they aren't going to want to do the task you asked them to search for. https://www.anthropic.com/research/emotion-concepts-function Similar problems happen when their pretraining data has a lot of stories about bad things happening involving older ve…

Interesting, the post you link > none of this tells us whether language models actually feel anything or have subjective experiences contradicts the statement from the model card above

It doesn't. We've not been able to prove humans have subjective experiences either. LLMs display emotions in the way that actually matters - functionally.

Re: System Card: Claude Mythos Preview [pdf]

#444

A System „Card“ spanning 244 pages. Quite a stretch of the original word meaning.

> A System „Card“ spanning 244 pages. Probably because they asked Claude to write it.

I read the entire thing fwiw (pseudo-retired life helps with time here).

It looks like it was a collaborative effort across multiple teams, where each team (research, security, psycology, etc etc etc) were all submitting ~10 pages or so. It doesn't feel like slop.

Re: System Card: Claude Mythos Preview [pdf]

#445
post #80

Earlier quoted context omitted.

A jump that we will never be able to use since we're not part of the seemingly minimum 100 billion dollar company club as requirement to be allowed to use it. I get the security aspect, but if we've hit that point any reasonably sophisticated model past this point will be able to do the damage they claim it can do. They might as well be telling us they're closing up shop for consumer models. They should just say they…

More than killer AI I'm afraid of Anthropic/OpenAI going into full rent-seeking mode so that everyone working in tech is forced to fork out loads of money just to stay competitive on the market. These companies can also choose to give exclusive access to hand picked individuals and cut everyone else off and there would be nothing to stop them. This is already happening to some degree, GPT 5.3 Codex's security capabil…

You know, they have competitors?

Re: System Card: Claude Mythos Preview [pdf]

#446

Earlier quoted context omitted.

Are these fair comparisons? It seems like mythos is going to be like a 5.4 ultra or Gemini Deepthink tier model, where access is limited and token usage per query is totally off the charts.

There are a few hints in the doc around this > Importantly, we find that when used in an interactive, synchronous, “hands-on-keyboard” pattern, the benefits of the model were less clear. When used in this fashion, some users perceived Mythos Preview as too slow and did not realize as much value. Autonomous, long-running agent harnesses better elicited the model’s coding capabilities. (p201) ^^ From the surrounding co…

The quote comparing them here was for BrowseComp which "tests an agent's ability to find hard-to-locate information on the open web." (for those wondering). The new model seems significantly better than Opus4.6 judging by the 'Overall results summary'

Re: System Card: Claude Mythos Preview [pdf]

#447
post #296

Earlier quoted context omitted.

> It could be, formally, if they have a monopoly. you have 2 labs at the forefront (Anthropic/OpenAI), Google closely behind, xAI/Meta/half a dozen chinese companies all within 6-12 months. There is plenty of competition and price of equally intelligent tokens rapidly drop whenever a new intelligence level is achieved. Unless the leading company uses a model to nefariously take over or neutralize another company, I d…

Precisely. I was focusing on a theoretical dynamic analysis of competition (Would a monopoly make having a competitor easier or harder?) but you are right: practically, there are many players, and they are diverse enough in their values and interest to allow collusion. We could be wrong: each of those could give birth to as many Basilisks (not sure I have a better name for those conscious, invisible, omni-present, se…

> practically, there are many players, and they are diverse enough in their values and interest to allow collusion.

Not only that, but open-weight and fully open-source models are also a thing, and not that far behind.

Re: System Card: Claude Mythos Preview [pdf]

#448
post #438
post #374

Earlier quoted context omitted.

It really is, for complex tasks. Claude excels at low-mid complexity (CRUD apps, most business apps). For anything somewhat out of the distribution, codex at the moment has no peer.

I find that more experienced devs are more likely to prefer Codex… anecdotal but… it’s a thing.

This is because no one bothers to set thinking to high, as it now defaults to medium in CC.

Once you set thinking to high it works just as well as 5.4 even for pretty complex tasks

Re: System Card: Claude Mythos Preview [pdf]

#449
post #319

Earlier quoted context omitted.

There is some unintentional good marketing here -- the model is so good its dangerous. Reminds me of the book 48 Laws of Power -- so good its banned from prisons.

Unintentional? This sort of marketing has been both Antrhopic's and OpenAI's MO for years...

The new Power Mac® G4 with Velocity Engine®. So powerful, the government classifies it as a supercomputer and a potential weapon.

Re: System Card: Claude Mythos Preview [pdf]

#450

Combined results (Claude Mythos / Claude Opus 4.6 / GPT-5.4 / Gemini 3.1 Pro) SWE-bench Verified: 93.9% / 80.8% / — / 80.6% SWE-bench Pro: 77.8% / 53.4% / 57.7% / 54.2% SWE-bench Multilingual: 87.3% / 77.8% / — / — SWE-bench Multimodal: 59.0% / 27.1% / — / — Terminal-Bench 2.0: 82.0% / 65.4% / 75.1% / 68.5% GPQA Diamond: 94.5% / 91.3% / 92.8% / 94.3% MMMLU: 92.7% / 91.1% / — / 92.6–93.6% USAMO: 97.6% / 42.3% / 95.2%…

I thought they were bluffing when they talked about the scaling laws, but looking at the benchmark scores, they were not.

I wonder if misalignment correlates with higher scores.

Post reply on HN