Live data from Hacker News

System Card: Claude Mythos Preview [pdf]

www-cdn.anthropic.com

661–670 of 687 posts

Re: System Card: Claude Mythos Preview [pdf]

#661
post #13

Combined results (Claude Mythos / Claude Opus 4.6 / GPT-5.4 / Gemini 3.1 Pro) SWE-bench Verified: 93.9% / 80.8% / — / 80.6% SWE-bench Pro: 77.8% / 53.4% / 57.7% / 54.2% SWE-bench Multilingual: 87.3% / 77.8% / — / — SWE-bench Multimodal: 59.0% / 27.1% / — / — Terminal-Bench 2.0: 82.0% / 65.4% / 75.1% / 68.5% GPQA Diamond: 94.5% / 91.3% / 92.8% / 94.3% MMMLU: 92.7% / 91.1% / — / 92.6–93.6% USAMO: 97.6% / 42.3% / 95.2%…

We're gonna need some new benchmarks... ARC-AGI-3 might be the only remaining benchmark below 50%

> We're gonna need some new benchmarks...

You can't consistently benchmark something that is qualitative by nature. I'm struggling to understand how people don't understand this.

Re: System Card: Claude Mythos Preview [pdf]

#662
post #633

Earlier quoted context omitted.

the point is that each question is something that a specialist in a field would be able to do, but deems challenging enough that the ability to solve it would imply significant general usefulness in that domain

I mean they could just feed the solutions into the training data. Then suddenly the bot will do real good at HLE.

Exactly. This is called overfitting and it's most definitely a thing.

Re: System Card: Claude Mythos Preview [pdf]

#663
post #165

Cool on not publicly releasing it. I would assume they've also not connected it to the internet yet? If they have I guess humanity should just keep our collective fingers crossed that they haven't created a model quite capable of escaping yet, or if it is, and may have escaped, lets hope it has no goals of it's own that are incompatible with our own. Also, maybe lets not continue running this experiment to see how fa…

It would have to "escape" to hardware capable of running it, which limits where it could go quite a bit, I'd imagine.

Re: System Card: Claude Mythos Preview [pdf]

#665
post #106

Earlier quoted context omitted.

Yeah this has always been the glaring blind spot for most of the "AI Safety" community; and most of the proposals for "improving" AI safety actually make these risks far worse and far more likely.

It makes quite a lot of sense to focus on reducing the risks of every human everywhere dying, rather than the risks of already existing oppression getting worse.

No, you are deeply misunderstanding the issue. Creating a rivalrous good that powers fight over for control, then use violence to maintain control of, creating a global feudalism, is not "existing oppression getting worse". It actually makes the risks of every human everywhere dying far higher, and even if that doesn't happen, decreases global utility by a similar percentage (99%, instead of 100%). It could actually be worse, if average human utility becomes negative.

Re: System Card: Claude Mythos Preview [pdf]

#666

Earlier quoted context omitted.

This is the notebook filled with exposition you find in post apocalyptic videogames.

It reminds me of Resident Evil in some way. Thank god they are researching AI and not bio-weapons! Then the AI will invent superduper ebola to help a random person have a faster commute or something.

'But wait! You are absolutely right! Distance is an invariant, as is top achievable speed. Let me find a way to actually reduce traffic ahead of you during the same-distance commute ...'

~ Churning ...

Re: System Card: Claude Mythos Preview [pdf]

#667
post #588
post #444

Earlier quoted context omitted.

I read the entire thing fwiw (pseudo-retired life helps with time here). It looks like it was a collaborative effort across multiple teams, where each team (research, security, psycology, etc etc etc) were all submitting ~10 pages or so. It doesn't feel like slop.

Did anything stand out across those 244 pages? Perhaps you have some of your take away thoughts written up somewhere?

Sorry very late reply to this, but ya. I posted here: https://x.com/pwnies/status/2041658034087457236

I'll copy the highlights here, but the tweets have imagery as well:

> The obvious hype - It crushes benchmarks across the board, and it does so with fewer tokens per task.

> Despite this, they don’t think it can self-improve on its own. There are still areas your average engineer does better with, and despite it accelerating tasks by 4x, that only translates to > They’re probably right to hold this back - its ability to exploit things is unprecedented. Any site running on an old stack right now or any traditional industry with outdated software should be terrified if this becomes accessible.

> Counterintuitively, while it’s the most dangerous model, it’s also the safest. They’ve also seen significant additional improvements in safety between their early versions of Mythos and the preview version.

> Anthropic does a really good job of documenting some of the rare dangerous behaviors the early models had. > Interestingly, Mythos itself leaked a recent internal “code related artifact” on github.

> Mythos is also RUTHLESS in Vending Bench. Agent-as-a-CEO might be viable?

> The last thing: Mythos has emergent humor. One of the first models I’ve seen that’s witty. The examples are puns it came up with and witty slack responses it had when operating as a bot.

Re: System Card: Claude Mythos Preview [pdf]

#668

Earlier quoted context omitted.

Are these fair comparisons? It seems like mythos is going to be like a 5.4 ultra or Gemini Deepthink tier model, where access is limited and token usage per query is totally off the charts.

There are a few hints in the doc around this > Importantly, we find that when used in an interactive, synchronous, “hands-on-keyboard” pattern, the benefits of the model were less clear. When used in this fashion, some users perceived Mythos Preview as too slow and did not realize as much value. Autonomous, long-running agent harnesses better elicited the model’s coding capabilities. (p201) ^^ From the surrounding co…

[flagged]

Re: System Card: Claude Mythos Preview [pdf]

#669

Earlier quoted context omitted.

Why couldn't it be unknowable? I am not saying that it is, but it could be. The human brain has its limits and things could me too complex for us to understand enough to be able to modify them at will. We could understand a lot, but not enough to manipulate it with certainty. Biology is not physics.

Because physics is knowable, and I don't think an unknowable thing can be created from a knowable thing.

Why not? Human mind has its limits. The complexity of physics is orders of magnitude smaller than biology, let alone any kind of social science. Physics is the exception, not the rule. The rest of sciences are way more messy.
Post reply on HN