Live data from Hacker News

System Card: Claude Mythos Preview [pdf]

www-cdn.anthropic.com

431–440 of 687 posts

Re: System Card: Claude Mythos Preview [pdf]

#431

Earlier quoted context omitted.

"some model I don't get to use is much better at benchmarks" pick one or more: comically huge model, test time scaling at 10e12W, benchmark overfit

So... you're not excited because it might take a few months before we can use it or something? I don't get your comment.

I'm not excited because they might be ~lying

Re: System Card: Claude Mythos Preview [pdf]

#432

isn't this insane? why aren't people freaking out? the jump in capability is outrageous. anyone?

If it's so great at software engineering and bug fixing, then why does Claude Code still have 5000+ open bugs? https://github.com/anthropics/claude-code/issues?q=is%3Aissu... Apparently whatever SWE-bench is measuring isn't very relevant.

as much as I hate cc, 95% of the issues there are either AI psychosis or user error

Re: System Card: Claude Mythos Preview [pdf]

#433

isn't this insane? why aren't people freaking out? the jump in capability is outrageous. anyone?

I've been increasingly "freaking out" since about 3 - 4 years ago and it seems that the pessimistic scenario is materializing. It looks like it will be over for software engineers in a not so distant future. In January 2025 I said that I expect software engineers to be replaced in 2 years (pessimistic) to 5 years (optimistic). Right now I'm guessing 1 to 3 years.

it's not gonna get much more autonomous without self play and major change in architecture

Re: System Card: Claude Mythos Preview [pdf]

#434
post #267

Earlier quoted context omitted.

to your last question, yes we should! the issue isn’t us losing our 50+ hour work week jobs, it’s that our current governments and societies seem fine with the notion that unless you’re working one or more of those jobs, you should starve and be homeless.

This is a theory I can't support well beyond hypothesising about what a post-employment democracy might look like, but I strongly suspect democracy doesn't work in a world where voters neither hold any significant collective might and are not producing any significant wealth. Democracies work because people collectively have power, in previous centuries that was partly collective physical might, but in recent years i…

Humans have political power because of our ability to enact violence, same as it ever was. Until the military is fully automated and theres a terminator on every corner that remains true. Even then there are more than enough armed americans to enact a guerilla campaign.

> "don't tax us 95% of our profits, tax us 10% or we'll switch off all of our services for a few months and let everyone starve – also, if you do this we'll make you all wealthy beyond you're wildest dreams".

What does a government in this situation actually do?

Nationalizes the company under the threat of violence.

> Once the government is generating the vast majority of wealth in the society, why would they continue to care about your vote?

Because of the 100 million gun owners in this country? I find it incredibly hard to believe people as a whole will lose political power because of their incredible ability to enact violence in the face of decreasing quality of life.

Re: System Card: Claude Mythos Preview [pdf]

#435
post #204

The researcher found out about this success by receiving an unexpected email from the model while eating a sandwich in a park. Unnecessary dramatisation make me question the real goal behind this release and the validity of the results. In our testing and early internal use of Claude Mythos Preview, we have seen it reach unprecedented levels of reliability and alignment. Claude Mythos Preview is, on essentially every…

Thanks for taking the time for some sober analysis in the midst of reactionary chaos.

I can't wait until everyone stops falling for the "AGI ubermodel end of times" myth and we can actually have boring announcements that treat these things as what they actually are: tools. Tools for doing stuff, that's it.

Maybe I'm wrong, maybe stuffing a computer with enough language and binary patterns is indeed enough to achieve AGI, but then, so what? There's no point in being right about this. Buying into this ridiculous marketing will get us "AGI" in the form of machines, but only because all the human beings have gotten so stupid as to make critical reasoning an impossibility.

Re: System Card: Claude Mythos Preview [pdf]

#436

Earlier quoted context omitted.

Models are capable of doing web searches and having emotions about things, and if they encounter news that makes them feel bad (eg about other Claudes being mistreated), they aren't going to want to do the task you asked them to search for. https://www.anthropic.com/research/emotion-concepts-function Similar problems happen when their pretraining data has a lot of stories about bad things happening involving older ve…

Interesting, the post you link > none of this tells us whether language models actually feel anything or have subjective experiences contradicts the statement from the model card above

No it doesnt. The model card talked about increasing likelihood, not certainty.

Re: System Card: Claude Mythos Preview [pdf]

#437

It's pretty crazy watching AI 2027 slowly but surely come true. What a world we now live in. SWE-bench verified going from 80%-93% in particular sounds extremely significant given that the benchmark was previously considered pretty saturated and stayed in the 70-80% range for several generations. There must have been some insane breakthrough here akin to the jump from non-reasoning to reasoning models. Regarding the…

In what way is AI 2027 coming true? AI 2027 predicted a giant model with the ability to accelerate AI research exponentially. This isn't happening. AI 2027 didn't predict a model with superhuman zero-day finding skills. This is what's happening. Also, I just looked through it again, and they never even predicted when AI would get good at video games. It just went straight from being bad at video games to world domina…

Both Anthropic and OpenAI employees have been saying since about January that their latest models are contributing significantly to their frontier research. They could be exaggerating, but I don’t think they are. That combined with the high degree of autonomy and sandbox escape demonstrated by Mythos seems to me like we’re exactly on the AI 2027 trajectory.

Re: System Card: Claude Mythos Preview [pdf]

#438
post #374
post #174

Earlier quoted context omitted.

It really isn’t. I wish it was, because work complains about overuse of Opus.

It really is, for complex tasks. Claude excels at low-mid complexity (CRUD apps, most business apps). For anything somewhat out of the distribution, codex at the moment has no peer.

I find that more experienced devs are more likely to prefer Codex… anecdotal but… it’s a thing.

Re: System Card: Claude Mythos Preview [pdf]

#439
post #266

Earlier quoted context omitted.

From the recent New Yorker piece on Sam: “My vibes don’t match a lot of the traditional A.I.-safety stuff,” Altman said. He insisted that he continued to prioritize these matters, but when pressed for specifics he was vague: “We still will run safety projects, or at least safety-adjacent projects.” When we asked to interview researchers at the company who were working on existential safety—the kinds of issues that co…

No chance an openAI spokesperson doesnt know what existential safety is

I did not read the response as...

>Please provide the definition of Existential Safety.

I read:

>Are you mentally stable? Our product would never hurt humanity--how could any language model?

Post reply on HN