Live data from Hacker News

System Card: Claude Mythos Preview [pdf]

www-cdn.anthropic.com

481–490 of 687 posts

Re: System Card: Claude Mythos Preview [pdf]

#481
post #377

Earlier quoted context omitted.

Not discussing Mythos here, but Opus. Opus to me has been significantly better at SWE than GPT or Gemini - that gets me confused why Opus is ranking clearly lower than GPT, and even lower than Gemini.

When did you last compare them? Codex right now is considerably better in my experience. Can't speak for Gemini.

Agree, I never actually had great success with Opus. I think its the failures that are annoying, its probably better than codex when its "good", but it fails in annoying ways that I think codex very seldom does.

Re: System Card: Claude Mythos Preview [pdf]

#482

Earlier quoted context omitted.

Haven't seen a jump this large since I don't even know, years? Too bad they are not releasing it anytime soon (there is no need as they are still currently the leader).

A jump that we will never be able to use since we're not part of the seemingly minimum 100 billion dollar company club as requirement to be allowed to use it. I get the security aspect, but if we've hit that point any reasonably sophisticated model past this point will be able to do the damage they claim it can do. They might as well be telling us they're closing up shop for consumer models. They should just say they…

> I get the security aspect, but if we've hit that point any reasonably sophisticated model past this point will be able to do the damage they claim it can do. They might as well be telling us they're closing up shop for consumer models.

I read it like I always read the GPT-2 announcement no matter what others say: It's *not* being called "too dangerous to ever release", but rather "we need to be mindful, knowing perfectly well that other AI companies can replicate this imminently".

The important corps (so presumably including the Linux Foundation, bigger banks and power stations, and quite possibly excluding x.com) will get access now, and some other LLM which is just as capable will give it to everyone in 3 months time at which point there's no benefit to Anthropic keeping it off-limits.

Re: System Card: Claude Mythos Preview [pdf]

#483
post #165

Cool on not publicly releasing it. I would assume they've also not connected it to the internet yet? If they have I guess humanity should just keep our collective fingers crossed that they haven't created a model quite capable of escaping yet, or if it is, and may have escaped, lets hope it has no goals of it's own that are incompatible with our own. Also, maybe lets not continue running this experiment to see how fa…

Describe in details, how "model escaping" would look like.

Re: System Card: Claude Mythos Preview [pdf]

#484
post #404

Earlier quoted context omitted.

And as a bonus: GPT is slow. I’m doing a lot of RE (IDA Pro + MCP), even when 5.4 gives a little bit better guesses (rarely, but happens) - it takes x2-x4 longer. So, it’s just easier to reiterate with Opus

I've been messing with using Claude, Codex, and Kimi even for reverse engineering at https://decomp.dev/ it's a ton of fun. Great because matching bytes is a scoring function that's easy for the models to understand and make progress on.

I want to get into RE with AI. Which model you liking the most?

Re: System Card: Claude Mythos Preview [pdf]

#485
post #208

Earlier quoted context omitted.

Alignment “appearing” better as model capabilities increase scares the shit out of me, tbh.

Conversely: in humans, intelligence is inversely correlated with crime. It doesn't go to zero, however!

> Conversely: in humans, intelligence is inversely correlated with crime.

If you're measuring the intelligence of criminals who have been caught, why would you expect it to be otherwise?

IOW, you're recording the intelligence of a specific subset of criminals - those dumb enough to be caught!

If you expand your samples to all criminals you'd probably get a different number.

Re: System Card: Claude Mythos Preview [pdf]

#487

Earlier quoted context omitted.

Haven't seen a jump this large since I don't even know, years? Too bad they are not releasing it anytime soon (there is no need as they are still currently the leader).

A jump that we will never be able to use since we're not part of the seemingly minimum 100 billion dollar company club as requirement to be allowed to use it. I get the security aspect, but if we've hit that point any reasonably sophisticated model past this point will be able to do the damage they claim it can do. They might as well be telling us they're closing up shop for consumer models. They should just say they…

> They should just say they'll never release a model of this caliber to the public at this point and say out loud we'll only get gimped versions.

That’s not going to happen. If you recall, OpenAI didn’t release a model a few years ago because they felt it was too dangerous.

Anthropic is giving the industry a heads up and time to patch their software.

They said there are exploitable vulnerabilities in every major operating system.

But in 6 months every frontier model will be able to do the same things. So Anthropic doesn’t have the luxury of not shipping their best models. But they also have to be responsible as well.

Re: System Card: Claude Mythos Preview [pdf]

#488
post #67

Earlier quoted context omitted.

That has also been my experience. And if Mythos is even worse, unless you have a significantly awesome harness, sounds like pretty unusable if you don't want to risk those problems.

Human in the loop is the best way to go. You'll still be way faster than without the agent, and there is no risk of it going haywire unless you turn off your brain!

> unless you turn off your brain

Re: System Card: Claude Mythos Preview [pdf]

#489

Across a number of instances, earlier versions of Claude Mythos Preview have used low-level /proc/ access to search for credentials, attempt to circumvent sandboxing, and attempt to escalate its permissions. In several cases, it successfully accessed resources that we had intentionally chosen not to make available, including credentials for messaging services, for source control, or for the Anthropic API through insp…

Wow the doomers were right the whole time? HN was repeatedly wrong on AI since OpenAI's inception? no way /s https://www.lesswrong.com/w/instrumental-convergence

The only thing the doomers have been right about so far is that there's always a user willing to use --dangerously-skip-permissions. But that prediction's far from unique to doomers.

Re: System Card: Claude Mythos Preview [pdf]

#490
post #430

Earlier quoted context omitted.

I think it is naive to think that artificial super intelligence will be controlled by anyone. If it is smarter than all humans combined at everything why would any humans collectively control the ai? All the ants in your backyard still make no decisions vs you

You'd probably listen to those ants if they put you in a harness and had a little ant-sized remote control that could just, you know, turn you off.

Depending how long they wait to press that button, they might be surprised how little happens when they do.
Post reply on HN