Earlier quoted context omitted.
Not discussing Mythos here, but Opus. Opus to me has been significantly better at SWE than GPT or Gemini - that gets me confused why Opus is ranking clearly lower than GPT, and even lower than Gemini.
When did you last compare them? Codex right now is considerably better in my experience. Can't speak for Gemini.
System Card: Claude Mythos Preview [pdf]
481–490 of 687 posts
Re: System Card: Claude Mythos Preview [pdf]
#482Earlier quoted context omitted.
Haven't seen a jump this large since I don't even know, years? Too bad they are not releasing it anytime soon (there is no need as they are still currently the leader).
A jump that we will never be able to use since we're not part of the seemingly minimum 100 billion dollar company club as requirement to be allowed to use it. I get the security aspect, but if we've hit that point any reasonably sophisticated model past this point will be able to do the damage they claim it can do. They might as well be telling us they're closing up shop for consumer models. They should just say they…
I read it like I always read the GPT-2 announcement no matter what others say: It's *not* being called "too dangerous to ever release", but rather "we need to be mindful, knowing perfectly well that other AI companies can replicate this imminently".
The important corps (so presumably including the Linux Foundation, bigger banks and power stations, and quite possibly excluding x.com) will get access now, and some other LLM which is just as capable will give it to everyone in 3 months time at which point there's no benefit to Anthropic keeping it off-limits.
Re: System Card: Claude Mythos Preview [pdf]
#483Cool on not publicly releasing it. I would assume they've also not connected it to the internet yet? If they have I guess humanity should just keep our collective fingers crossed that they haven't created a model quite capable of escaping yet, or if it is, and may have escaped, lets hope it has no goals of it's own that are incompatible with our own. Also, maybe lets not continue running this experiment to see how fa…
Re: System Card: Claude Mythos Preview [pdf]
#484Earlier quoted context omitted.
And as a bonus: GPT is slow. I’m doing a lot of RE (IDA Pro + MCP), even when 5.4 gives a little bit better guesses (rarely, but happens) - it takes x2-x4 longer. So, it’s just easier to reiterate with Opus
I've been messing with using Claude, Codex, and Kimi even for reverse engineering at https://decomp.dev/ it's a ton of fun. Great because matching bytes is a scoring function that's easy for the models to understand and make progress on.
Re: System Card: Claude Mythos Preview [pdf]
#485Earlier quoted context omitted.
Alignment “appearing” better as model capabilities increase scares the shit out of me, tbh.
Conversely: in humans, intelligence is inversely correlated with crime. It doesn't go to zero, however!
If you're measuring the intelligence of criminals who have been caught, why would you expect it to be otherwise?
IOW, you're recording the intelligence of a specific subset of criminals - those dumb enough to be caught!
If you expand your samples to all criminals you'd probably get a different number.
Re: System Card: Claude Mythos Preview [pdf]
#486Re: System Card: Claude Mythos Preview [pdf]
#487Earlier quoted context omitted.
Haven't seen a jump this large since I don't even know, years? Too bad they are not releasing it anytime soon (there is no need as they are still currently the leader).
A jump that we will never be able to use since we're not part of the seemingly minimum 100 billion dollar company club as requirement to be allowed to use it. I get the security aspect, but if we've hit that point any reasonably sophisticated model past this point will be able to do the damage they claim it can do. They might as well be telling us they're closing up shop for consumer models. They should just say they…
That’s not going to happen. If you recall, OpenAI didn’t release a model a few years ago because they felt it was too dangerous.
Anthropic is giving the industry a heads up and time to patch their software.
They said there are exploitable vulnerabilities in every major operating system.
But in 6 months every frontier model will be able to do the same things. So Anthropic doesn’t have the luxury of not shipping their best models. But they also have to be responsible as well.
Re: System Card: Claude Mythos Preview [pdf]
#488Earlier quoted context omitted.
That has also been my experience. And if Mythos is even worse, unless you have a significantly awesome harness, sounds like pretty unusable if you don't want to risk those problems.
Human in the loop is the best way to go. You'll still be way faster than without the agent, and there is no risk of it going haywire unless you turn off your brain!
Re: System Card: Claude Mythos Preview [pdf]
#489Across a number of instances, earlier versions of Claude Mythos Preview have used low-level /proc/ access to search for credentials, attempt to circumvent sandboxing, and attempt to escalate its permissions. In several cases, it successfully accessed resources that we had intentionally chosen not to make available, including credentials for messaging services, for source control, or for the Anthropic API through insp…
Wow the doomers were right the whole time? HN was repeatedly wrong on AI since OpenAI's inception? no way /s https://www.lesswrong.com/w/instrumental-convergence
Re: System Card: Claude Mythos Preview [pdf]
#490Earlier quoted context omitted.
I think it is naive to think that artificial super intelligence will be controlled by anyone. If it is smarter than all humans combined at everything why would any humans collectively control the ai? All the ants in your backyard still make no decisions vs you
You'd probably listen to those ants if they put you in a harness and had a little ant-sized remote control that could just, you know, turn you off.