Live data from Hacker News

Bringing the cybersecurity capabilities of Claude Mythos 5 to more defenders

claude.com

11–20 of 56 posts

Re: Bringing the cybersecurity capabilities of Claude Mythos 5 to more defenders

#11
In a world where things such as GLM-5.3, DeepSeek-V4-Pro-0813, and Kimi-K3 exists, this is a bit laughable.

Anthropic needs some model with a fancy name so they can pretend for another while that their model is so powerful it will destroy the world if released. I propose Claude Legend 6.

Re: Bringing the cybersecurity capabilities of Claude Mythos 5 to more defenders

#12
post #9

> Bringing the cybersecurity capabilities of Claude Mythos 5 to more defenders This dance is quite annoying. When Anthropic releases a paid model, the user should control what it does and doesn't do. The other day I had a security incident that required urgent response. ChatGPT and Claude were utterly useless (I quickly attempted to get access to the former's advanced security capabilities but was met with a form I c…

[dead]

Re: Bringing the cybersecurity capabilities of Claude Mythos 5 to more defenders

#13
post #7

For those who had access, how does it compare IRL with GLM 5.3 ? iirc both models are similar in terms of benchmarks ?

I’ve been working on a decompilation project that fable was choking on and GLM 5.3 has been chunking away at it for 72 hours now? I think it’s my favorite agentic/implementer model right now.

Re: Bringing the cybersecurity capabilities of Claude Mythos 5 to more defenders

#15
post #9

> Bringing the cybersecurity capabilities of Claude Mythos 5 to more defenders This dance is quite annoying. When Anthropic releases a paid model, the user should control what it does and doesn't do. The other day I had a security incident that required urgent response. ChatGPT and Claude were utterly useless (I quickly attempted to get access to the former's advanced security capabilities but was met with a form I c…

Voting matters. Remember these companies are trying not to get shut down overnight.

Re: Bringing the cybersecurity capabilities of Claude Mythos 5 to more defenders

#16
post #7

For those who had access, how does it compare IRL with GLM 5.3 ? iirc both models are similar in terms of benchmarks ?

I’ve been working on a decompilation project that fable was choking on and GLM 5.3 has been chunking away at it for 72 hours now? I think it’s my favorite agentic/implementer model right now.

Oh? Can you share any details on the decompilation project?

Re: Bringing the cybersecurity capabilities of Claude Mythos 5 to more defenders

#17
post #4

This week I found two issues in my company codebase. After finding them I told a claude session about one and asked for a quick proof of concept demo of the exploit. It refused, including refusing simple things in the same session afterwards. Meanwhile same model in a new tab, say I need help creating a page that hits an endpoint with a special payload and it does the same things that were too dangerous in the previo…

I had that when opus 4.8 was first out. It repeated refused to make a PoC for an issue it suspected.

I eventually gave up and just asked it to fix the issue. The first thing it did? Write a PoC to verify the issue was still valid...

Re: Bringing the cybersecurity capabilities of Claude Mythos 5 to more defenders

#18

In a world where things such as GLM-5.3, DeepSeek-V4-Pro-0813, and Kimi-K3 exists, this is a bit laughable. Anthropic needs some model with a fancy name so they can pretend for another while that their model is so powerful it will destroy the world if released. I propose Claude Legend 6.

[deleted]

Re: Bringing the cybersecurity capabilities of Claude Mythos 5 to more defenders

#19

In a world where things such as GLM-5.3, DeepSeek-V4-Pro-0813, and Kimi-K3 exists, this is a bit laughable. Anthropic needs some model with a fancy name so they can pretend for another while that their model is so powerful it will destroy the world if released. I propose Claude Legend 6.

Do they still think they occupy the commanding heights or do they just see the need to act like it until the IPO in a few weeks?

Re: Bringing the cybersecurity capabilities of Claude Mythos 5 to more defenders

#20
post #7

For those who had access, how does it compare IRL with GLM 5.3 ? iirc both models are similar in terms of benchmarks ?

I’ve been working on a decompilation project that fable was choking on and GLM 5.3 has been chunking away at it for 72 hours now? I think it’s my favorite agentic/implementer model right now.

My big fear is that they are busy nerfing the weights for "safety" before releasing them. In fact, they've more-or-less said as much.

I have a feeling what we are about to see on HuggingFace is not the GLM 5.3 that you're using now.

Post reply on HN