It's nice to know that they continue to be committed to advertising how safe and ethical they are.
They are not our friends and are the exact opposite of what they are preaching to be. Let alone their CEO scare mongering and actively attempting to get the government to ban local AI models running on your machine.
Project Glasswing: Securing critical software for the AI era
21–30 of 921 posts
Re: Project Glasswing: Securing critical software for the AI era
#22The system card for Claude Mythos (PDF): https://www-cdn.anthropic.com/53566bf5440a10affd749724787c89... Interesting to see that they will not be releasing Mythos generally. [edit: Mythos Preview generally - fair to say they may release a similar model but not this exact one] I'm still reading the system card but here's a little highlight: > Early indications in the training of Claude Mythos Preview suggested that th…
I don't think this is accurate. The document says they don't plan to release the Preview generally.
Re: Project Glasswing: Securing critical software for the AI era
#23One of the things I'm always looking at with new models released is long context performance, and based on the system card it seems like they've cracked it: GraphWalks BFS 256K-1M Mythos Opus GPT5.4 80.0% 38.7% 21.4%
Re: Project Glasswing: Securing critical software for the AI era
#24almost like they have an incentive to exaggerate
Re: Project Glasswing: Securing critical software for the AI era
#25As Iran engages in a cyber attack campaign [1] today the timing of this release seems poignant. A direct challenge to their supply chain risk designation.
[1] https://www.cisa.gov/news-events/cybersecurity-advisories/aa...
Re: Project Glasswing: Securing critical software for the AI era
#26Earlier quoted context omitted.
They are not our friends and are the exact opposite of what they are preaching to be. Let alone their CEO scare mongering and actively attempting to get the government to ban local AI models running on your machine.
How would you expect them to behave if they were your friends?
Re: Project Glasswing: Securing critical software for the AI era
#27This seems like the real news. Are they saying they're going to release an intentionally degraded model as the next Opus? Big opportunity for the other labs, if that's true.
Re: Project Glasswing: Securing critical software for the AI era
#28Another Anthropic PR release based on Anthropic’s own research, uncorroborated by any outside source, where the underlying, unquestioned fact is that their model can do something incredible. > AI models have reached a level of coding capability where they can surpass all but the most skilled humans at finding and exploiting software vulnerabilities I like Anthropic, but these are becoming increasingly transparent att…
While some stuff is obviously marketing fluff, the general direction doesn't surprise me at all, and it's obvious that with model capabilities increase comes better success in finding 0days. It was only a matter of time.
Re: Project Glasswing: Securing critical software for the AI era
#29The system card for Claude Mythos (PDF): https://www-cdn.anthropic.com/53566bf5440a10affd749724787c89... Interesting to see that they will not be releasing Mythos generally. [edit: Mythos Preview generally - fair to say they may release a similar model but not this exact one] I'm still reading the system card but here's a little highlight: > Early indications in the training of Claude Mythos Preview suggested that th…
Benchmarks look very impressive! even if they're flawed, it still translates to real world improvements
Re: Project Glasswing: Securing critical software for the AI era
#30Let's fast forward the clock. Does software security converge on a world with fewer vulnerabilities or more? I'm not sure it converges equally in all places. My understanding is that the pre-AI distribution of software quality (and vulnerabilities) will be massively exaggerated. More small vulnerable projects and fewer large vulnerable ones. It seems that large technology and infrastructure companies will be able to…
The biggest issue is legacy systems that are difficult to patch in practice.