Show HN: We post-trained a model that pen tests instead of refusing
21–30 of 49 posts
Re: Show HN: We post-trained a model that pen tests instead of refusing
#22> This won't be made available to anyone and everyone, but we do believe that responsible SMEs and midmarket companies also need access to these tools in order to identify key vulnerabilities in their systems; not just enterprises. So this is the same policy that Anthropic and OpenAI have, it is just based on your criteria rather than theirs.
I think the policy universally makes sense, who would want to give a tool like this to bad actors? But it does leave a big section of the market underserved. Particularly when Mythos was made accessible to very large orgs and then Fable was pulled on export grounds.
Re: Show HN: We post-trained a model that pen tests instead of refusing
#23Earlier quoted context omitted.
The tool is live, you can test it.
No, you can’t. This page is a sales funnel to schedule a 30 minute video chat with Cosine.ai or argusred or whatever. The thing you can test is not the thing that the headline is talking about. It’s just more “We’re so smart we invented the boogeyman, trust us” slop marketing that’s been happening since gpt-2
Re: Show HN: We post-trained a model that pen tests instead of refusing
#24Earlier quoted context omitted.
No, you can’t. This page is a sales funnel to schedule a 30 minute video chat with Cosine.ai or argusred or whatever. The thing you can test is not the thing that the headline is talking about. It’s just more “We’re so smart we invented the boogeyman, trust us” slop marketing that’s been happening since gpt-2
Did you follow the link? There is a brew install binary you can install and test. It's live.
If I wanted to show off a “model that pen tests” I’d at least include a gif of it running against Juice Shop or something before the spooky language and “schedule a sales call”
Re: Show HN: We post-trained a model that pen tests instead of refusing
#25> This won't be made available to anyone and everyone, but we do believe that responsible SMEs and midmarket companies also need access to these tools in order to identify key vulnerabilities in their systems; not just enterprises. So this is the same policy that Anthropic and OpenAI have, it is just based on your criteria rather than theirs.
I think the policy universally makes sense, who would want to give a tool like this to bad actors? But it does leave a big section of the market underserved. Particularly when Mythos was made accessible to very large orgs and then Fable was pulled on export grounds.
Re: Show HN: We post-trained a model that pen tests instead of refusing
#26as a side note - I think it's very unprofessional and very shitty to not mention kimi2.6 at all in your marketing copy. and i feel that you posted that in this hn post begrudgingly since the hn crowd would have flagged that. confirmed with a google search too: https://www.google.com/search?q=kimi+site%3Aargusred.com
All around your marketing website you keep mentioning - 'A model lab built it'. A fintune does not maketh you a model lab - some humility please :)
finally - doesn't Kimi's licensing prohibit you from not mentioning them? Didn't cursor run into the same issue?
Re: Show HN: We post-trained a model that pen tests instead of refusing
#27This in its own right proves that the defenses of Fable and others are temporary blocks, and AI based hacking is going to be effectively available to all parties regardless of stop gaps, as long as open models exist.
Re: Show HN: We post-trained a model that pen tests instead of refusing
#28Earlier quoted context omitted.
Did you follow the link? There is a brew install binary you can install and test. It's live.
> Gated because the security implications are real; access is via booking If I wanted to show off a “model that pen tests” I’d at least include a gif of it running against Juice Shop or something before the spooky language and “schedule a sales call”
Re: Show HN: We post-trained a model that pen tests instead of refusing
#29Re: Show HN: We post-trained a model that pen tests instead of refusing
#30IMO the most interesting thing about this is Kimi K2.6, an extremely capable model, can be relatively easily post-trained to allow pen tests. This in its own right proves that the defenses of Fable and others are temporary blocks, and AI based hacking is going to be effectively available to all parties regardless of stop gaps, as long as open models exist.