Earlier quoted context omitted.
A jump that we will never be able to use since we're not part of the seemingly minimum 100 billion dollar company club as requirement to be allowed to use it. I get the security aspect, but if we've hit that point any reasonably sophisticated model past this point will be able to do the damage they claim it can do. They might as well be telling us they're closing up shop for consumer models. They should just say they…
More than killer AI I'm afraid of Anthropic/OpenAI going into full rent-seeking mode so that everyone working in tech is forced to fork out loads of money just to stay competitive on the market. These companies can also choose to give exclusive access to hand picked individuals and cut everyone else off and there would be nothing to stop them. This is already happening to some degree, GPT 5.3 Codex's security capabil…
System Card: Claude Mythos Preview [pdf]
221–230 of 687 posts
Re: System Card: Claude Mythos Preview [pdf]
#222"Claude Mythos Preview’s large increase in capabilities has led us to decide not to make it generally available." Disappointing that AGI will be for the powerful only. We are heading for an AI dystopia of Sci-Fi novels.
Unless governments nationalise the companies involved, but then there’s no way our governments of today give this power out to the masses either.
Re: System Card: Claude Mythos Preview [pdf]
#223Earlier quoted context omitted.
Haven't seen a jump this large since I don't even know, years? Too bad they are not releasing it anytime soon (there is no need as they are still currently the leader).
There's speculation that next Tuesday will be a big day for OpenAI and possibly GPT 6. Anthropic showed their hand today.
Re: System Card: Claude Mythos Preview [pdf]
#224While we still have months to a year or two left, I will once again remind people that it's not too late to change our current trajectory. You are not "anti-progress" to not want this future we are building, as you are not "anti-progress" for not wanting your kids to grow up on smart phones and social media. We should remember that not all technology is net-good for humanity, and this technology in particular poses u…
Funny, I was about to say the same thing to you! Life is full of little coincidences.
Re: System Card: Claude Mythos Preview [pdf]
#225Re: System Card: Claude Mythos Preview [pdf]
#226Earlier quoted context omitted.
Are these fair comparisons? It seems like mythos is going to be like a 5.4 ultra or Gemini Deepthink tier model, where access is limited and token usage per query is totally off the charts.
There are a few hints in the doc around this > Importantly, we find that when used in an interactive, synchronous, “hands-on-keyboard” pattern, the benefits of the model were less clear. When used in this fashion, some users perceived Mythos Preview as too slow and did not realize as much value. Autonomous, long-running agent harnesses better elicited the model’s coding capabilities. (p201) ^^ From the surrounding co…
Re: System Card: Claude Mythos Preview [pdf]
#227Earlier quoted context omitted.
A jump that we will never be able to use since we're not part of the seemingly minimum 100 billion dollar company club as requirement to be allowed to use it. I get the security aspect, but if we've hit that point any reasonably sophisticated model past this point will be able to do the damage they claim it can do. They might as well be telling us they're closing up shop for consumer models. They should just say they…
More than killer AI I'm afraid of Anthropic/OpenAI going into full rent-seeking mode so that everyone working in tech is forced to fork out loads of money just to stay competitive on the market. These companies can also choose to give exclusive access to hand picked individuals and cut everyone else off and there would be nothing to stop them. This is already happening to some degree, GPT 5.3 Codex's security capabil…
Re: System Card: Claude Mythos Preview [pdf]
#228> Claude Mythos Preview is, on essentially every dimension we can measure, the best-aligned model that we have released to date by a significant margin. We believe that it does not have any significant coherent misaligned goals, and its character traits in typical conversations closely follow the goals we laid out in our constitution. Even so, we believe that it likely poses the greatest alignment-related risk of any…
Translation: yay, more paternalism.
funny because they do it every time like clockwork acting like their ai is a thunderstorm coming to wipe out the world
Re: System Card: Claude Mythos Preview [pdf]
#229Re: System Card: Claude Mythos Preview [pdf]
#230Earlier quoted context omitted.
And as a bonus: GPT is slow. I’m doing a lot of RE (IDA Pro + MCP), even when 5.4 gives a little bit better guesses (rarely, but happens) - it takes x2-x4 longer. So, it’s just easier to reiterate with Opus
Yeah, need some good RE benchmarks for the LLMs. :) RE is very interesting problem. A lot more that SWE can be RE'd. I've found the LLMs are reluctant to assist, though you can workaround.