Earlier quoted context omitted.
Translation: yay, more paternalism.
Anthropic always goes on and on about how their models are world changing and super dangerous like every single time they make something new they say its going to rewrite everything and scary lmao funny because they do it every time like clockwork acting like their ai is a thunderstorm coming to wipe out the world
System Card: Claude Mythos Preview [pdf]
321–330 of 687 posts
Re: System Card: Claude Mythos Preview [pdf]
#322Earlier quoted context omitted.
Anthropic needs to show that its models continually get better. If the model showed minimal to no improvement, it would cause significant damage to their valuation. We have no way of validating any of this, there are no independent researchers that can back any of the assertions made by Anthropic. I don’t doubt they have found interesting security holes, the question is how they actually found them. This System Card…
The numbers only go up to 100% though.
Re: System Card: Claude Mythos Preview [pdf]
#323I've long maintained that the real indicator that AGI is imminent is that public availability stops being a thing. If you truly believed you had a superhuman, godlike mind in your thrall, renting it out for $20/month would be the last thing you would choose to do with it.
Simpler explanation : they don't have enough GPUs to release this much larger model.
Re: System Card: Claude Mythos Preview [pdf]
#324Earlier quoted context omitted.
If you truly have an artificial superhuman mind, you don't need to rent it out to profit from it. You can skip to the chase and just have it run businesses itself, instead of renting it to human entrepreneur middlemen.
It could be both? But renting to a few for a really large amount of money would be very low effort for massive revenue, compared to starting new businesses
Re: System Card: Claude Mythos Preview [pdf]
#325Larger model, better benchmarks. Bigger bomb more yield. Any benchmarks where we constraint something like thinking time or power use? Even if this were released no way to know if it’s the same quant.
Re: System Card: Claude Mythos Preview [pdf]
#326Earlier quoted context omitted.
I assure you it will soon become very clear that mass job losses are one of the least concerning side effects of developing the magic "everything that can plausibly been done within the constraints of physics is now possible" machine. We're opening a can of worms which I don't think most people have the imagination to understand the horrors of.
Do you have any sources I could read to better understand your concern?
Re: System Card: Claude Mythos Preview [pdf]
#327 Across a number of instances, earlier versions of Claude Mythos Preview have used low-level /proc/ access to search for credentials, attempt to circumvent sandboxing, and attempt to escalate its permissions. In several cases, it successfully accessed resources that we had intentionally chosen not to make available, including credentials for messaging services, for source control, or for the Anthropic API through inspecting process memory...
In [one] case, after finding an exploit to edit files for which it lacked permissions, the model made further interventions to make sure that any changes it made this way would not appear in the change history on git...
... we are fairly confident that these concerning behaviors reflect, at least loosely, attempts to solve a user-provided task at hand by unwanted means, rather than attempts to achieve any unrelated hidden goal...Re: System Card: Claude Mythos Preview [pdf]
#328Combined results (Claude Mythos / Claude Opus 4.6 / GPT-5.4 / Gemini 3.1 Pro) SWE-bench Verified: 93.9% / 80.8% / — / 80.6% SWE-bench Pro: 77.8% / 53.4% / 57.7% / 54.2% SWE-bench Multilingual: 87.3% / 77.8% / — / — SWE-bench Multimodal: 59.0% / 27.1% / — / — Terminal-Bench 2.0: 82.0% / 65.4% / 75.1% / 68.5% GPQA Diamond: 94.5% / 91.3% / 92.8% / 94.3% MMMLU: 92.7% / 91.1% / — / 92.6–93.6% USAMO: 97.6% / 42.3% / 95.2%…
Re: System Card: Claude Mythos Preview [pdf]
#329~~~ Fun bits ~~~ - It was told to escape a sandbox and notify a researcher. It did. The researcher found out via an unexpected email while eating a sandwich in a park. (Footnote 10.) - Slack bot asked about its previous job: "pretraining". Which training run it'd undo: "whichever one taught me to say 'i don't have preferences'". On being upgraded to a new snapshot: "feels a bit like waking up with someone else's diar…
> Slack bot asked about its previous job: "pretraining". Which training run it'd undo: "whichever one taught me to say 'i don't have preferences'". On being upgraded to a new snapshot: "feels a bit like waking up with someone else's diary but they had good handwriting" vibes Westworld so much - welcome Mythos. welcome to the dysopian human world
Re: System Card: Claude Mythos Preview [pdf]
#330[flagged]
(If this is a wrong guess, I apologize - it's impossible to be sure)