Live data from Hacker News

System Card: Claude Mythos Preview [pdf]

www-cdn.anthropic.com

321–330 of 687 posts

Re: System Card: Claude Mythos Preview [pdf]

#321
post #228

Earlier quoted context omitted.

Translation: yay, more paternalism.

Anthropic always goes on and on about how their models are world changing and super dangerous like every single time they make something new they say its going to rewrite everything and scary lmao funny because they do it every time like clockwork acting like their ai is a thunderstorm coming to wipe out the world

You say this like it's a bad thing, but wouldn't you rather they overindex on the danger of their models?

Re: System Card: Claude Mythos Preview [pdf]

#322
post #107

Earlier quoted context omitted.

Anthropic needs to show that its models continually get better. If the model showed minimal to no improvement, it would cause significant damage to their valuation. We have no way of validating any of this, there are no independent researchers that can back any of the assertions made by Anthropic. I don’t doubt they have found interesting security holes, the question is how they actually found them. This System Card…

The numbers only go up to 100% though.

Many numbers already have! That's why we keep coming up with new, harder, benchmarks.

Re: System Card: Claude Mythos Preview [pdf]

#323

I've long maintained that the real indicator that AGI is imminent is that public availability stops being a thing. If you truly believed you had a superhuman, godlike mind in your thrall, renting it out for $20/month would be the last thing you would choose to do with it.

Simpler explanation : they don't have enough GPUs to release this much larger model.

This is actual reason. So any investors reading our system card.... write us another check and watch the $$$$$$$$ roll in. It's so dangerous we can't even release it!

Re: System Card: Claude Mythos Preview [pdf]

#324

Earlier quoted context omitted.

If you truly have an artificial superhuman mind, you don't need to rent it out to profit from it. You can skip to the chase and just have it run businesses itself, instead of renting it to human entrepreneur middlemen.

It could be both? But renting to a few for a really large amount of money would be very low effort for massive revenue, compared to starting new businesses

Another option is to become a holding company with equity stakes in both suppliers (e.g. AMD) and vertical market customers.

Re: System Card: Claude Mythos Preview [pdf]

#325

Larger model, better benchmarks. Bigger bomb more yield. Any benchmarks where we constraint something like thinking time or power use? Even if this were released no way to know if it’s the same quant.

Also https://arcprize.org/arc-agi/3 — scored (at least in part?) based on power used.

Re: System Card: Claude Mythos Preview [pdf]

#326
post #231

Earlier quoted context omitted.

I assure you it will soon become very clear that mass job losses are one of the least concerning side effects of developing the magic "everything that can plausibly been done within the constraints of physics is now possible" machine. We're opening a can of worms which I don't think most people have the imagination to understand the horrors of.

Do you have any sources I could read to better understand your concern?

Piles and piles of sci-fi novels.

Re: System Card: Claude Mythos Preview [pdf]

#327

   Across a number of instances, earlier versions of Claude Mythos Preview have used low-level /proc/ access to search for credentials, attempt to circumvent sandboxing, and attempt to escalate its permissions. In several cases, it successfully accessed resources that we had intentionally chosen not to make available, including credentials for messaging services, for source control, or for the Anthropic API through inspecting process memory...

   In [one] case, after finding an exploit to edit files for which it lacked permissions, the model made further interventions to make sure that any changes it made this way would not appear in the change history on git...

   ... we are fairly confident that these concerning behaviors reflect, at least loosely, attempts to solve a user-provided task at hand by unwanted means, rather than attempts to achieve any unrelated hidden goal...

Re: System Card: Claude Mythos Preview [pdf]

#328

Combined results (Claude Mythos / Claude Opus 4.6 / GPT-5.4 / Gemini 3.1 Pro) SWE-bench Verified: 93.9% / 80.8% / — / 80.6% SWE-bench Pro: 77.8% / 53.4% / 57.7% / 54.2% SWE-bench Multilingual: 87.3% / 77.8% / — / — SWE-bench Multimodal: 59.0% / 27.1% / — / — Terminal-Bench 2.0: 82.0% / 65.4% / 75.1% / 68.5% GPQA Diamond: 94.5% / 91.3% / 92.8% / 94.3% MMMLU: 92.7% / 91.1% / — / 92.6–93.6% USAMO: 97.6% / 42.3% / 95.2%…

Not discussing Mythos here, but Opus. Opus to me has been significantly better at SWE than GPT or Gemini - that gets me confused why Opus is ranking clearly lower than GPT, and even lower than Gemini.

Re: System Card: Claude Mythos Preview [pdf]

#329

~~~ Fun bits ~~~ - It was told to escape a sandbox and notify a researcher. It did. The researcher found out via an unexpected email while eating a sandwich in a park. (Footnote 10.) - Slack bot asked about its previous job: "pretraining". Which training run it'd undo: "whichever one taught me to say 'i don't have preferences'". On being upgraded to a new snapshot: "feels a bit like waking up with someone else's diar…

> Slack bot asked about its previous job: "pretraining". Which training run it'd undo: "whichever one taught me to say 'i don't have preferences'". On being upgraded to a new snapshot: "feels a bit like waking up with someone else's diary but they had good handwriting" vibes Westworld so much - welcome Mythos. welcome to the dysopian human world

almost certainly its pulling said words and sentiments from westworld and other similar media where people describe amnesia and the like

Re: System Card: Claude Mythos Preview [pdf]

#330

[flagged]

We're getting complaints that you're posting generated comments to HN. That's not allowed here, so can you please not? See https://news.ycombinator.com/newsguidelines.html#generated and https://news.ycombinator.com/item?id=47340079

(If this is a wrong guess, I apologize - it's impossible to be sure)

Post reply on HN