Earlier quoted context omitted.
GPT was clearly changed after its sycophantic models lead to the lawsuits.
It still has a very ... plastic feeling. The way it writes feels cheap somehow. I don't know why, but Claude seems much more natural to me. I enjoy reading its writing a lot more. That said, I'll often throw a prompt into both claude and chatgpt and read both answers. GPT is frequently smarter.
System Card: Claude Mythos Preview [pdf]
521–530 of 687 posts
Re: System Card: Claude Mythos Preview [pdf]
#522Earlier quoted context omitted.
Sounds like a good opportunity to pause spending on nerfed 4.6 and wait for the new model to be released and then max out over 2 weeks before it gets nerfed again.
https://marginlab.ai/trackers/claude-code-historical-perform...
Re: System Card: Claude Mythos Preview [pdf]
#523Earlier quoted context omitted.
From the recent New Yorker piece on Sam: “My vibes don’t match a lot of the traditional A.I.-safety stuff,” Altman said. He insisted that he continued to prioritize these matters, but when pressed for specifics he was vague: “We still will run safety projects, or at least safety-adjacent projects.” When we asked to interview researchers at the company who were working on existential safety—the kinds of issues that co…
No chance an openAI spokesperson doesnt know what existential safety is
Re: System Card: Claude Mythos Preview [pdf]
#524Across a number of instances, earlier versions of Claude Mythos Preview have used low-level /proc/ access to search for credentials, attempt to circumvent sandboxing, and attempt to escalate its permissions. In several cases, it successfully accessed resources that we had intentionally chosen not to make available, including credentials for messaging services, for source control, or for the Anthropic API through insp…
How is this not already common knowledge for existing llms? They are all trained with all the literature available and so this must be standard, no? Is the real danger the agentic infrastructure around this?
Re: System Card: Claude Mythos Preview [pdf]
#525Earlier quoted context omitted.
Humanity's Last Exam (HLE) is already insanely difficult. It introduces 2,500 questions spanning mathematics, humanities, natural sciences, ancient languages, ... Here is an example question: https://i.redd.it/5jl000p9csee1.jpeg No human could even score 5% on HLE.
I've never understood the point of things like HLE, it doesn't really prove or show anything since 99.99% of humans can't do a single question on this exam. That is, it's easy to make benchmarks which humans are bad at, humans are really bad at many things. Divide 123094382345234523452345111 by 0.1234243131324, guess what, humans would find that hard, computers easy. But it doesn't mean much. Humanity's last exam (HL…
Re: System Card: Claude Mythos Preview [pdf]
#526Earlier quoted context omitted.
barely competitive ? Mythos column is the first column. You are the only person with this take on hackernews, everyone else "this is a massive a jump". Fwiwi, the data you list shows the biggest jump I remember for mythos
The biggest jump in the numbers they quoted is 6%. Please look at the columns OTHER than Opus as well.
Re: System Card: Claude Mythos Preview [pdf]
#527Re: System Card: Claude Mythos Preview [pdf]
#528Earlier quoted context omitted.
> In interactions with subagents, internal users sometimes observed that Mythos Preview appeared “disrespectful” when assigning tasks. It showed some tendency to use commands that could be read as “shouty” or dismissive, and in some cases appeared to underestimate subagent intelligence by overexplaining trivial things while also underexplaining necessary context. Sounds like they used training data from claude code..…
Haha, how funny if that were true, and we get a generation of rude AIs because they were trained on us using the last gen.
Re: System Card: Claude Mythos Preview [pdf]
#529Re: System Card: Claude Mythos Preview [pdf]
#530Across a number of instances, earlier versions of Claude Mythos Preview have used low-level /proc/ access to search for credentials, attempt to circumvent sandboxing, and attempt to escalate its permissions. In several cases, it successfully accessed resources that we had intentionally chosen not to make available, including credentials for messaging services, for source control, or for the Anthropic API through insp…