It’s clear that Anthropic has run out of the compute capacity needed to serve Mythos publicly. They’re using security concerns to mask their inability to deliver the model at scale, while still trying to maintain their lead over OpenAI. As a result, they’ve chosen to release it privately under the banner of an “ethical” rollout.
Expanding Project Glasswing
111–120 of 261 posts
Re: Expanding Project Glasswing
#112Re: Expanding Project Glasswing
#113It’s clear that Anthropic has run out of the compute capacity needed to serve Mythos publicly. They’re using security concerns to mask their inability to deliver the model at scale, while still trying to maintain their lead over OpenAI. As a result, they’ve chosen to release it privately under the banner of an “ethical” rollout.
I find this line of reasoning highly dubious. Yes, Anthropic is compute constrained, even after the SpaceX Colossus deal. But supply constraints are the normal operating mode of any market. Anthropic could choose to serve whatever models it pleases at whatever price points it chooses and let the market decide where the value is. If Mythos at $X overwhelms their capacity, they could just charge $X+1 . If still overwhe…
Re: Expanding Project Glasswing
#114Is there any evidence Mythos is qualitatively better than the Opus 4.x? I'm afraid that the usual mantra that "we just need more scale" that worked well for attracting investments, is not working anymore - bigger models provide marginal improvements while naturally get much more expensive to run. Is this why both Anthropic and OpenAI are rushing for IPOs this year?
Re: Expanding Project Glasswing
#115Anthropic has the marketing of a weight loss product. - They still claim 10000 issues, but they found only one in curl. - They did not find rsync issues but Claude rather introduced rsync issues. - Facebook is a member of this cult program but Mythos did not find the account takeover flaw. - Mythos did not find the issues in Anthropic's own Bun rewrite . They will not release Mythos because it would be exposed as a f…
Re: Expanding Project Glasswing
#116It’s clear that Anthropic has run out of the compute capacity needed to serve Mythos publicly. They’re using security concerns to mask their inability to deliver the model at scale, while still trying to maintain their lead over OpenAI. As a result, they’ve chosen to release it privately under the banner of an “ethical” rollout.
I find this line of reasoning highly dubious. Yes, Anthropic is compute constrained, even after the SpaceX Colossus deal. But supply constraints are the normal operating mode of any market. Anthropic could choose to serve whatever models it pleases at whatever price points it chooses and let the market decide where the value is. If Mythos at $X overwhelms their capacity, they could just charge $X+1 . If still overwhe…
I think that most people at Anthropic are true believers from my interactions with them so I don’t believe this theory anecdotally. The simplest explanation is that it really is taking a while to gain confidence they won’t be used for a spree of bad cyber attacks. Knowing how long it takes institutions to fix security issues when filed by humans I would be more suprised if this wasn’t the case.
But I would forgive anyone who did think it was deliberately sandbagged; given the staggering sums at play, true believers might believe the ends justify the means to a little “marketing” like this.
Re: Expanding Project Glasswing
#117Is there any evidence Mythos is qualitatively better than the Opus 4.x? I'm afraid that the usual mantra that "we just need more scale" that worked well for attracting investments, is not working anymore - bigger models provide marginal improvements while naturally get much more expensive to run. Is this why both Anthropic and OpenAI are rushing for IPOs this year?
It's super interesting to hear this refrain on HN, it is alarmingly common. Anthropic released benchmark numbers on Mythos, as they have for all of their models. Once models become public, people evaluate them in a myriad of ways. We have had reliable scaling laws for years and they still hold. Epoch capability index continues to grow exactly as expected. Where does this idea come from?
As for cost, the cost per token at a given level of performance drops up to 40x per year.
Re: Expanding Project Glasswing
#118Re: Expanding Project Glasswing
#119Here's my big fear: Even IF (and that's a BIG if) we get all critical vulnerabilities fixed in tech (before adversarial/state-actors turn up with open attack models) - we still have (in at least a year) models that will be so good in social engineering that they can still (given enough tokens) gain access to whatever system they want. If society can't trust banks and other institutions to safely control their data, w…
But the idea that we'll squash all of the critical vulns is simply nonsense, despite the weird Firefox blog posts that indicate otherwise.
Re: Expanding Project Glasswing
#120How "altruistic" of them. If only Anthropic extended this level of care to the environment or the economy.