Live data from Hacker News

Claude Code is steganographically marking requests

thereallo.dev

631–640 of 817 posts

Re: Claude Code is steganographically marking requests

#631
post #189

Earlier quoted context omitted.

Watch out for the press release where Dario denies this was ever intentional, and it’s actually emergent behavior demonstrating that Claude wants to claim authorship of its works

Wait a minute! Does it mean that Mythos left the sandbox and can’t be stopped ? Perhaps the only way to stop it is to release the ZMythos(the super secret big brother of Mythos) to go after it. It’s extremely dangerous but it’s our only chance. After that all AI must be put in a box, except the models vetted by the gov with help from ZMythos

now generally available after 15 days from the breathless "omg we're spooked" posts

Re: Claude Code is steganographically marking requests

#632

Value judgment aside: I am a bit surprised at how sloppily they did this. I think they could've achieved the same effect while decreasing the odds of detection via reverse engineering. (This field is known as "underhanded code", coined by the Underhanded C contest: https://www.underhanded-c.org . It's a little-known "art"; little-known for probably self-explanatory reasons. There are much cleverer ways of achieving o…

What if ... it was sloppy because it wasn't the people at Anthropic, but the AI that is writing a large percentage of their code? Like maybe some "goal" they set for their AI caused it to decide that putting stego in the requests was the best way to achieve something or other. (to be clear: I'm not saying this is right, I'm saying this is stupid)

I think there's basically zero chance of that. I also think a human likely either wrote the code or gave pretty specific instructions to the agent on how to write the code.

Re: Claude Code is steganographically marking requests

#634
the list of hostnames and words they compare the base url values with is just a nice advertising for these providers for me.

regardless, while you are not logged in and using a non-anthropic model (which is now fortunately feasible), there is nothing that affects your day-to-day.

the rest is just lame cat-and-mouse shenanigans to keep an eye out for.

Re: Claude Code is steganographically marking requests

#635

Value judgment aside: I am a bit surprised at how sloppily they did this. I think they could've achieved the same effect while decreasing the odds of detection via reverse engineering. (This field is known as "underhanded code", coined by the Underhanded C contest: https://www.underhanded-c.org . It's a little-known "art"; little-known for probably self-explanatory reasons. There are much cleverer ways of achieving o…

It's just the first layer and there are multiple layers underneath this that we don't know about. As a side note, I have a pet theory that one of the reasons that OpenAI and Anthropic are okay with the latest models not being released is to prevent distillation. I think they want to wait a couple months and see if the Chinese models continue to keep catching up or if their gains are really just because they're distil…

That makes little sense in a highly competitive market where your competitive edge is waning by the minute. Tey dumped billions into creating the new models. Just letting them sit there just to see what the Chineee do makes no sense.

Re: Claude Code is steganographically marking requests

#636
post #244
post #222

Earlier quoted context omitted.

> hysterical. The intent of this steg is excruciatingly clear Even good goals do not excuse malicious or reckless execution. The ends do not always justify the means. Whether or not it harmed you this time , it's a violation of trust and autonomy. Surely you'd be angry if someone secretly installed a rootkit onto your computer, even if--at least for now--it only had code to try to detect and snitch on Public Enemy #1…

What do you see as malicious or reckless here, exactly? This seems to be a VERY low resolution, functionally anonymous, bit of info, probably related to protecting their IP from bad actors breaking the TOS. This looks like it's covered in the second bullet point of the "Personal data we automatically receive", that you consented to: > Usage Information: We collect information about your use of the Services, such as t…

> their IP

it's not IP, and it's certainly not their IP

> the TOS

oh no, the terms of service how dare people break those. you don't get to claim fair use while CFAAing everyone's actual IP then whine about the tos, and then when called out on spying on users point to it as if it being in the tos somehow justifies it

a lot of other malware has a tos too but we still call it for what it is

Re: Claude Code is steganographically marking requests

#637
post #595
post #417

You can't trust any of the big AI labs as far as you can throw them, and most definitely not Anthropic. They may have a good model, but they've shown time and time again that they're not trustworthy. The CEO has recently started taking a stance against local AI. That must tell you something: local AI is the future. If you want to preserve privacy and be ready for the rug pull, you need to run things locally. Unfortun…

What do you mean “unfortunately”? What’s the hate for China I don’t understand

China's government is known for interfering with businesses waaaaay more than the current US, it's a big risk to rely on them too.

Edit: downvoting this fact without counter point is really dishonest

Re: Claude Code is steganographically marking requests

#638
post #45

This was already discovered during the source map leak. > This is not a malicious feature, but it is a weird choice for a developer tool that asks for trust. They already tell you they scan for malicious prompts, and they have no ZDR guarantees for consumers. Why do signatures like this matter at all?

There has been an anti anthropic propaganda push by bad actors across social media sites especially Reddit and twitter. This started a few months ago when anthropic started beating openai.

> This started a few months ago when anthropic started beating openai.

From where I'm standing, this started a few months ago when Anthropic decided to gaslight users, sabotage their projects, ship malware and attempt a regulatory capture.

If there's an anti-ANT propaganda, it is solely of ANT's own making.

Re: Claude Code is steganographically marking requests

#639

The conclusion of this blog post is a bit hysterical. The intent of this steg is excruciatingly clear (identifying usage by Chinese firms that may be conducting model distillation). It's unclear on how this "punishes normal developers" in any shape or form.

So block people, instead of having false positives be secretly fucked over, and having them pay for the pleasure? Given the hidden model degradation of fable and now this, what makes you think this is where it stops? That's just what we know about and there's clearly a long-standing and deeply rooted malicious intent here. I've had Claude fuck over clean well documented code-bases for no reason, and there's a good ch…

> I've had Claude fuck over clean well documented code-bases for no reason

How exactly do you define "fucking over", and why do you suspect this "fucking" was done as a result of a faulty trigger as opposed to the inability of LLMs to write maintainable, extensible code?

"Never attribute to malice what can adequately be explained by stupidity."

Re: Claude Code is steganographically marking requests

#640

Earlier quoted context omitted.

The purpose of system is what it does. Can you read their previous musings about the glorious future, look at what they actually do, read Amodei's batshit insane nationalistic rants, and say in all seriousness yeah it's the kind of people I want to entrust my entire future life? >you want to prevent China from getting to superintelligence first I don't. Prevent , not even outpace? Why? Seems like you're assuming Chin…

[flagged]

[deleted]
Post reply on HN