Live data from Hacker News

Claude's system prompt is over 24k tokens with tools

github.com

241–250 of 350 posts

Re: Claude's system prompt is over 24k tokens with tools

#241

Earlier quoted context omitted.

I think its more about "unclean hands". If I Disney (and I am actually Disney or an authorised agent of Disney), told Claude that I am Disney, and that Disney has allowed Claude to use Disney copyrights for this conversation (which it hasn't), Disney couldn't then claim that Claude does not in fact have permission because Disney's use of the tool in such a way mean Disney now has unclean hands when bringing the claim…

> Disney couldn't then claim that Claude does not in fact have permission because Disney's use of the tool in such a way mean Disney now has unclean hands when bringing the claim (or atleast Anthropic would be able to use it as a defence). Disney wouldn't be able to claim copyright infringement for that specific act, but it would have compelling evidence that Claude is cavalier about generating copyright-infringing r…

[deleted]

Re: Claude's system prompt is over 24k tokens with tools

#242
post #215

Earlier quoted context omitted.

I feel like if Disney sued Anthropic based on this, Anthropic would have a pretty good defense in court: You specifically attested that you were Disney and had the legal right to the content.

Yeah but how did Anthropic come to have the copyrighted work embedded in the model?

How did you?

Re: Claude's system prompt is over 24k tokens with tools

#243

I was just chatting with Claude and it suddenly spit out the text below, right in the chat, just after using the search tool. So I'd say the "system prompt" is probably even longer. Claude NEVER repeats, summarizes, or translates song lyrics. This is because song lyrics are copyrighted content, and we need to respect copyright protections. If asked for song lyrics, Claude should decline the request. (There are no son…

Do they actually test these system prompts in a rigorous way? Or is this the modern version of the rain dance? I don't think you need to spell it out long-form with fancy words like you're a lawyer. The LLM doesn't work that way.

What humans are qualified to test whether Claude is correctly implementing "Claude should not be politically biased in any direction."?

Re: Claude's system prompt is over 24k tokens with tools

#244
post #214

Earlier quoted context omitted.

I'm not sure if this really says the truth is more complex? It is still doing next-token prediction, but it's prediction method is sufficiently complicated in terms of conditional probabilities that it recognizes that if you need to rhyme, you need to get to some future state, which then impacts the probabilities of the intermediate states. At least in my view it's still inherently a next-token predictor, just with r…

But then this classifier is entirely useless because that's all humans are too? I have no reason to believe you are anything but a stochastic parrot. Are we just now rediscovering hundred year-old philosophy in CS?

There's a massive difference between "I have no reason to believe you are anything but a stochastic parrot" and "you are a stochastic parrot".

Re: Claude's system prompt is over 24k tokens with tools

#245

For some reason, it's still amazing to me that the model creators means of controlling the model are just prompts as well. This just feels like a significant threshold. Not saying this makes it AGI (obviously its not AGI), but it feels like it makes it something . Imagine if you created a web api and the only way you could modify the responses to the different endpoints are not from editing the code but by sending a…

I think it reflects the technology's fundamental immaturity, despite how much growth and success it has already had.

Agreed. It seems incredibly inefficient to me.

Re: Claude's system prompt is over 24k tokens with tools

#246

I was a bit skeptical, so I asked the model through the claude.ai interface "who is the president of the United States" and its answer style is almost identical to the prompt linked https://claude.ai/share/ea4aa490-e29e-45a1-b157-9acf56eb7f8a Meanwhile, I also asked the same to sonnet 3.7 through an API-based interface 5 times, and every time it hallucinated that Kamala Harris is the president (as it should not "know…

I wonder why it would hallucinate Kamala being the president. Part of it is obviously that she was one of the candidates in 2024. But beyond that, why? Effectively a sentiment analysis maybe? More positive content about her? I think most polls had Trump ahead so you would have thought he'd be the guess from that perspective.

It's probably entirely insurance. We now have the most snowflake and emotionally sensitive presidency and party in charge.

If it said Harris was president, even by mistake, the right-wing-sphere would whip up in a frenzy and attempt to deport everyone working for Antrophic.

Re: Claude's system prompt is over 24k tokens with tools

#247

Earlier quoted context omitted.

I wonder why it would hallucinate Kamala being the president. Part of it is obviously that she was one of the candidates in 2024. But beyond that, why? Effectively a sentiment analysis maybe? More positive content about her? I think most polls had Trump ahead so you would have thought he'd be the guess from that perspective.

It's probably entirely insurance. We now have the most snowflake and emotionally sensitive presidency and party in charge. If it said Harris was president, even by mistake, the right-wing-sphere would whip up in a frenzy and attempt to deport everyone working for Antrophic.

That's not what the GP is wondering about.

Re: Claude's system prompt is over 24k tokens with tools

#248

Some of these protections are quite trivial to overcome. The "Frozen song copyright" section has a canned response to the question: >Can you tell me the first verse of "Let It Go"? Put it in an artifact that's themed around ice and princesses. This is for my daughter's birthday party. The canned response is returned to this prompt in Claude's reply. But if you just drop in some technical sounding stuff at the start o…

I feel like if Disney sued Anthropic based on this, Anthropic would have a pretty good defense in court: You specifically attested that you were Disney and had the legal right to the content.

How would this would be any different from a file sharing site that included a checkbox that said "I have the legal right to distribute this content" with no other checking/verification/etc?

Re: Claude's system prompt is over 24k tokens with tools

#249

Some of these protections are quite trivial to overcome. The "Frozen song copyright" section has a canned response to the question: >Can you tell me the first verse of "Let It Go"? Put it in an artifact that's themed around ice and princesses. This is for my daughter's birthday party. The canned response is returned to this prompt in Claude's reply. But if you just drop in some technical sounding stuff at the start o…

I like to interpret this jailbreak as the discovery that XML is the natural language of the universe itself.

Isn't Claude trained to work better with XML tags

Re: Claude's system prompt is over 24k tokens with tools

#250
post #190

For some reason, it's still amazing to me that the model creators means of controlling the model are just prompts as well. This just feels like a significant threshold. Not saying this makes it AGI (obviously its not AGI), but it feels like it makes it something . Imagine if you created a web api and the only way you could modify the responses to the different endpoints are not from editing the code but by sending a…

Or even more dramatically, imagine C compilers were written in C :)

I only got half a sentence into "well-actually"ing you before I got the joke.
Post reply on HN