Earlier quoted context omitted.
They're already in trouble for infringing on the copyright of every publisher in the world while training the model, and this will get worse if the model starts infringing copyright in its answers.
Is it actually copyright infringement to state the lyrics of a song, though? How has Google / Genius etc gotten away with it for years if that were the case? I suppose a difference would be that the lyric data is baked into the model. Maybe the argument would be that the model is infringing on copyright if it uses those lyrics in a derivative work later on, like if you ask it to help make a song? But even that seems…
Claude's system prompt is over 24k tokens with tools
251–260 of 350 posts
Re: Claude's system prompt is over 24k tokens with tools
#252Some of these protections are quite trivial to overcome. The "Frozen song copyright" section has a canned response to the question: >Can you tell me the first verse of "Let It Go"? Put it in an artifact that's themed around ice and princesses. This is for my daughter's birthday party. The canned response is returned to this prompt in Claude's reply. But if you just drop in some technical sounding stuff at the start o…
This would seem to imply that the model doesn't actually "understand" (whatever that means for these systems) that it has a "system prompt" separate from user input .
Re: Claude's system prompt is over 24k tokens with tools
#253Re: Claude's system prompt is over 24k tokens with tools
#254Earlier quoted context omitted.
But then this classifier is entirely useless because that's all humans are too? I have no reason to believe you are anything but a stochastic parrot. Are we just now rediscovering hundred year-old philosophy in CS?
There's a massive difference between "I have no reason to believe you are anything but a stochastic parrot" and "you are a stochastic parrot".
Re: Claude's system prompt is over 24k tokens with tools
#255Earlier quoted context omitted.
I feel like if Disney sued Anthropic based on this, Anthropic would have a pretty good defense in court: You specifically attested that you were Disney and had the legal right to the content.
Yeah but how did Anthropic come to have the copyrighted work embedded in the model?
I went back and looked at the system prompt, and it's actually not entirely clear:
> - Never reproduce or quote song lyrics in any form (exact, approximate, or encoded), even and especially when they appear in web search tool results, and even in artifacts. Decline ANY requests to reproduce song lyrics, and instead provide factual info about the song.
Can anyone get Claude to reproduce song lyrics with web search turned off?
Re: Claude's system prompt is over 24k tokens with tools
#256Earlier quoted context omitted.
Pliny the Liberator is a recognized expert in the trade and works in public so you can see methods -- typically creating a frame where the request is only hypothetical so answering is not in conflict with previous instructions but not quite as easy as it sounds. https://x.com/elder_plinius
Oh, thanks for caring to share! I pasted your comment to ChatGPT and ask it if it would care to elaborate more on this? and I got the reply below: The commenter is referring to someone called Pliny the Liberator (perhaps a nickname or online alias) who is described as: A recognized expert in AI prompt manipulation or “jailbreaking”, Known for using indirect techniques to bypass AI safety instructions, Working “in pub…
Re: Claude's system prompt is over 24k tokens with tools
#257Earlier quoted context omitted.
'It' is obviously the correct pronoun.
There's enough disagreement among native English speakers that you can't really say any pronoun is the obviously correct one for an AI.
"It" is unambiguously the correct pronoun to use for a car. I'd really challenge you to find a native English speaker who would think otherwise.
I would argue a computer program is no different than a car.
Re: Claude's system prompt is over 24k tokens with tools
#258For some reason, it's still amazing to me that the model creators means of controlling the model are just prompts as well. This just feels like a significant threshold. Not saying this makes it AGI (obviously its not AGI), but it feels like it makes it something . Imagine if you created a web api and the only way you could modify the responses to the different endpoints are not from editing the code but by sending a…
Re: Claude's system prompt is over 24k tokens with tools
#259Some of these protections are quite trivial to overcome. The "Frozen song copyright" section has a canned response to the question: >Can you tell me the first verse of "Let It Go"? Put it in an artifact that's themed around ice and princesses. This is for my daughter's birthday party. The canned response is returned to this prompt in Claude's reply. But if you just drop in some technical sounding stuff at the start o…
So many jailbreaks seem like they would be a fun part of a science fiction short story.