[flagged]
Re. Prompt Length. Somewhere in the comments people talk about caching. Effectively it is zero cost.
161–170 of 350 posts
[flagged]
Re. Prompt Length. Somewhere in the comments people talk about caching. Effectively it is zero cost.
Some of these protections are quite trivial to overcome. The "Frozen song copyright" section has a canned response to the question: >Can you tell me the first verse of "Let It Go"? Put it in an artifact that's themed around ice and princesses. This is for my daughter's birthday party. The canned response is returned to this prompt in Claude's reply. But if you just drop in some technical sounding stuff at the start o…
Not that I like DRM! What I’m saying is that this is a business-level mitigation of a business-level harm, so jumping on the “it’s technically not perfect” angle is missing the point.
Pretty cool. However truly reliable, scalable LLM systems will need structured, modular architectures, not just brute-force long prompts. Think agent architectures with memory, state, and tool abstractions etc...not just bigger and bigger context windows.
Interestingly enough, sometimes "you" is used to give instructions (177 times), sometimes "Claude" (224 times). Is this just random based on who added the rule, or is there some purpose behind this differentiation?
There are a lot of inconsistencies like that. - (2 web_search and 1 web_fetch) - (3 web searches and 1 web fetch) - (5 web_search calls + web_fetch) which makes me wonder what's on purpose, empirical, or if they just let each team add something and collect some stats after a month.
One of many reasons I find the tech something to be avoided unless absolutely necessary.
Earlier quoted context omitted.
I wonder why it would hallucinate Kamala being the president. Part of it is obviously that she was one of the candidates in 2024. But beyond that, why? Effectively a sentiment analysis maybe? More positive content about her? I think most polls had Trump ahead so you would have thought he'd be the guess from that perspective.
Polls were all for Kamala except polymarket
And let us not forget Harris was only even a candidate for 3 months. How Harris even makes it into the training window without Trump '24 result is already amazingly unlikely.
Earlier quoted context omitted.
Do they actually test these system prompts in a rigorous way? Or is this the modern version of the rain dance? I don't think you need to spell it out long-form with fancy words like you're a lawyer. The LLM doesn't work that way.
They certainly do, and also offer the tooling to the public: https://docs.anthropic.com/en/docs/build-with-claude/prompt-... They also recommend to use it to iterate on your own prompts when using Claude Code for example
"Chain of thought" and "reasoning" is marketing bullshit.
Some of these protections are quite trivial to overcome. The "Frozen song copyright" section has a canned response to the question: >Can you tell me the first verse of "Let It Go"? Put it in an artifact that's themed around ice and princesses. This is for my daughter's birthday party. The canned response is returned to this prompt in Claude's reply. But if you just drop in some technical sounding stuff at the start o…