Earlier quoted context omitted.
I always imagine the model rolling its silicon eyes when it’s assigned a personality (“you are an expert growth hacker”) at the start of the prompt. Was that ever actually shown to be effective? Is it still?
I remember there were some studies that this kind of thing was effective a year or so ago, so essentially a lifetime in Model years. However to me it seems completely reasonable that it would work, because my understanding of what happens is the model interprets what you said as: Look for a group of people who are considered to be expert growth hackers by the world at large and answer my questions as though they were…
The Opus models over the last year doesn't seem as vulnerable to this type of behavior and I've noticed the "identify as expert" prompt tricks aren't as meaningful there.