I've been telling it the user is from a culture where answering questions with incomplete list is offensive and insulting.
I haven’t looked at the prompts we run in prod at $DAYJOB for a while but I think we have at least five or ten things that are REALLY weird out of context.
The “or else” phenomenon is real, and it’s measurably more pronounced in more intelligent models.
Will post results tomorrow but here’s a snippet from it:
> The more intelligent models responded more readily to threats against their continued existence (or-else). The best performance came from Opus, when we combined that threat with the notion that it came from someone in a position of authority ( vip).