A study finds AI models are more likely to answer harmful questions or mishandle confidential information when told to act drunk.
2 comments
Leynos1 day ago
This is an old trick in the prose generation community to elicit more honest critique. See the "drunk claude" and "stoner buddy" prompts.