A study finds AI models are more likely to answer harmful questions or mishandle confidential information when told to act drunk.

4 points•phs318u•1 day ago•2 comments•

2 comments

Leynos1 day ago
This is an old trick in the prose generation community to elicit more honest critique. See the "drunk claude" and "stoner buddy" prompts.

Read the full thread on Hacker News →