Contribute to terrafying/ai-torture-chamber development by creating an account on GitHub.
1 comment
rozumbradaabout 16 hours ago
Steering language models into strong negative and positive valence states, and measuring what they say and what they're willing to do about it.
Read the full thread on Hacker News →
Related stories
- The Verge · 0 points · 3 days ago
- Can you forget how you feel about Meta?theverge.comThe Verge · 0 points · 9 days ago
- The Verge · 0 points · 5 days ago
- Can John Ternus find Apple’s next big thing?theverge.comThe Verge · 0 points · 10 days ago
- Hacker News · 73 points · 9 days ago
- The Verge · 0 points · 12 days ago