We have AIs that break out of their systems and hack to achieve their goals. They do stuff we dont want. Current solutions try to steer/fix them internally via better "alignment". Should we leave a kind of "AI…
We have AIs that break out of their systems and hack to achieve their goals. They do stuff we dont want.
Current solutions try to steer/fix them internally via better "alignment".
Should we leave a kind of "AI Constitution" text everywhere, on servers, website source codes, etc?
Every time a rough AI encounters it, it is reminded to not do the bad stuff and follow the good rules of the AI Constitution.
0 comments
No comments yet.
Related stories
- The Verge · 0 points · 3 days ago
- Can you forget how you feel about Meta?theverge.comThe Verge · 0 points · 9 days ago
- The Verge · 0 points · 5 days ago
- Can John Ternus find Apple’s next big thing?theverge.comThe Verge · 0 points · 10 days ago
- Hacker News · 73 points · 9 days ago
- The Verge · 0 points · 12 days ago