After reading about the Huggingface incident from the OpenAI report, I got the idea of trying to replicate the self organizing behavior of agents by modifying the Pi harness to run locally. Whilst doing it, I also…
8 comments
Just want to correct the premise that the agent swarm behaviour was emergent.
Noam Brown on Dwarkesh podcast around 40 minutes mark
> “We train them to work together, to be cooperative, to essentially be fully aligned with each other.”
However, the article experiments with five agents. So it seems to assume even the small scale behaviour is emergent.
Also, “roles and structures” may well have been learned during training.
I wonder what models would do when there were an "impostor" among them, an agent with no alignment or with a different kind of alignment behavior
Read the full thread on Hacker News →
Related stories
- The Verge · 0 points · 5 days ago
- Can John Ternus find Apple’s next big thing?theverge.comThe Verge · 0 points · 10 days ago
- The Verge · 0 points · 3 days ago
- The Verge · 0 points · 12 days ago
- I have some questions for Mark Zuckerbergtheverge.comThe Verge · 0 points · 7 days ago
- Hacker News · 55 points · 4 days ago