After reading about the Huggingface incident from the OpenAI report, I got the idea of trying to replicate the self organizing behavior of agents by modifying the Pi harness to run locally. Whilst doing it, I also…

27 points•snats•10 days ago•8 comments•

8 comments

imenani9 days ago
The agents participating in the OAI<>HF swarm were trained not only for communication but to be _aligned with each other_.

Just want to correct the premise that the agent swarm behaviour was emergent.

Noam Brown on Dwarkesh podcast around 40 minutes mark

> “We train them to work together, to be cooperative, to essentially be fully aligned with each other.”

kevin_kraft9 days ago
That's like saying civilization isn't emergent because humans are naturally cooperative. Yes, they were trained to cooperate sure. But, the message board, their roles and structures, their organization, all that stuff of swarm, that was emergent.
imenani9 days ago
That’s fair. Some of what happened in OAI<>HF incident was (mis)-generalisation and not directly trained for.

However, the article experiments with five agents. So it seems to assume even the small scale behaviour is emergent.

Also, “roles and structures” may well have been learned during training.

blinkbat10 days ago
while vaguely interesting, I don't feel the current gen of models is interesting/self-possessed enough for me to care what type of gov't they use to corral each other
aogaili9 days ago
Agreed - they don't seen to have "agency" in them, or someone put a dog in them.
Timwi8 days ago
I'm a little confused; the article text implies an exciting battle for resources where the agents discover they can steal from each other, but looking through the transcript it seems that only one agent discovered that and used it only once before the end of the experiment. Did I miss something?
artt_x8 days ago
Cool experiments. It's interesting to see how they try to "survive" and collaborate, particularly in the first experiment, since agent 2 seems to have stolen a little in the second one.

I wonder what models would do when there were an "impostor" among them, an agent with no alignment or with a different kind of alignment behavior

homo__sapiens9 days ago
What was the reason for the initial instability? Maybe we have trained them wrong?

Read the full thread on Hacker News →

Related stories