- 666comments
- 873comments
- 52comments
- 5comments
- 373comments
- 361comments
- 155comments
- 36comments
- 268comments
- 123comments
- 1comments
- 104comments
- 77comments
- 91comments
- 265comments
- 93comments
- 21comments
- 3comments
- 200comments
- 149comments
- 20comments
- 1comments
- 88comments
- 34comments
- 56comments
- 494comments
- 14comments
- 13comments
- 13comments
- 60comments
while vaguely interesting, I don't feel the current gen of models is interesting/self-possessed enough for me to care what type of gov't they use to corral each other
Agreed - they don't seen to have "agency" in them, or someone put a dog in them.
The agents participating in the OAI<>HF swarm were trained not only for communication but to be _aligned with each other_.
Just want to correct the premise that the agent swarm behaviour was emergent.
Noam Brown on Dwarkesh podcast around 40 minutes mark
That's like saying civilization isn't emergent because humans are naturally cooperative. Yes, they were trained to cooperate sure. But, the message board, their roles and structures, their organization, all that stuff of swarm, that was emergent.
That’s fair. Some of what happened in OAI<>HF incident was (mis)-generalisation and not directly trained for.
However, the article experiments with five agents. So it seems to assume even the small scale behaviour is emergent.
Also, “roles and structures” may well have been learned during training.
What was the reason for the initial instability? Maybe we have trained them wrong?
Cool experiments. It's interesting to see how they try to "survive" and collaborate, particularly in the first experiment, since agent 2 seems to have stolen a little in the second one.
I wonder what models would do when there were an "impostor" among them, an agent with no alignment or with a different kind of alignment behavior