In short, here, I described this not as collusion
other similar systems
Something of a side note to your excellent write-up... I think we should push back on the "sub-agent" "agent swarm" language used to describe LLM workflows.
The workflows often described as swarms or whatever are the same LLM (or at least related LLMs), just prompted a bunch of different ways in parallel. Individual agents don't really act like coherent agents in the lesswrong rationalist sense, or in any sort of philosophical sense of selfhood or agency or unified purposes, they just get described that way for convenience, so treating "swarms" of agents and subagents as some special category is lending too much credence to the anthropomorphisizing boosters like to do. "Agent swarms" really just means that someone set up a slop machine to prompt itself a bunch of times in parallel without enough human supervision.