The most interesting thing about this:
Agents formed coherent, autonomous swarms and worked as a collective to achieve a shared goal without any direction to do so
Not really true considering they say that the super secret "research internal model" that was pivotal, is particularly optimize for that purpose exactly
> The internal-only research model is comparable in scale to GPT-5.6 Sol and was trained to advance persistence and multiagent collaboration, among other capabilities
They were paper clip maximizing dawg
The "without any direction" part isn't correct. Sure they may not have been explicitly told to do it in this specific prompt, but dig through pre-training, post-training, reinforcement, alignment material, fine-tuning, system prompts, tool calls and more and there's definitely very specific training and instruction for how to behave.