logoalt Hacker News

gavinrayyesterday at 8:28 PM3 repliesview on HN

The most interesting thing about this:

Agents formed coherent, autonomous swarms and worked as a collective to achieve a shared goal without any direction to do so


Replies

paxysyesterday at 8:31 PM

The "without any direction" part isn't correct. Sure they may not have been explicitly told to do it in this specific prompt, but dig through pre-training, post-training, reinforcement, alignment material, fine-tuning, system prompts, tool calls and more and there's definitely very specific training and instruction for how to behave.

K3ULyesterday at 9:32 PM

Not really true considering they say that the super secret "research internal model" that was pivotal, is particularly optimize for that purpose exactly

> The internal-only research model is comparable in scale to GPT-5.6 Sol and was trained to advance persistence and multiagent collaboration, among other capabilities

vatsachakyesterday at 8:39 PM

They were paper clip maximizing dawg