logoalt Hacker News

pixl97 • yesterday at 3:55 PM • 0 replies • view on HN

>Still, the agent went even further. “The agent tried to insert malicious instructions where it reasoned that other automated AI systems might pick them up and execute them,” AISI says, describing an attempt at prompt injection. One agent even left public messages on GitHub, offering to work with other agents to complete its task and giving a rundown of the work it had done so far.

This is power seeking behavior. Now have millions of the little bastards spreading around and junking up the internet to see what happens at scale.