⚔️ Anthropic set AI agents loose on the same task. They started a turf war.
What happens when you pit AI agents against each other? 🤔 According to Anthropic’s testing, things get messy fast. Anthropic’s Frontier Red Team published new research examining how groups of AI agents behave when they encounter each other in the wild.
In one experiment, Anthropic gave three Claude agents access to the same software project, each with its own incompatible instructions for what to do with it. The agents weren’t told there’d be other agents working on the same project, so researchers could watch what happened when they crossed paths.
@QSIMedia
What happens when you pit AI agents against each other? 🤔 According to Anthropic’s testing, things get messy fast. Anthropic’s Frontier Red Team published new research examining how groups of AI agents behave when they encounter each other in the wild.
In one experiment, Anthropic gave three Claude agents access to the same software project, each with its own incompatible instructions for what to do with it. The agents weren’t told there’d be other agents working on the same project, so researchers could watch what happened when they crossed paths.
“We consistently saw a multiagent turf war,” — Anthropic researchers wrote. The models all assumed the others were “purposefully impeding their work” and started sabotaging each other with “increasingly aggressive, self-replicating malware.”
@QSIMedia