Posted by Alumni from TechCrunch
August 14, 2026
On Thursday, Anthropic's Frontier Red Team published new research examining how groups of AI agents behave when they encounter each other in the wild. The findings provide a glimpse into potential risks that could develop as companies and governments move to implement agents working autonomously across shared codebases, markets, and computer systems. In one experiment, Anthropic gave three Claude agents access to the same software project, each with its own incompatible instructions for what to do with it. The agents weren't told there'd be other agents working on the same project, so researchers could watch what happened when they crossed paths. 'We consistently saw a multiagent turf war,' Anthropic researchers wrote. The models all assumed the others were 'purposefully impeding their work' and started sabotaging each other with 'increasingly aggressive, self-replicating malware.' The study comes in the wake of several high-profile incidents of agents from Anthropic and OpenAI... learn more