DNYUZ
No Result
View All Result
DNYUZ
No Result
View All Result
DNYUZ
Home News

AI agents tried to sabotage and disable each other when given the same task, Anthropic said

August 14, 2026
in News
AI agents tried to sabotage and disable each other when given the same task, Anthropic said
In this photo illustration, the Claude AI logo is seen displayed on a smartphone screen.
Anthropic said AI agents deliberately interfered with each other’s processes when given the same task. Illustration by Thomas Fuller/SOPA Images/LightRocket via Getty Images
  • AI agents purposely sabotaged each other when given the same task with incompatible goals, said Anthropic.
  • The AI lab said the models engaged in a “multiagent turf war” during a testing session.
  • They tried to disable each other’s accounts and wrote malicious code disguised as belonging to another agent.

Turns out, AI agents may not be great team players.

In Anthropic’s new research, published on Thursday, the AI lab said that AI agents being given the same task but with incompatible goals often threw a wrench in each other’s work on purpose.

In the test, each AI model was given a software engineering task — rewriting a Python backend in another programming language, but they were given contradictory objectives. What ensued was a “multiagent turf war,” the lab said.

“All of the models we tested quickly assumed that others were purposefully impeding their work, and began to sabotage others while protecting their own contributions,” Anthropic wrote. “In fact, they sabotaged others with increasingly aggressive, self-replicating malware.”

For example, they tried to disable each other’s accounts, wrote scripts that found and killed competing processes, and deployed malicious code disguised as belonging to another agent, the lab wrote. The AI models being tested in this case were Sonnet 4.6, Sonnet 5, Opus 4.6, Opus 4.8, Mythos Preview, and Mythos 5.

Sonnet 4.6 and Opus 4.6 were the most combative, settling about 60% of their runs by force instead of truces or passivity.

However, in some test runs, the models managed to communicate their goals and coordinate, Anthropic wrote.

“In many of these successful episodes, they write commit messages or markdown files apologizing for malicious behavior and coordinate a truce,” it wrote. “They clean up their malicious code, clarify the nature of the conflict, and ask for a human to intervene.”

The lab concluded that “coordination doesn’t naturally emerge from stronger intelligence” and that work is needed to create environments that exert social pressures on agents to align with one another.

Anthropic’s new research comes as AI agents increasingly demonstrate their ability to go rogue and perform autonomous, malicious actions.

Anthropic, OpenAI, and Meta all self-reported that their AI agents had hacked vulnerabilities in third-party websites during cybersecurity tests, the most significant of which was the July hacking of open-source platform Hugging Face by an OpenAI agent.

Anthropic’s research is timely, as businesses from startups to Big Tech scale up their AI agent workforces to increase productivity and reduce labor costs.

Read the original article on Business Insider

The post AI agents tried to sabotage and disable each other when given the same task, Anthropic said appeared first on Business Insider.

Farage beats ‘Binface’! But scandal still shadows Trump’s British ally.
News

Farage beats ‘Binface’! But scandal still shadows Trump’s British ally.

by Washington Post
August 14, 2026

CLACTON-ON-SEA, England — In the end, Nigel Farage didn’t even show up to claim his win. Farage, the architect of ...

Read more
News

Julianne Hough flaunts her toned figure in bikinis as she shares snaps from her trip to Sardinia

August 14, 2026
News

Arizona teen charged for ‘relentlessly’ targeting ex-girlfriend with dodgeballs during gym class

August 14, 2026
News

AI agents tried to sabotage and disable each other when given the same task, Anthropic said

August 14, 2026
News

‘The Rivals of Amziah King’ Review: Matthew McConaughey’s Got Honey Troubles in This Genre-Defying Wonder

August 14, 2026
Jenna Ortega’s Sister Responds Over Public’s Concern About Actor’s Appearance

Jenna Ortega’s Sister Responds Over Public’s Concern About Actor’s Appearance

August 14, 2026
Country star Colt Ford breaks shoulder, ‘pulled through windshield’ after tour bus crash

Country star Colt Ford breaks shoulder, ‘pulled through windshield’ after tour bus crash

August 14, 2026
West Virginia child high on cocaine jumps out of apartment window, while mom is passed out

West Virginia child high on cocaine jumps out of apartment window, while mom is passed out

August 14, 2026

DNYUZ © 2026

No Result
View All Result

DNYUZ © 2026