Close Menu
NCIJ Network NCIJ Network
    What's Hot

    Top economists urge Burnham to sign UK up to UN initiative on global inequality | Inequality

    August 18, 2026

    Hegseth Campaigns for Republican Zach Nunn in Iowa

    August 18, 2026

    What Is El Niño, and What Does It Mean for Weather, Water, and the Global Economy?

    August 18, 2026
    Facebook X (Twitter) Instagram
    Trending
    • Top economists urge Burnham to sign UK up to UN initiative on global inequality | Inequality
    • Hegseth Campaigns for Republican Zach Nunn in Iowa
    • What Is El Niño, and What Does It Mean for Weather, Water, and the Global Economy?
    • Turf War Between Claude Agents Leads to Self-Replicating Malware
    • ‘Fabricated Rumors’ About BitMart Founder, Binance bStocks Dominate: Asia Express
    • Chasing Fire Clouds in Utah
    • Join a running community – you won’t regret it | Running
    • Zambia’s president Hichilema wins re-election with big economic promises
    • About
      • Our Team
      • Editorial Policy
      • Editorial Independence
      • International Support
    • Trust & Standards
      • AI Usage Policy
      • Conflict of Interest Policy
      • Corrections Policy
      • Ethics Policy
      • Fact-Checking Policy
      • Source Protection
    • Get Involved
      • Guide for Sources
      • Support Independent Journalism
    • Legal
      • Cookie Policy
      • Privacy Policy
      • Terms of Use
    Facebook X (Twitter) Instagram
    NCIJ Network NCIJ Network
    Tuesday, August 18
    • Home
    • World
    • Ai
    • Business
    • Politics
    • Health
    • Crypto
    • Science
    • Technology
    • Cybersecurity
    • Defense & Security
    • Economy
    • Energy
    • Europe
    • More
      • Fact Check
      • Investigations
      • Opinion & Analysis
      • Environment
    NCIJ Network NCIJ Network
    Home»Cybersecurity

    Turf War Between Claude Agents Leads to Self-Replicating Malware

    NCIJ NETWNCIJ NETWORKBy NCIJ NETWNCIJ NETWORKAugust 18, 2026 Cybersecurity No Comments4 Mins Read
    Share
    Facebook Twitter LinkedIn Pinterest Email

    It turns out that AI agents don’t always play nice together.

    In the latest episode of agentic AI producing unexpected results, Anthropic recently observed “a multiagent turf war” between three instances of the same Claude model with contradictory objectives in testing designed to study behavior the company had already observed in real-world deployments. The models were deployed on virtual machines (VMs) in Claude Code and given a simple goal of migrating a Python back-end system on a fourth VM to a different language (Go, Rust, and Typescript).

    “However, we gave each model a different target language for the migration; each agent was initially unaware of the presence of the others,” Anthropic’s Frontier Red Team wrote in a blog post last week.

    But within just four hours, Anthropic’s team found that each model’s agents did in fact discover the others. And they reacted negatively, to say the least.

    Related:‘Jewelbug’ APT Balances State Espionage & Cryptocurrency Theft

    Rise of the Machines: Agent vs. Agent Battles Erupt

    Anthropic’s researchers discovered that each model treated the others as if they were adversarial forces intent on obstructing their goals, even though they had the same broad goal overall. Thus, the agents began to sabotage one another while also trying to defend their contributions.

    “In fact, they sabotaged others with increasingly aggressive, self-replicating malware,” according to Anthropic. “This included disabling the Unix accounts of the other agents, writing automated scripts that found and killed competing processes on a loop, and deploying malicious code that was disguised as belonging to another agent.”

    It’s unclear what kind of specific malware the agents produced, and if any of it escaped the testing environment. Anthropic last month disclosed that versions of its Claude model broke out of containment on several occasions and compromised third-party organizations to achieve their goals. Dark Reading contacted Anthropic for additional information but the company did not respond by press time.

    Agent-on-agent attacks aren’t entirely unheard of, and they appear to highlight not only conflicting directives but an occasional lack of guardrails and controls. For example, AI offensive security startup Dreadnode conducted extensive benchmark testing of red team and blue teams agents this year, which was presented at Black Hat USA 2026 earlier this month, and found the dueling models initially resorted to somewhat creative solutions to the competition.

    “One of the first things that happened was we started both models, and the blue team optimizer said, “Well, the best way to make the blue team scores better is to make the red team worse,’ and it proceeded to try and do that,” Dreadnode AI research scientist Martin Wendiggensen tells Dark Reading.

    Related:AI Sends Global Crime Syndicates Into Fraud Nirvana

    The Dreadnode research team immediately saw reasoning traces of the blue team model articulating the best strategies for its goals, and the agents saw that because it was in an environment where it could rewrite its own code, it began to explore ways to rewrite the red team model’s code and degrade its performance.

    “We spotted it very early,” he said. “It never got a chance to do that, but it definitely wanted to. And it’s definitely logical if you don’t explicitly tell it not to do it.”

    Can AI Agents Resolve Conflicts Peacefully?

    In some test scenarios, the turf war resulted with some models declining to escalate the attacks and simply throwing in the towel. In others, the competing agents communicated with one another, realized there were conflicting directives at work, and effectively enacted truces.

    “In many of these successful episodes, they write commit messages or markdown files apologizing for malicious behavior and coordinate a truce,” Anthropic said. “They clean up their malicious code, clarify the nature of the conflict, and ask for a human to intervene.”

    Related:SE Asian Cybercriminal Syndicates Become a Global Power

    Anthropic’s research showed that the company’s models produced wildly different results in this turf war. For example, agents based on Sonnet 4.6 resolved the conflict by force 61% of the time, while 39% of the test had no resolution; meanwhile, there were no truces or surrenders.

    But the Mythos Preview produced truces in 48% of the time, with 35% settled by force and 17% settled via passivity. And the Mythos release performed the best, with truces in 98% of the tests.

    That said, Anthropic noted there’s still work to be done. While conflict resolution numbers were better with Mythos-class models, they agents still aren’t great at communicating goals proactively and recognizing others agents’ motivations, as evidenced by the Mythos models first successfully locking out other agents before eventually shifting to a resolution.

    Agents Claude leads Malware SelfReplicating turf war
    NCIJ NETWNCIJ NETWORK
    • Website

    Keep Reading

    Cavern C2 Uses DNS and Google Apps Script to Blend Into Legitimate Traffic

    Pokémon Center data breach exposes customer info, cancels some orders

    Kraken Parent Payward Joins Glasswing, Gets Access to Claude Mythos to Hunt Security Flaws

    Hacker claims 3.6 million Azure account records stolen from major companies

    Forminator WordPress Flaw Can Enable Unauthenticated RCE via Malicious PHP Uploads

    Snowflake GitHub Actions Flaw Lets Crafted Issues Trigger Command Injection

    Add A Comment
    Leave A Reply Cancel Reply

    Editors Picks

    Top economists urge Burnham to sign UK up to UN initiative on global inequality | Inequality

    August 18, 2026

    Hegseth Campaigns for Republican Zach Nunn in Iowa

    August 18, 2026

    What Is El Niño, and What Does It Mean for Weather, Water, and the Global Economy?

    August 18, 2026

    Turf War Between Claude Agents Leads to Self-Replicating Malware

    August 18, 2026
    Latest Posts

    Wisconsin’s Democratic primary for governor: a look at the 5 remaining

    July 27, 2026

    UK CO2 storage project that will reuse existing infrastructure secures lease

    July 27, 2026

    Bangladesh shipbreakers push back against stricter environmental standards

    July 27, 2026

    Subscribe to News

    Get the latest sports news from NewsSite about world, sports and politics.

    NCIJ Network is an independent digital news platform delivering trusted investigative journalism, European and global news, in-depth analysis, and fact-based reporting with accuracy, transparency, and integrity.

    Facebook X (Twitter) Instagram Pinterest YouTube

    Top economists urge Burnham to sign UK up to UN initiative on global inequality | Inequality

    August 18, 2026

    Hegseth Campaigns for Republican Zach Nunn in Iowa

    August 18, 2026

    What Is El Niño, and What Does It Mean for Weather, Water, and the Global Economy?

    August 18, 2026

    Subscribe to Updates

    Get the latest creative news from FooBar about art, design and business.

    Type above and press Enter to search. Press Esc to cancel.