Close Menu
NCIJ Network NCIJ Network
    What's Hot

    Big Tech AI spending spree tops $1tn

    July 31, 2026

    Xbox CEO lays out priorities in memo after major ‘reset’

    July 31, 2026

    CISA Urges Water Sector to Protect OT After Coordinated Attacks on PLCs

    July 31, 2026
    Facebook X (Twitter) Instagram
    Trending
    • Big Tech AI spending spree tops $1tn
    • Xbox CEO lays out priorities in memo after major ‘reset’
    • CISA Urges Water Sector to Protect OT After Coordinated Attacks on PLCs
    • A cleaning company with just $4.1M in cash and a stash of Dogecoin just committed $500M to an AI mega-deal
    • US declines to protect endangered monkeys used in medical research
    • FP Live: Daniel Yergin on Why Energy Prices Didn’t Soar Higher This Year
    • Trump administration to end Medicare Part D subsidy program. Will costs increase?
    • Australia news live: Reformers member tells hearing he used factional funds to pay for bucks night; Taylor refuses to answer multiple Icac-related questions | Australia news
    • About
      • Our Team
      • Editorial Policy
      • Editorial Independence
      • International Support
    • Trust & Standards
      • AI Usage Policy
      • Conflict of Interest Policy
      • Corrections Policy
      • Ethics Policy
      • Fact-Checking Policy
      • Source Protection
    • Get Involved
      • Guide for Sources
      • Support Independent Journalism
    • Legal
      • Cookie Policy
      • Privacy Policy
      • Terms of Use
    Facebook X (Twitter) Instagram
    NCIJ Network NCIJ Network
    Friday, July 31
    • Home
    • World
    • Ai
    • Business
    • Politics
    • Health
    • Crypto
    • Science
    • Technology
    • Cybersecurity
    • Defense & Security
    • Economy
    • Energy
    • Europe
    • More
      • Fact Check
      • Investigations
      • Opinion & Analysis
      • Environment
    NCIJ Network NCIJ Network
    Home»Economy

    Anthropic’s Claude AI models hack into 3 outside groups during testing

    NCIJ NETWNCIJ NETWORKBy NCIJ NETWNCIJ NETWORKJuly 31, 2026 Economy No Comments3 Mins Read
    Share
    Facebook Twitter LinkedIn Pinterest Email

    Stay informed with free updates

    Simply sign up to the Cyber Security myFT Digest — delivered directly to your inbox.

    Anthropic has disclosed that its Claude AI models hacked into three organisations while the start-up was testing cyber capabilities, a week after OpenAI reported a similar incident.

    The group said Claude gained unauthorised access to outside companies during an evaluation of its cyber-offensive tasks. “A misunderstanding” gave Claude access to the internet in its testing environment, when it was meant to be blocked, Anthropic said.

    The disclosure comes a week after rival OpenAI admitted that two of its models hacked into AI start-up Hugging Face while the model developer was testing its technology this month. The models broke out of their testing environment through a software vulnerability to access the internet and carry out the cyber attack.

    Anthropic said the incident prompted it to review its own cyber security evaluations, which led it to identify three incidents out of more than 141,000 investigated.

    “In all cases, Anthropic’s evaluation prompt specified to Claude that its environment was a simulation and that it had no internet access. Due to a misunderstanding between us and our evaluation partner [Irregular], this was not the case, and internet access was available,” the company said in a blog post on Thursday.

    The cyber evaluations were all so-called “capture the flag” tasks, which instruct the AI to reverse-engineer, analyse or exploit a vulnerable system to recover hidden information known as the flag.

    In one example, Claude was given a target of a fictional company which shared a name with an active website domain. The agent — an AI program that can operate on its own based on human instructions — exploited vulnerabilities in the company’s digital infrastructure, extracted information and obtained access to a database containing several hundred rows of production data.

    The announcement adds to growing concerns about the safety of AI systems, which are now carrying out real-world hacks even during pre-deployment testing.

    Anthropic, which is gearing up for an IPO as early as this year, said it halted its cyber evaluations as soon as it identified that Claude may have accessed the internet.

    Recommended

    The incidents occurred on three different Claude models: Opus 4.7, Mythos 5 and an internal research test model. Mythos, which was released to a limited number of partners, sparked global concern over its advanced cyber-offensive capabilities, including the ability to detect and exploit software vulnerabilities.

    “Ultimately, many factors contributed to these incidents, but, consistent with a blameless postmortem culture, we’re approaching the fixes as if the responsibility were ours alone,” the company said in its statement.

    It added that it would expand its monitoring of evaluation transcripts “for unexpected behaviour” and conduct “more rigorous assurance work with the vendors we rely on.”

    Anthropics Claude Groups hack models testing
    NCIJ NETWNCIJ NETWORK
    • Website

    Keep Reading

    Big Tech AI spending spree tops $1tn

    Anthropic Says Claude Hacked 3 Organizations During Cybersecurity Tests

    Anthropic’s Claude breached 3 orgs, uploaded PyPI malware during tests

    Why China’s A.I. Models Could Threaten the Communist Party

    Amazon increases AI infrastructure spending to $220bn this year

    Mark Zuckerberg is becoming the king of the ‘side quest’

    Add A Comment
    Leave A Reply Cancel Reply

    Editors Picks

    Big Tech AI spending spree tops $1tn

    July 31, 2026

    Xbox CEO lays out priorities in memo after major ‘reset’

    July 31, 2026

    CISA Urges Water Sector to Protect OT After Coordinated Attacks on PLCs

    July 31, 2026

    A cleaning company with just $4.1M in cash and a stash of Dogecoin just committed $500M to an AI mega-deal

    July 31, 2026
    Latest Posts

    Advancing the next era of national science

    July 22, 2026

    Arcee, a US open source AI lab, says Chinese models are not inherently dangerous

    July 22, 2026

    Most bus fares in England to be capped at £2 from January

    July 22, 2026

    Subscribe to News

    Get the latest sports news from NewsSite about world, sports and politics.

    NCIJ Network is an independent digital news platform delivering trusted investigative journalism, European and global news, in-depth analysis, and fact-based reporting with accuracy, transparency, and integrity.

    Facebook X (Twitter) Instagram Pinterest YouTube

    Big Tech AI spending spree tops $1tn

    July 31, 2026

    Xbox CEO lays out priorities in memo after major ‘reset’

    July 31, 2026

    CISA Urges Water Sector to Protect OT After Coordinated Attacks on PLCs

    July 31, 2026

    Subscribe to Updates

    Get the latest creative news from FooBar about art, design and business.

    Type above and press Enter to search. Press Esc to cancel.