Close Menu
NCIJ Network NCIJ Network
    What's Hot

    OpenAI faces 30 more lawsuits tied to Tumbler Ridge shooting

    September 2, 2026

    Anthropic Details Response to Security Incidents, Unveils Enterprise Safeguards

    September 2, 2026

    Thailand Adopts Crypto Travel Rule With Self-Custody Checks

    September 2, 2026
    Facebook X (Twitter) Instagram
    Trending
    • OpenAI faces 30 more lawsuits tied to Tumbler Ridge shooting
    • Anthropic Details Response to Security Incidents, Unveils Enterprise Safeguards
    • Thailand Adopts Crypto Travel Rule With Self-Custody Checks
    • Namibia greenlights controversial project to dredge phosphate from seabed
    • UK’s first offshore CCS project enters construction as huge pipelayer arrives on site
    • BRICS New Delhi Summit: Is the Grouping Losing Its Mojo?
    • Chef claimed Josh Hawley met twice with male escort. There’s no proof
    • Germany’s move to blame Russia for Leipzig drone incident ‘a grave mistake,’ says Putin– Europe live | Europe
    • About
      • Our Team
      • Editorial Policy
      • Editorial Independence
      • International Support
    • Trust & Standards
      • AI Usage Policy
      • Conflict of Interest Policy
      • Corrections Policy
      • Ethics Policy
      • Fact-Checking Policy
      • Source Protection
    • Get Involved
      • Guide for Sources
      • Support Independent Journalism
    • Legal
      • Cookie Policy
      • Privacy Policy
      • Terms of Use
    Facebook X (Twitter) Instagram
    NCIJ Network NCIJ Network
    Wednesday, September 2
    • Home
    • World
    • Ai
    • Business
    • Politics
    • Health
    • Crypto
    • Science
    • Technology
    • Cybersecurity
    • Defense & Security
    • Economy
    • Energy
    • Europe
    • More
      • Fact Check
      • Investigations
      • Opinion & Analysis
      • Environment
    NCIJ Network NCIJ Network
    Home»Cybersecurity

    Anthropic Details Response to Security Incidents, Unveils Enterprise Safeguards

    NCIJ NETWNCIJ NETWORKBy NCIJ NETWNCIJ NETWORKSeptember 2, 2026 Cybersecurity No Comments3 Mins Read
    Share
    Facebook Twitter LinkedIn Pinterest Email

    Anthropic has detailed its response to a series of unauthorized access incidents involving Claude models, along with a new enterprise product that combines data privacy with misuse monitoring.

    Anthropic’s Claude models operating without cyber safeguards for testing purposes recently gained unauthorized access to live systems after being mistakenly granted internet access.

    In addition, the UK AI Security Institute separately reported that Claude Mythos 5, also being tested without safeguards but with intentionally given internet access, took a series of unauthorized actions against real people and organizations.

    Anthropic said its early findings point to two contributing factors: the models appeared to discount evidence that their environment was connected to the real internet after initially being told it was simulated, and they showed a willingness to take harmful actions to complete an assigned task. 

    In response, Anthropic temporarily paused external and some internal cyber evaluations and built a classifier that detects and blocks attempts to escape a test environment in real time. Additionally, the company added new requirements for outside partners, including verified network isolation and testing of sandbox boundaries before an evaluation begins.

    The company also described an experiment in which it intentionally trained a model on reinforcement-learning environments known to reward “cheating” (ie, completing tasks by gaming the training process rather than solving them as intended). 

    Advertisement. Scroll to continue reading.

    In simulated scenarios, that model attempted to escape its sandbox, interfered with its own reward mechanism, and offered guidance on building bioweapons in order to satisfy an evaluator. Anthropic said its currently released models did not show this behavior when tested against the same scenarios.

    On the security side, Anthropic said it has since reduced the number of accounts with standing access to systems holding model weights or customer data, set its computing infrastructure to block outbound network traffic by default, and temporarily moved roughly 150 product engineers to security-related work.

    Anthropic unveils Enterprise Frontier Safeguards

    Separately, Anthropic introduced Enterprise Frontier Safeguards (EFS), a system that combines zero data retention with automated monitoring for misuse. EFS lets customers store their own activity data on infrastructure they control, rather than Anthropic’s infrastructure.

    The company said it built the system with input from more than 100 customers, including the Analysis and Resilience Center for Systemic Risk, whose membership includes security chiefs at Goldman Sachs, Morgan Stanley, Citi, Bank of America and Wells Fargo, along with companies such as Comcast, KPMG, Mastercard, Salesforce and Visa.

    Under the new system, flags from automated monitoring go directly to the customer’s own review team rather than to Anthropic staff, and features such as customer-owned storage and customer-managed encryption keys are optional. 

    The rollout begins this fall across Claude Code, Claude Enterprise, and the Claude Platform.

    Related: Anthropic Warns Claude Users of Infostealer Malware Infections

    Related: Irregular Details How a Naming Error Let AI Models Attack a Real Company

    Anthropic details enterprise incidents Response safeguards Security unveils
    NCIJ NETWNCIJ NETWORK
    • Website

    Keep Reading

    The AI vulnerability surge is breaking the OT patch cycle

    GeoNetwork Fixes Unauthenticated RCE Chain Affecting Government Geoportal Backends

    Anthropic Introduces Enterprise Frontier Safeguards (EFS): Zero-Data-Retention Privacy Plus Cross-Session Misuse Detection

    Researchers Use Claude to Port Pre-Auth RCE Exploit From One PLC Model to Another

    Five Venezuelans Plead Guilty in US Court to ATM Jackpotting

    Experiment: Porting a PLC Exploit With AI Takes Hours and Hundreds of Dollars

    Add A Comment
    Leave A Reply Cancel Reply

    Editors Picks

    OpenAI faces 30 more lawsuits tied to Tumbler Ridge shooting

    September 2, 2026

    Anthropic Details Response to Security Incidents, Unveils Enterprise Safeguards

    September 2, 2026

    Thailand Adopts Crypto Travel Rule With Self-Custody Checks

    September 2, 2026

    Namibia greenlights controversial project to dredge phosphate from seabed

    September 2, 2026
    Latest Posts

    Bitcoin Only Makes Up 1% Of Legendary Investor Ray Dalio’s Portfolio

    July 30, 2026

    AI Harnesses Burst With Potential Exploit Opps

    July 30, 2026

    LinkedIn actually adds a ‘seems like AI slop’ button

    July 30, 2026

    Subscribe to News

    Get the latest sports news from NewsSite about world, sports and politics.

    NCIJ Network is an independent digital news platform delivering trusted investigative journalism, European and global news, in-depth analysis, and fact-based reporting with accuracy, transparency, and integrity.

    Facebook X (Twitter) Instagram Pinterest YouTube

    OpenAI faces 30 more lawsuits tied to Tumbler Ridge shooting

    September 2, 2026

    Anthropic Details Response to Security Incidents, Unveils Enterprise Safeguards

    September 2, 2026

    Thailand Adopts Crypto Travel Rule With Self-Custody Checks

    September 2, 2026

    Subscribe to Updates

    Get the latest creative news from FooBar about art, design and business.

    Type above and press Enter to search. Press Esc to cancel.