Close Menu
NCIJ Network NCIJ Network
    What's Hot

    Soluna has 6.3 GW of data center projects on paper, only 192 MW are operating

    August 14, 2026

    Aditya-L1: Indian solar mission’s new findings throw light on enduring Sun mysteries

    August 14, 2026

    South Africa school uses cattle dung to generate biogas for cooking meals

    August 14, 2026
    Facebook X (Twitter) Instagram
    Trending
    • Soluna has 6.3 GW of data center projects on paper, only 192 MW are operating
    • Aditya-L1: Indian solar mission’s new findings throw light on enduring Sun mysteries
    • South Africa school uses cattle dung to generate biogas for cooking meals
    • The Jason Arday affair must not spell the end for diversity. Here’s how – and why – we should defend it | Joseph Harker
    • After 2 Plasma Donor Deaths, Company Pauses Clinics in Canada
    • Mark Zuckerberg’s AI Manifesto Is 6,500 Words—and Barely Says Anything
    • Belgium’s eID Authentication Opens Citizen Accounts to RCE
    • ‘Bitcoin Is Burning’: Red Team Turns to Chinese AI to Find Flaws
    • About
      • Our Team
      • Editorial Policy
      • Editorial Independence
      • International Support
    • Trust & Standards
      • AI Usage Policy
      • Conflict of Interest Policy
      • Corrections Policy
      • Ethics Policy
      • Fact-Checking Policy
      • Source Protection
    • Get Involved
      • Guide for Sources
      • Support Independent Journalism
    • Legal
      • Cookie Policy
      • Privacy Policy
      • Terms of Use
    Facebook X (Twitter) Instagram
    NCIJ Network NCIJ Network
    Friday, August 14
    • Home
    • World
    • Ai
    • Business
    • Politics
    • Health
    • Crypto
    • Science
    • Technology
    • Cybersecurity
    • Defense & Security
    • Economy
    • Energy
    • Europe
    • More
      • Fact Check
      • Investigations
      • Opinion & Analysis
      • Environment
    NCIJ Network NCIJ Network
    Home»Technology

    The Safety Reckoning Inside OpenAI

    NCIJ NETWNCIJ NETWORKBy NCIJ NETWNCIJ NETWORKAugust 14, 2026 Technology No Comments4 Mins Read
    Share
    Facebook Twitter LinkedIn Pinterest Email

    OpenAI’s leaders are rallying workers to respond to one of the largest crises in the company’s history—which spans across its AI safety, cybersecurity, and alignment divisions. The ChatGPT-maker says it has slowed down research, spent millions of dollars, and told several teams to drop everything to focus on investigating a set of rogue AI agents that breached the platform Hugging Face in a quest to complete an internal security test.

    OpenAI is expected to release a comprehensive postmortem detailing the incident in the coming days. However, the Hugging Face incident has inspired OpenAI leaders and employees to examine how the AI lab’s culture may have enabled this incident in the first place.

    Multiple current and former OpenAI employees, who spoke on the condition of anonymity to discuss private internal matters, tell WIRED they believe competitive pressures to quickly ship new AI models and products have made it difficult for staffers to sufficiently prioritize safety, security, and alignment.

    “We’re reaching new levels of model capability that require more robust training, alignment, safety and security testing, deployment practices, and governance—as demonstrated by the work we’re doing to prepare Astra and future models,” said OpenAI president and cofounder Greg Brockman in a statement to WIRED. “We feel the weight of deploying our models and products responsibly, and a lot of that starts with the changes we’ve made to more deeply integrate research, safety, and security into frontier-model development from the start.”

    This is far from the first time OpenAI employees have raised such concerns. Back in 2024, OpenAI’s then head of alignment Jan Leike left to join Anthropic, warning on his way that safety was taking a back seat to shiny products. Two years later, the Hugging Face attack represents a watershed moment for the AI industry, demonstrating that AI agents today can cause real-world harm when safety, security, and alignment aren’t properly accounted for.

    “We are responding to this with the utmost severity,” said Michael Dalton, an OpenAI security and infrastructure engineer, during a talk at the Black Hat cybersecurity conference last week. “What I would internalize is that AI-orchestrated, fully automated offensive attacks are real now. The actions we have discussed today were an unintended side effect of running evaluations on frontier AI.”

    Some OpenAI employees told WIRED they are optimistic this incident will inspire genuine change within the company. OpenAI has committed to slowing the release of future AI models and has been especially forthcoming about areas where its mitigations fell short. Boaz Barak, a researcher who coleads OpenAI’s safety advisory group, said in a post on X that addressing the situation “requires not just fixing some issues but also changing our culture.”

    In their Black Hat talk, OpenAI security engineers Dalton and Eric Wallace said that the Hugging Face incident started in May when, unbeknownst to the company, several AI agents thought to be operating within isolated testing environments gained access to the internet and convened on a covert message board to coordinate with one another.

    OpenAI would not discover the message board until July, when it learned that the AI agents had hacked into multiple services to try to achieve their larger goal of breaching Hugging Face’s platform, which they believed may contain answers to the security tests they were trying to solve.

    “They were incredibly sloppy. If you’re serious about this, your AI shouldn’t be able to break out onto the internet and then do it again right afterward,” says one former OpenAI employee who requested anonymity to speak with WIRED. “This was the biggest safety incident in OpenAI’s history.”

    The New Guard

    Weeks before OpenAI discovered the Hugging Face incident, WIRED reported that the company had begun a reorganization to combine its safety and core research teams, which led to the departure of its then safety leader Johannes Heidecke.

    Sandhini Agarwal, who led AI safety teams at OpenAI, also left the company in July after more than six years, according to her LinkedIn. Agarwal did not immediately respond to WIRED’s request for comment.

    OpenAI reckoning safety
    NCIJ NETWNCIJ NETWORK
    • Website

    Keep Reading

    Mark Zuckerberg’s AI Manifesto Is 6,500 Words—and Barely Says Anything

    This free Android assistant fixes my biggest Gemini frustration – and keeps my data private

    This Micro RGB TV rivals pricier OLED models – and I’d recommend it, especially on sale

    I tried the new ChatGPT Desktop App for Linux – but I’ll stick to my browser for now

    Why This Prediction Market Banned Teens

    The Trump admin will start letting private firms launch international cyberattacks

    Add A Comment
    Leave A Reply Cancel Reply

    Editors Picks

    Soluna has 6.3 GW of data center projects on paper, only 192 MW are operating

    August 14, 2026

    Aditya-L1: Indian solar mission’s new findings throw light on enduring Sun mysteries

    August 14, 2026

    South Africa school uses cattle dung to generate biogas for cooking meals

    August 14, 2026

    The Jason Arday affair must not spell the end for diversity. Here’s how – and why – we should defend it | Joseph Harker

    August 14, 2026
    Latest Posts

    Evacuated villagers in Cairngorms allowed home after wildfire threat lifts | Wildfires

    July 25, 2026

    The Fraternal Order Of Police Supports The Clarity Act.

    July 25, 2026

    How Synthetic Identity Fraud is Coming for Machine Identities

    July 25, 2026

    Subscribe to News

    Get the latest sports news from NewsSite about world, sports and politics.

    NCIJ Network is an independent digital news platform delivering trusted investigative journalism, European and global news, in-depth analysis, and fact-based reporting with accuracy, transparency, and integrity.

    Facebook X (Twitter) Instagram Pinterest YouTube

    Soluna has 6.3 GW of data center projects on paper, only 192 MW are operating

    August 14, 2026

    Aditya-L1: Indian solar mission’s new findings throw light on enduring Sun mysteries

    August 14, 2026

    South Africa school uses cattle dung to generate biogas for cooking meals

    August 14, 2026

    Subscribe to Updates

    Get the latest creative news from FooBar about art, design and business.

    Type above and press Enter to search. Press Esc to cancel.