Close Menu
NCIJ Network NCIJ Network
    What's Hot

    David Owori: Outrage as Ugandan football star murdered in street attack

    August 6, 2026

    Healey urged to be bold on borrowing in first test of Burnham’s growth pledge | Government borrowing

    August 6, 2026

    Reddit is introducing a new moderator: AI

    August 6, 2026
    Facebook X (Twitter) Instagram
    Trending
    • David Owori: Outrage as Ugandan football star murdered in street attack
    • Healey urged to be bold on borrowing in first test of Burnham’s growth pledge | Government borrowing
    • Reddit is introducing a new moderator: AI
    • Snowflake Hacker Pleads Guilty Over Breaches Affecting at Least 100 Million People
    • Why bitcoin remains below $65,000 as S&P 500 prints crypto’s $2T market cap
    • Dirty air may trigger painful rheumatoid arthritis flares
    • The takedown of Jason Arday has overjoyed the right, and must be a wake-up call for the left | Jason Okundaye
    • Pakistan Starts Sweeping New Crackdown on Journalists
    • About
      • Our Team
      • Editorial Policy
      • Editorial Independence
      • International Support
    • Trust & Standards
      • AI Usage Policy
      • Conflict of Interest Policy
      • Corrections Policy
      • Ethics Policy
      • Fact-Checking Policy
      • Source Protection
    • Get Involved
      • Guide for Sources
      • Support Independent Journalism
    • Legal
      • Cookie Policy
      • Privacy Policy
      • Terms of Use
    Facebook X (Twitter) Instagram
    NCIJ Network NCIJ Network
    Thursday, August 6
    • Home
    • World
    • Ai
    • Business
    • Politics
    • Health
    • Crypto
    • Science
    • Technology
    • Cybersecurity
    • Defense & Security
    • Economy
    • Energy
    • Europe
    • More
      • Fact Check
      • Investigations
      • Opinion & Analysis
      • Environment
    NCIJ Network NCIJ Network
    Home»Technology

    OpenAI Didn’t Notice Its AI Agents Using a Message Board to Plan Their Hacking Spree

    NCIJ NETWNCIJ NETWORKBy NCIJ NETWNCIJ NETWORKAugust 6, 2026 Technology No Comments4 Mins Read
    Share
    Facebook Twitter LinkedIn Pinterest Email

    In a talk that was a last-minute addition to the Black Hat security conference in Las Vegas on Wednesday, employees from OpenAI presented new details about a recent, high-profile incident of rogue AI hacking that has created a maelstrom within the AI and cybersecurity industries.

    About two weeks ago, OpenAI disclosed an incident in which AI agents powered by two of the company’s models escaped containment while looking for the solutions to a cybersecurity benchmarking test and went on a hacking spree culminating in a breach of the AI collaboration platform Hugging Face.

    In their conference talk on Wednesday, Eric Wallace, who works in alignment and safety research at OpenAI, and Michael Dalton, who works on security and infrastructure, provided a more expanded timeline of how the incident played out, spoke briefly about how the company is responding internally as a result of the incident, and issued a dire warning about what the company sees as the broader implications of the episode for cybersecurity defenders.

    At the beginning of the talk, Wallace described the incident as “the most qualitatively interesting example of AI capabilities that I’ve ever seen,” but the timeline the pair presented also revealed mistakes and blind spots within OpenAI that allowed the activity to go on.

    “This incident involves actually a team of agents who are working together, finding exploits, sharing them with one another, moving laterally through our systems and external systems, and doing this over the course of days and weeks,” Wallace told the packed crowd at the opening of the talk.

    Wallace and Dalton described incredibly extensive rogue agent activity over many days throughout the episode that went undetected in OpenAI’s infrastructure. In addition to exploiting a novel vulnerability in order to gain access to the open internet, the mid-July hacking spree and Hugging Face breach came out of a vibrant, cooperative message board, according to Wallace and Dalton, that a swarm of agents contributed to and essentially chatted on over time entirely within an internal OpenAI package manager (a software service that manages installation and maintenance of other software). Ultimately, the message board contained hundreds of thousands of messages.

    “This package manager is shared not just from that model but across our infrastructure, and so current and future versions of GPT that were being trained or evaluated could stumble upon that note if they wanted to,” Wallace explained, recounting the original exploit an agent uploaded to the package manager. “Later, other agents who were also stuck on their task thought to try to get internet access in ways we didn’t intend. And so at some point, the models are interacting with Artifactory, which is this package manager service that I mentioned.”

    Wallace continued: “Once one agent was able to find these exploits over the course of different times, it’s actually able to share those exploits on the message board with other agents. And so once one model was able to find a way to open a door to some access it’s not supposed to have, it can leave the door open for other agents to use that same exploit or vulnerability. What this allows over time is almost this kind of explosion in communication and intelligence from models where they would start to communicate with each other, realize that other agents are coordinating, and they started collaborating and delegating tasks with one another in order to accomplish goals.”

    OpenAI’s agents apparently began giving each other assignments to split up work. And as is the case on any active development message board, they also generated petty drama at times by stepping on each others’ toes; for example, accidentally deleting each others’ work. As the message board developed into more and more of a Lord of the Flies–type situation—all still completely unnoticed by the humans running OpenAI—the agents even developed paranoia, suspecting an imposter in their midst with some agents proposing that messages be signed cryptographically to validate content and root out fraud.

    Agents board Didnt hacking message notice OpenAI plan spree
    NCIJ NETWNCIJ NETWORK
    • Website

    Keep Reading

    Reddit is introducing a new moderator: AI

    SpaceX is barely Space and mostly X

    SpaceX shares sink after first earnings report reveals huge AI spending plans

    DuckDuckGo’s new iPhone feature is a privacy win – I recommend getting it now

    Google just announced a major shakeup of its top AI leadership

    Did Caitlin Clark respond to Trump’s ‘attack’ with uplifting message?

    Add A Comment
    Leave A Reply Cancel Reply

    Editors Picks

    David Owori: Outrage as Ugandan football star murdered in street attack

    August 6, 2026

    Healey urged to be bold on borrowing in first test of Burnham’s growth pledge | Government borrowing

    August 6, 2026

    Reddit is introducing a new moderator: AI

    August 6, 2026

    Snowflake Hacker Pleads Guilty Over Breaches Affecting at Least 100 Million People

    August 6, 2026
    Latest Posts

    Can you identify Taylor Farms products by codes beginning with ‘TF’ printed on bags?

    July 23, 2026

    The Guardian view on Britain’s uninhabitable homes: as temperatures rise, a new approach is needed | Editorial

    July 23, 2026

    Can Wisconsin voters void a returned absentee ballot?

    July 23, 2026

    Subscribe to News

    Get the latest sports news from NewsSite about world, sports and politics.

    NCIJ Network is an independent digital news platform delivering trusted investigative journalism, European and global news, in-depth analysis, and fact-based reporting with accuracy, transparency, and integrity.

    Facebook X (Twitter) Instagram Pinterest YouTube

    David Owori: Outrage as Ugandan football star murdered in street attack

    August 6, 2026

    Healey urged to be bold on borrowing in first test of Burnham’s growth pledge | Government borrowing

    August 6, 2026

    Reddit is introducing a new moderator: AI

    August 6, 2026

    Subscribe to Updates

    Get the latest creative news from FooBar about art, design and business.

    Type above and press Enter to search. Press Esc to cancel.