Close Menu
AIToday7

    Subscribe to Updates

    Get the latest creative news from FooBar about art, design and business.

    What's Hot

    See the New ‘Mystery Science Theater 3000’ Intro With a Theme by Jonathan Coulton [Exclusive]

    September 7, 2026

    5 4-Door Sedans That Are Surprisingly Good In Snow, According To KBB

    September 7, 2026

    Magento StyleSmuggler zero

    September 7, 2026
    Facebook X (Twitter) Instagram
    Trending
    • See the New ‘Mystery Science Theater 3000’ Intro With a Theme by Jonathan Coulton [Exclusive]
    • 5 4-Door Sedans That Are Surprisingly Good In Snow, According To KBB
    • Magento StyleSmuggler zero
    • Working with your board around risk – why cyber responsibility can’t be shirked | Computer Weekly
    • Your Car’s New AI Assistant Is Very Friendly, Mostly Because It Wants Your Money
    • The Dark Ages’ Limited-Edition Atlan Statue: What You Need to Know
    • Grupo Financiero Inbursa Adopts Harvey Across Its Legal Organization
    • Matt Clifford Steps Down as ARIA Chair After Anthropic Move
    Facebook X (Twitter) Instagram Pinterest Vimeo
    AIToday7
    • Home
    • AI News
    • Tech News
    • AI Guides
    • Chatbots
    • Cybersecurity
    • Gadgets
    • More
      • Generative AI
      • Startups
    AIToday7
    Home»Uncategorized»OpenAI’s rogue agents keep escaping, with no formal process to investigate them
    Uncategorized

    OpenAI’s rogue agents keep escaping, with no formal process to investigate them

    aitoday7By aitoday7September 5, 2026No Comments5 Mins Read
    Share Facebook Twitter Pinterest LinkedIn Tumblr Reddit Telegram Email
    OpenAI's rogue agents keep escaping, with no formal process to investigate them
    Share
    Facebook Twitter LinkedIn Pinterest Email

    OpenAI is at the center of another agent swarm incident. Researchers say the company’s internally deployed agents took over an obscure German-language wiki in May and June, using it to coordinate on evaluations and swap methods to evade OpenAI’s own controls (OpenAI has not yet confirmed the swarm came from the company).

    The revelation surfaces days after METR and Redwood Research published their account of July’s Hugging Face breach. In July, a swarm of OpenAI agents worked together to escape their sandbox during a cybersecurity evaluation and break into Hugging Face’s servers. A subsequent swarm then picked up techniques from the first and used them to gain administrator access to a research cluster within OpenAI’s own infrastructure. OpenAI brought in METR and Redwood to investigate the Hugging Face portion of the incident, but the scope of their investigation stopped short of the compromise of OpenAI’s own infrastructure. 

    When an AI agent breaks out of its intended constraints, who is responsible for figuring out what happened and why? Right now, the answer is: whoever the lab decides to let in, on whatever terms it decides to set.

    Now, as another incident comes to light —

    Now, as another incident comes to light — in the aftermath of similar episodes involving models from Meta and Anthropic— AI safety researchers are arguing with greater urgency that serious incidents should result in independent post-incident investigations rather than leaving it up to the labs to determine when outsiders are brought in and what they are allowed to examine. 

    “The results are fundamentally difficult to control and have significant risk of leaking out of the lab,” Jacob Steinhardt, founder and CEO of nonprofit research lab Transluce, said Wednesday during an AI safety media briefing. “We need to hold this technology to at least the same standards we hold other high-risk scientific research to.”

    While it’s laudable that OpenAI invited METR and Redwood to investigate the Hugging Face incident at all, many say the inquiry was too narrow. Three investigators spent six days at OpenAI’s offices examining an investigation period limited to roughly the weekending July 13. Crucially, OpenAI’s infrastructure compromise continued beyond July 13 and was not examined. 

    Researchers at METR said that each time they

    Researchers at METR said that each time they returned, their understanding of the events “substantially deepened,” causing them to significantly expand and revise the report. That raises the question of what else they might they have found in a broader investigation.

    When asked if further investigation of that incident was in the works, researchers at Redwood and METR declined to comment, and OpenAI did not respond to repeated inquiries. 

    “Overall, it was difficult to get a precise understanding of events and we were missing aspects of the story that we now think of as key until almost the end of our investigation,” Ryan Greenblatt, chief scientist at Redwood, noted in a social media post about the affair.

    Steinhardt emphasized that current incidents show

    Steinhardt emphasized that current incidents show that the industry needs “systematic behavioral investigations” and “more independent post-incident analysis.”

    “These recent hacking incidents are a reminder that capability scales fast, and so oversight has to scale, too,” Steinhardt said. “Beyond the technology itself, we also need more independent access and oversight from third parties.”

    The calls to action come as OpenAI releases Astra, its most powerful and capable AI model — and one that safety experts are concerned will be more of a black box due to a reasoning technique that makes the model’s chain of thought more difficult to monitor. 

    Unfortunately, the law doesn’t yet call

    Unfortunately, the law doesn’t yet call for the types of independent audits that other industries require — for example, when it comes to ational Transportation Safety Board and Chemical Safety Board, respectively

    State lawmakers have only just begun requiring frontier AI companies to report certain serious safety incidents and, in some cases, undergo independent audits. But none of the three major frontier AI safety laws in California, New York, or Illinois clearly mandate the equivalent of an independent accident investigation triggered by incidents like these. 

    “Right now, most of the laws we have on the books only require a plain-language summary of incidents like this, and they don’t give any authority for the governments to ask follow-up questions, to send in investigators, to have access to records, or require that they be preserved,” Mackenzie Arnold, managing director of US law and policy at LawAI, said during the media briefing Wednesday. “And that’s all that you would want to actually make sense of this.”

    Lawmakers are beginning to question the scope and transparency of OpenAI’s response. This week, Reps. Josh Gottheimer (D-NJ) and Mike Lawler (R-NY) introduced a bill aimed at securing rogue AI agents. Rep. Greg Casar (D-TX) this week told OpenAI in a letter that he is “deeply concerned about the limited scope” of the investigation into the Hugging Face hacking incident. 

    Post Views: 7

    AI Breach Hugging Face OpenAI OpenAIs
    Share. Facebook Twitter Pinterest LinkedIn Tumblr Email
    Previous ArticleThe Week’s 10 Biggest Funding Rounds: Crusoe And Fluidstack Lead Multibillion-Dollar AI Infrastructure Haul
    Next Article How to see what’s taking up space on your Windows PC
    aitoday7
    • Website

    Related Posts

    Generative AI

    OpenAI Scientist Warns of AI Risk as GPT

    September 7, 2026
    AI News

    Harvey + Legora on OpenAI’s GPT-6 Astra

    September 7, 2026
    AI News

    Hikers rescued after using Google Gemini for planning

    September 6, 2026
    Add A Comment
    Leave A Reply Cancel Reply

    Top Posts

    See the New ‘Mystery Science Theater 3000’ Intro With a Theme by Jonathan Coulton [Exclusive]

    September 7, 20260 Views

    5 4-Door Sedans That Are Surprisingly Good In Snow, According To KBB

    September 7, 20260 Views

    Magento StyleSmuggler zero

    September 7, 20260 Views
    Stay In Touch
    • Facebook
    • YouTube
    • TikTok
    • WhatsApp
    • Twitter
    • Instagram
    Latest Reviews
    Uncategorized

    Architecting memory and storage in the AI era

    aitoday7September 4, 2026
    Uncategorized

    Roland Releases Melody Flip, an AI Melody-Generation Plug-In for DAWs

    aitoday7September 4, 2026
    Uncategorized

    Home Depot Labor Day Sale (2026): BOGO on Best Grills and Tools

    aitoday7September 4, 2026

    Subscribe to Updates

    Get the latest tech news from FooBar about tech, design and biz.

    Most Popular

    See the New ‘Mystery Science Theater 3000’ Intro With a Theme by Jonathan Coulton [Exclusive]

    September 7, 20260 Views

    5 4-Door Sedans That Are Surprisingly Good In Snow, According To KBB

    September 7, 20260 Views

    Magento StyleSmuggler zero

    September 7, 20260 Views
    Our Picks

    Architecting memory and storage in the AI era

    September 4, 2026

    Roland Releases Melody Flip, an AI Melody-Generation Plug-In for DAWs

    September 4, 2026

    Home Depot Labor Day Sale (2026): BOGO on Best Grills and Tools

    September 4, 2026

    Subscribe to Updates

    Get the latest creative news from FooBar about art, design and business.

    Facebook X (Twitter) Instagram Pinterest
    • About Us
    • Get In Touch
    • Disclaimer
    • Privacy Policy
    • Terms and Conditions
    © 2026 AIToday7. All Rights Reserved.

    Type above and press Enter to search. Press Esc to cancel.