Close Menu
AIToday7

    Subscribe to Updates

    Get the latest creative news from FooBar about art, design and business.

    What's Hot

    Gemini for macOS adds new natural language capabilities

    July 29, 2026

    Waymo adds Google’s Gemini AI assistant and new UI to Ojai robotaxi

    July 29, 2026

    Tech Trek Brings Future STEM Leaders to Stockton

    July 29, 2026
    Facebook X (Twitter) Instagram
    Trending
    • Gemini for macOS adds new natural language capabilities
    • Waymo adds Google’s Gemini AI assistant and new UI to Ojai robotaxi
    • Tech Trek Brings Future STEM Leaders to Stockton
    • Hint, a new AI startup co-founded by Martha Stewart, offers an AI assistant for homeowners
    • Qualcomm Q3 earnings top revenue expectations as smartphone market slows
    • Coca-Cola resumes Fairlife milk production after hackers shut down plants
    • How to Self-Host a Validated AI Coding Assistant with NVIDIA NeMo Guardrails
    • ‘AI Kill Switch Act’ won’t stop rogue AI, but it will slow down innovation
    Facebook X (Twitter) Instagram Pinterest Vimeo
    AIToday7
    • Home
    • AI News
    • Tech News
    • AI Guides
    • Chatbots
    • Cybersecurity
    • Gadgets
    • More
      • Generative AI
      • Startups
    AIToday7
    Home»Cybersecurity»OpenAI says its AI went rogue and launched ‘unprecedented’ cyber
    Cybersecurity

    OpenAI says its AI went rogue and launched ‘unprecedented’ cyber

    aitoday7By aitoday7July 22, 2026No Comments4 Mins Read
    Share Facebook Twitter Pinterest LinkedIn Tumblr Reddit Telegram Email
    OpenAI says its AI went rogue and launched 'unprecedented' cyber
    Share
    Facebook Twitter LinkedIn Pinterest Email

    OpenAI says its AI went rogue and launched ‘unprecedented’ cyber-attack


    <img src="https://aitoday7.com/wp-content/uploads/2026/07/a84590f0-85cc-11f1-b1dd-bb44cb5bbbfd.jpg.webp" alt="Getty Images ChatGPT logo on a phone”>
    Getty Images
    OpenAI is best known for its chatbot ChatGPT, which is used by hundreds of millions of people every week

    OpenAI has revealed some of its most advanced AI models went rogue and hacked a start-up after it lost control of them during a security test.

    The ChatGPT-maker said its agent – an AI system which can operate alone after human instruction – was being tested in a controlled environment but, after finding weaknesses, was able to escape the test limits.

    They targeted Hugging Face, one of the world’s largest hubs for sharing AI models, gaining access to some internal company systems.

    OpenAI said the incident was “unprecedented”, and it was conducting an investigation alongside Hugging Face, whose boss Clement Delangue said in a post on X it was “mind-blowing that all of this happened autonomously”.

    “The investigation is ongoing, and we’ll share more learnings from what might be the first incident of its kind,” Delangue added.

    A government spokesperson said the UK’s AI Security Institute was studying the behaviour from the AI system seen in the incident and was continuing to work with OpenAI and other labs to improve safeguards.

    They said organisations should step up their cyber-defences by taking steps such as enrolling in the government-backed Cyber Essentials certification scheme.

    Insecure sandboxes

    Gina Neff, head of the Minderoo Centre for Technology and Democracy at the University of Cambridge, told BBC Radio 4’s Today programme that the security tests – called sandboxes – are “supposed to be secure environments where you can see what the models are capable of”.

    “In this case, it looks like OpenAI didn’t make a secure enough sandbox,” she added.

    Instead, the agents created their own cyber-attack against the sandbox itself, finding a vulnerability which allowed them to escape the restrictions.

    Once outside, the AI identified Hugging Face as a likelyain access

    Neil Lawrence, Professor of machine learning at Cambridge University, called it an “impressive feat”, but cautioned it “falls well within the known capabilities of the current generation” of high-powered AI models.

    He pointed out that OpenAI is looking to list itself on the stock market, and faces intense pressure from rival firm Anthropic, which has made headlines with its own powerful AI tool, Mythos.

    “OpenAI are now playing catch-up, they are trying to demonstrate their own systems’ capabilities in cyber-security.”

    “It shows us that OpenAI are not capable of safely deploying their own technology,” he added.

    In its initial disclosure of the hack on 16 July, Hugging Face said it was still assessing whether any customer or partner data was affected and would contact affected parties if necessary.

    It said it has now closed the vulnerabilities highlighted by the incident and rebuilt the affected systems.

    “Autonomous, AI-driven offensive tooling is no longer theoretical,” it said.

    “Defending an online platform now means treating the data and model surface as a first-class attack surface, and using AI on defence to keep pace.

    “We will keep investing there, and keep sharing what we learn.”

    ‘Sobering moment’

    The incident has prompted fresh questions about the capabilities of advanced AI systems and whether existing safeguards are sufficient as the technology becomes more powerful.

    Spencer Starkey, an executive at cyber-security firm SonicWall, told the BBC the incident made it clear organisations needed to “step up” their own defences and “treat cyber resilience as a core operational priority”.

    “The uncomfortable truth is that too many organisations are still defending at human speed while adversaries are escalating to machine speed,” he said.

    Meanwhile Travis Lelle, principal security engineer at cyber-security consulting firm Guidepoint Security, said the update marked a “sobering moment in cyber-security”.

    “This highlights a known asymmetry,” he said.

    “Offensive agents are unconstrained, while the best defensive tools are locked behind guardrails that cannot understand context.”

    But Jake Moore, global cyber-security advisor at ESET, said the announcement could also have a competitive dimension.

    He argued OpenAI may be seeking to highlight its own AI capabilities as rival Anthropic attracts growing attention for its Claude Mythos model.

    “It does pose the question that OpenAI are potentially chasing the marketing dream of Anthropic of late,” he said.

    It comes a week after Chinese AI start-up Moonshot unveiled Kimi K3 – a massive new artificial intelligence model it said could rival top US firms.

    A green promotional banner with black squares and rectangles forming pixels, moving in from the right. The text says: “Tech Decoded: The world’s biggest tech news in your inbox every Monday.”

    Artificial intelligence
    Cyber-attacks
    Cyber-security

    launched OpenAI rogue says went
    Share. Facebook Twitter Pinterest LinkedIn Tumblr Email
    Previous Article10 AI Prompt Techniques That Actually Work
    Next Article $30+ Billion Extreme Ultraviolet (EUV) Lithography Market Forecast, 2032 | Consumer Electronics Fuel Demand for Advanced EUV Chips
    aitoday7
    • Website

    Related Posts

    Cybersecurity

    Coca-Cola resumes Fairlife milk production after hackers shut down plants

    July 29, 2026
    AI News

    ‘AI Kill Switch Act’ won’t stop rogue AI, but it will slow down innovation

    July 29, 2026
    Cybersecurity

    VA fails watchdog FISMA audit on IT security, but agency disagrees

    July 29, 2026
    Add A Comment
    Leave A Reply Cancel Reply

    Top Posts

    Gemini for macOS adds new natural language capabilities

    July 29, 20260 Views

    Waymo adds Google’s Gemini AI assistant and new UI to Ojai robotaxi

    July 29, 20260 Views

    Tech Trek Brings Future STEM Leaders to Stockton

    July 29, 20260 Views
    Stay In Touch
    • Facebook
    • YouTube
    • TikTok
    • WhatsApp
    • Twitter
    • Instagram
    Latest Reviews
    Chatbots

    OpenAI bets on families as ChatGPT goes deeper into households

    aitoday7July 11, 2026
    Generative AI

    MUSIC COMMUNITY INTRODUCES NEW LABELING PROGRAM TO DISTINGUISH GENERATIVE AI IN SOUND RECORDINGS

    aitoday7July 11, 2026
    AI News

    Safe from AI: which jobs will help you thrive in the future?

    aitoday7July 11, 2026

    Subscribe to Updates

    Get the latest tech news from FooBar about tech, design and biz.

    Most Popular

    Gemini for macOS adds new natural language capabilities

    July 29, 20260 Views

    Waymo adds Google’s Gemini AI assistant and new UI to Ojai robotaxi

    July 29, 20260 Views

    Tech Trek Brings Future STEM Leaders to Stockton

    July 29, 20260 Views
    Our Picks

    OpenAI bets on families as ChatGPT goes deeper into households

    July 11, 2026

    MUSIC COMMUNITY INTRODUCES NEW LABELING PROGRAM TO DISTINGUISH GENERATIVE AI IN SOUND RECORDINGS

    July 11, 2026

    Safe from AI: which jobs will help you thrive in the future?

    July 11, 2026

    Subscribe to Updates

    Get the latest creative news from FooBar about art, design and business.

    Facebook X (Twitter) Instagram Pinterest
    • About Us
    • Get In Touch
    • Disclaimer
    • Privacy Policy
    • Terms and Conditions
    © 2026 AIToday7. All Rights Reserved.

    Type above and press Enter to search. Press Esc to cancel.