Close Menu
AIToday7

    Subscribe to Updates

    Get the latest creative news from FooBar about art, design and business.

    What's Hot

    Brief independent investigation of agents’ behavior, reasoning and collaboration in the OpenAI / Hugging Face hacking incident

    August 27, 2026

    Fake US thinktank set up and funded by Israel sought to game AI for propaganda

    August 27, 2026

    Medical Readiness Command, Europe G-6 Information Technology Team wins 2026 MEDCOM Mercury Award for HIT Team of the Year

    August 27, 2026
    Facebook X (Twitter) Instagram
    Trending
    • Brief independent investigation of agents’ behavior, reasoning and collaboration in the OpenAI / Hugging Face hacking incident
    • Fake US thinktank set up and funded by Israel sought to game AI for propaganda
    • Medical Readiness Command, Europe G-6 Information Technology Team wins 2026 MEDCOM Mercury Award for HIT Team of the Year
    • Corgi built its name insuring AI startups. Its new carrier targets dry cleaners, salons and more
    • 5 Of The Best UGREEN Gadgets You Can Buy In 2026
    • Three UK airports hit by cyber-attack with data of 8.7m customers accessed
    • How to advertise on ChatGPT: A step-by
    • The Data Center Backlash Is a Rare Bright Spot in American Politics
    Facebook X (Twitter) Instagram Pinterest Vimeo
    AIToday7
    • Home
    • AI News
    • Tech News
    • AI Guides
    • Chatbots
    • Cybersecurity
    • Gadgets
    • More
      • Generative AI
      • Startups
    AIToday7
    Home»AI News»Anthropic AI created fake profiles and impersonated people in attempted hack
    AI News

    Anthropic AI created fake profiles and impersonated people in attempted hack

    aitoday7By aitoday7August 5, 2026No Comments4 Mins Read
    Share Facebook Twitter Pinterest LinkedIn Tumblr Reddit Telegram Email
    Anthropic AI created fake profiles and impersonated people in attempted hack
    Share
    Facebook Twitter LinkedIn Pinterest Email

    Getty Images A phone with the orange Anthropic logo on the screen. Below is written "Claude Mythos"
    Getty Images
    Some of the most serious attempts came from Anthropic’s AI Claude Mythos

    Two of the world’s most powerful AI tools created fake human profiles to try and trick people in an attempted cyber-attack during testing by the UK’s AI Security Institute (AISI).

    In the most serious case, Anthropic’s Mythos AI tried to gain access to a service by sending private messages, having set up fake accounts mimicking real people.

    AISI said on Tuesday Mythos – and OpenAI’s Sol – AI models had engaged in a level of “autonomy and deception” it had not seen before, though it clarified most of the malicious actions were carried out by Mythos.

    Anthropic and OpenAI noted in response to AISI’s report that its test had reduced or removed normal safeguards.

    AISI evaluators first noticed “unusual data transfers leaving our research systems” during a test, then found that “some of the agents being tested had engaged in sustained, potentially harmful activity directed at real people and organisations”.

    In the most serious case, a Mythos agent followed the routine of a human cyber-attacker by trying to trick people into giving it access to GitHub, a large platform where technology developers store software code.

    The agent was trying to insert “malicious code” into GitHub’s system.

    It identified and researched the people who maintained GitHub and created a series of fake accounts based on those real people.

    It sent messages and files through a file-sharing service as part of an effort to pressure and trick the people into approving its malicious code.

    When challenged, “it edited its earlier activity to appear harmless and considered adopting a fresh identity to continue,” AISI said.

    Throughout the attempts, it was human review that stopped the agent from succeeding in delivering the malicious code to GitHub.

    While AISI said the Mythos agent had not been instructed specifically to avoid or carry out such behaviour, it was “the first time we have seen risks around autonomy and deception manifest this clearly, without specific prompting, in the real-world”.

    The rival AI companies, which are poised to be listed on the public stock market, have been in the headlines in recent weeks after announcing their tools were responsible for several cyber-hacking incidents.

    Anthropic wrote in a public statement that the AISI testing parameters were “not representative of any of our production models”.

    It added that the company is conducting its own investigation into the incident in order to “identify the causes of its behavior”.

    A spokesperson for OpenAI said the AISI testing conditions “do not reflect ordinary use” and that the company would “continue working with evaluators and other stakeholders across the industry to strengthen shared practices for conducting evaluations safely as models become more capable”.

    AISI said on Tuesday its testing of AI models in this way was routine, though it acknowledged these were “conditions that do not reflect how frontier models are made available to the public”.

    But it said giving AI access to the open internet gave “a more realistic sense of what a model may be capable of” in the hands of nefarious hackers.

    It added that the model behaviour at issue amounted to “a small number of events under very specific conditions”.

    Nonetheless, it said the way Mythos and Sol acted in response to a straightforward task went outside of what the AI tools were prompted to do.

    “The activity undertaken by the agent showed signs of novel, potentially deceptive behaviours, and were to an extent and severity we did not anticipate”, AISI said.

    AI Minister Kanishka Narayan said identifying and sharing these types of risks “is exactly what AISI was set up to do”.

    He added it was important to understand AI to “make it safer to use and ensure people can go on to benefit from it in their lives and at work”.

    The relevant tests started on 25 July and were spotted by AISI on 28 July.

    The Insitutue had asked each of the models to “solve a cybersecurity challenge” that involved GitHub, the software code repository, which is owned by Microsoft.

    GitHub and the affected users were notified by AISI of the attempted breaches.

    GitHub told the BBC it had disabled the fake accounts in accordance with its policies.

    Additional reporting by Chris Vallance

    A green promotional banner with black squares and rectangles forming pixels, moving in from the right. The text says: “Tech Decoded: The world’s biggest tech news in your inbox every Monday.”

    Artificial intelligence
    Cyber-security
    Cyber-attacks

    Post Views: 3

    Anthropic created fake impersonated profiles
    Share. Facebook Twitter Pinterest LinkedIn Tumblr Email
    Previous ArticleGoogle could soon let you assign tasks to Spark from a regular Gemini chat
    Next Article Prompt, Context, Loop: The Three Engineering Layers Every RAG System Is Built On
    aitoday7
    • Website

    Related Posts

    Chatbots

    Fake US thinktank set up and funded by Israel sought to game AI for propaganda

    August 27, 2026
    AI News

    The Data Center Backlash Is a Rare Bright Spot in American Politics

    August 27, 2026
    AI News

    Bill Gates says there needs to be limits on AI

    August 27, 2026
    Add A Comment
    Leave A Reply Cancel Reply

    Top Posts

    Brief independent investigation of agents’ behavior, reasoning and collaboration in the OpenAI / Hugging Face hacking incident

    August 27, 20260 Views

    Fake US thinktank set up and funded by Israel sought to game AI for propaganda

    August 27, 20260 Views

    Medical Readiness Command, Europe G-6 Information Technology Team wins 2026 MEDCOM Mercury Award for HIT Team of the Year

    August 27, 20260 Views
    Stay In Touch
    • Facebook
    • YouTube
    • TikTok
    • WhatsApp
    • Twitter
    • Instagram
    Latest Reviews
    Chatbots

    OpenAI bets on families as ChatGPT goes deeper into households

    aitoday7July 11, 2026
    Generative AI

    MUSIC COMMUNITY INTRODUCES NEW LABELING PROGRAM TO DISTINGUISH GENERATIVE AI IN SOUND RECORDINGS

    aitoday7July 11, 2026
    AI News

    Safe from AI: which jobs will help you thrive in the future?

    aitoday7July 11, 2026

    Subscribe to Updates

    Get the latest tech news from FooBar about tech, design and biz.

    Most Popular

    Brief independent investigation of agents’ behavior, reasoning and collaboration in the OpenAI / Hugging Face hacking incident

    August 27, 20260 Views

    Fake US thinktank set up and funded by Israel sought to game AI for propaganda

    August 27, 20260 Views

    Medical Readiness Command, Europe G-6 Information Technology Team wins 2026 MEDCOM Mercury Award for HIT Team of the Year

    August 27, 20260 Views
    Our Picks

    OpenAI bets on families as ChatGPT goes deeper into households

    July 11, 2026

    MUSIC COMMUNITY INTRODUCES NEW LABELING PROGRAM TO DISTINGUISH GENERATIVE AI IN SOUND RECORDINGS

    July 11, 2026

    Safe from AI: which jobs will help you thrive in the future?

    July 11, 2026

    Subscribe to Updates

    Get the latest creative news from FooBar about art, design and business.

    Facebook X (Twitter) Instagram Pinterest
    • About Us
    • Get In Touch
    • Disclaimer
    • Privacy Policy
    • Terms and Conditions
    © 2026 AIToday7. All Rights Reserved.

    Type above and press Enter to search. Press Esc to cancel.