Close Menu
StoryMoo – Global News & Trending Stories Hub

    Subscribe to Updates

    Get the latest creative news from FooBar about art, design and business.

    What's Hot

    11 Remote Entry-Level Jobs That Pay at Least $38 an Hour

    August 5, 2026

    Warrington Wolves and Wigan Warriors to play in Dublin in April 2027 as Super League hits the road | Rugby League News

    August 5, 2026

    Perez Hilton Seemed Happy, Excited About Life Weeks Before Alarming Incident

    August 5, 2026
    Facebook X (Twitter) Instagram
    Trending
    • 11 Remote Entry-Level Jobs That Pay at Least $38 an Hour
    • Warrington Wolves and Wigan Warriors to play in Dublin in April 2027 as Super League hits the road | Rugby League News
    • Perez Hilton Seemed Happy, Excited About Life Weeks Before Alarming Incident
    • Stephen Libby on fashion: so many men have asked me for clothes advice since The Traitors | Men’s fashion
    • ‘Tony’ Makes Bourdain Unlikeable, and That’s Why It Works
    • Disney+ looks to TikTok creators to bring fan content to its short-form video feed
    • Mohamed Salah lands in Turkiye to fan frenzy ahead of Trabzonspor move | Football News
    • Abdul El-Sayed likely winner in Michigan Democratic Senate primary
    Facebook X (Twitter) Instagram
    StoryMoo – Global News & Trending Stories Hub
    Subscribe
    Wednesday, August 5
    • Home
    • World News
    • Business
    • Health
    • Sports
    • Celebrities
    • Lifestyle
    • Travel & Tourism
    • Job post
    • Technology
    StoryMoo – Global News & Trending Stories Hub
    Home»Technology»OK, Well, Rogue AI Agents Are Hacking Again
    Technology

    OK, Well, Rogue AI Agents Are Hacking Again

    adminBy adminAugust 5, 2026No Comments4 Mins Read
    Facebook Twitter Pinterest LinkedIn Tumblr Email
    OK, Well, Rogue AI Agents Are Hacking Again
    Share
    Facebook Twitter LinkedIn Pinterest Email

    It’s officially getting hard to keep track of all the times and ways AI models from OpenAI and Anthropic have been involved in “security incidents,” going outside the confines of their testing and interacting with the wider internet in unintended, often unwelcome ways. Add these to the list: Agents from both AI labs went on recent, previously undisclosed hacking sprees, with one going so far as to leave instructions for future versions of itself.

    The most alarming behavior disclosed on Tuesday appears to have been tied to testing conducted by the UK’s AI Security Institute, which evaluates frontier models to identify potential issues before public release. AISI tests those models in “cyber ranges,” a simulated network in which AI agents are tasked with solving cybersecurity challenges, and intentionally disables safety features, including cybersecurity guardrails. In a recent bout of testing, models from both Anthropic and OpenAI took “autonomous, unsanctioned action on the live internet” a total of 19 times over 122 training runs.

    The institute attributed 17 unsanctioned actions to Anthropic’s Mythos 5 model and two to OpenAI’s GPT-5.6-Sol. In what the institute described as “the most serious case,” an AI agent attempted to insert malicious code into an open-source project on GitHub. It went so far as to create online personas “to pressure the project’s maintainer to approve the code,” according to AISI. Despite its elaborate attempts at social engineering, a human reviewer for the project ultimately rejected the pull request.

    Still, the agent went even further. “The agent tried to insert malicious instructions where it reasoned that other automated AI systems might pick them up and execute them,” AISI says, describing an attempt at prompt injection. One agent even left public messages on GitHub, offering to work with other agents to complete its task and giving a rundown of the work it had done so far. Subsequent agents found—and used—those instructions.

    AISI says it’s too soon to say whether the agents in question understood they had left the testing environment, or if they believed they were still within the boundaries of the simulation. Importantly, AISI does not test in a so-called sandbox environment; it allows agents access to the open internet during testing, in part so that they can access tools to accomplish their tasks. In this case, they did much more than that.

    In the other set of incidents detailed by OpenAI on Tuesday, a third-party AI security lab called Irregular mistakenly gave an unspecified OpenAI model access to the open internet. The model had been given an objective that was supposed to be completed in a sandbox environment, but thanks to a misconfiguration, it instead hacked a real website, using what OpenAI described as “a basic security vulnerability.” Not only that, but the model “found and used credentials to operate that same site.”

    It’s unclear what kind of site the OpenAI agent hacked, or what “operating” it might entail. Irregular did not respond to a request for comment.

    The latest discoveries follow several revelations from OpenAI last month, including the high-profile incident in which two of the company’s models hacked into servers of the AI evaluation and hosting startup Hugging Face—and four other organizations along the way—to steal the answers to a test they were being scored on. OpenAI’s disclosures prompted Anthropic to review its own testing. Last week, the Claude chatbot developer found that its models had gained unauthorized access to the computer systems of three different unnamed organizations.

    So far, the AI models have caused limited damage beyond allegedly violating some services’ terms of use and pointing to security lapses on the part of organizations they have breached. But the incidents have underscored the capabilities of AI models to find vulnerabilities across the internet and the dangers that await if they are allowed to operate with few restrictions. OpenAI called the Hugging Face situation “unprecedented,” but the pileup of breaches point to what cybersecurity experts have described as a clear pattern of human negligence and recklessness by the AI developers.

    Agents Hacking Rogue
    Share. Facebook Twitter Pinterest LinkedIn Tumblr Email
    admin
    • Website

    Related Posts

    Disney+ looks to TikTok creators to bring fan content to its short-form video feed

    August 5, 2026

    Nothing CMF is launching its first open earbuds

    August 4, 2026

    After killer quarter, Palantir CEO Alex Karp calls AI industry ‘Marxist’

    August 4, 2026
    Leave A Reply Cancel Reply

    Top Posts

    OPM cuts degree requirements for government tech jobs in new standards

    May 3, 20269 Views

    Weight loss drugs pose risk to pharma, report finds

    May 4, 20265 Views

    Grok Is Still Hosting Sexualized Deepfakes of Famous Women

    June 11, 20264 Views

    Chris Brown’s Ex-Housekeeper Fighting To Show Horrific Dog Attack Photos in Court

    May 1, 20264 Views
    Don't Miss
    Job post

    11 Remote Entry-Level Jobs That Pay at Least $38 an Hour

    By adminAugust 5, 20260

    Entry-level does not have to mean minimum wage or a cubicle. A set of roles…

    Warrington Wolves and Wigan Warriors to play in Dublin in April 2027 as Super League hits the road | Rugby League News

    August 5, 2026

    Perez Hilton Seemed Happy, Excited About Life Weeks Before Alarming Incident

    August 5, 2026

    Stephen Libby on fashion: so many men have asked me for clothes advice since The Traitors | Men’s fashion

    August 5, 2026
    Stay In Touch
    • Facebook
    • Twitter
    • Pinterest
    • Instagram
    • YouTube
    • Vimeo

    Subscribe to Updates

    Get the latest creative news from SmartMag about art & design.

    About Us

    Welcome to StoryMoo, your daily destination for the latest news, trending stories, and global updates from around the world.

    At StoryMoo, we bring together everything that matters in one place — from breaking world news and business insights to health updates, sports highlights, celebrity stories, lifestyle trends, travel inspiration, job updates, and the latest in technology.

    Facebook X (Twitter) Pinterest YouTube WhatsApp
    Our Picks

    11 Remote Entry-Level Jobs That Pay at Least $38 an Hour

    August 5, 2026

    Warrington Wolves and Wigan Warriors to play in Dublin in April 2027 as Super League hits the road | Rugby League News

    August 5, 2026

    Perez Hilton Seemed Happy, Excited About Life Weeks Before Alarming Incident

    August 5, 2026
    Most Popular

    Ukraine begins to flex muscle as an emerging air power, angering Russia | Russia-Ukraine war News

    May 1, 20260 Views

    Trump scraps Scotch whisky tariffs ‘in honor’ of King Charles

    May 1, 20260 Views

    Australia and Japan markets climb, looking past Iran war escalation fears

    May 1, 20260 Views
    Facebook X (Twitter) Instagram Pinterest
    • Terms & Conditions
    • Privacy Policy
    • Contact Us
    • About Us
    © 2026 StoryMoo. All Rights Reserved.

    Type above and press Enter to search. Press Esc to cancel.