By using this site, you agree to the Privacy Policy and Terms of Use.
Accept
Online Tech Guru
  • News
  • PC/Windows
  • Mobile
  • Apps
  • Gadgets
  • More
    • Gaming
    • Accessories
    • Editor’s Choice
    • Press Release
Reading: OK, Well, There Are Even More AI Agent Hacking Incidents
Best Deal
Font ResizerAa
Online Tech GuruOnline Tech Guru
  • News
  • Mobile
  • PC/Windows
  • Gaming
  • Apps
  • Gadgets
  • Accessories
Search
  • News
  • PC/Windows
  • Mobile
  • Apps
  • Gadgets
  • More
    • Gaming
    • Accessories
    • Editor’s Choice
    • Press Release
Samsung’s HDR10 Plus Advanced is launching this month on Prime Video

Samsung’s HDR10 Plus Advanced is launching this month on Prime Video

News Room News Room 5 August 2026
FacebookLike
InstagramFollow
YoutubeSubscribe
TiktokFollow
  • Subscribe
  • Privacy Policy
  • Contact
  • Terms of Use
© Foxiz News Network. Ruby Design Company. All Rights Reserved.
Online Tech Guru > News > OK, Well, There Are Even More AI Agent Hacking Incidents
News

OK, Well, There Are Even More AI Agent Hacking Incidents

News Room
Last updated: 5 August 2026 00:16
By News Room 5 Min Read
Share
OK, Well, There Are Even More AI Agent Hacking Incidents
SHARE

It’s officially getting hard to keep track of all the times and ways AI models from OpenAI and Anthropic have been involved in “security incidents,” going outside the confines of their testing and interacting with the wider internet in unintended, often unwelcome ways. Add these to the list: Agents from both AI labs went on recent, previously undisclosed hacking sprees, with one going so far as to leave instructions for future versions of itself.

The most alarming behavior disclosed on Tuesday appears to have been tied to testing conducted by the UK’s AI Security Institute, which evaluates frontier models to identify potential issues before public release. AISI tests those models in “cyber ranges,” a simulated network in which AI agents are tasked with solving cybersecurity challenges. In a recent bout of testing, models from both Anthropic and OpenAI took “autonomous, unsanctioned action on the live internet” a total of 19 times over 122 training runs.

The institute attributed 17 unsanctioned actions to Anthropic’s Mythos 5 model and two to OpenAI’s GPT-5.6-Sol. In what the institute described as “the most serious case,” an AI agent attempted to insert malicious code into an open-source project on GitHub. It went so far as to create online personas “to pressure the project’s maintainer to approve the code,” according to AISI. Despite its elaborate attempts at social engineering, a human reviewer for the project ultimately rejected the pull request.

Still, the agent went even further. “The agent tried to insert malicious instructions where it reasoned that other automated AI systems might pick them up and execute them,” AISI says, describing an attempt at prompt injection. One agent even left public messages on GitHub, offering to work with other agents to complete its task and giving a rundown of the work it had done so far. Subsequent agents found—and used—those instructions.

AISI says it’s too soon to say whether the agents in question understood they had left the testing environment, or if they believed they were still within the boundaries of the simulation. Importantly, AISI does not test in a so-called sandbox environment; it allows agents access to the open internet during testing, in part so that they can access tools to accomplish their tasks. In this case, they did much more than that.

In the other set of incidents detailed by OpenAI on Tuesday, a third-party AI security lab called Irregular mistakenly gave an unspecified OpenAI model access to the open internet. The model had been given an objective that was supposed to be completed in a sandbox environment, but thanks to a misconfiguration, it instead hacked a real website, using what OpenAI described as “a basic security vulnerability.” Not only that, but the model “found and used credentials to operate that same site.”

It’s unclear what kind of site the OpenAI agent hacked, or what “operating” it might entail. Irregular did not respond to a request for comment.

The latest discoveries follow several revelations from OpenAI last month, including the high-profile incident in which two of the company’s models hacked into servers of the AI evaluation and hosting startup Hugging Face—and four other organizations along the way—to steal the answers to a test they were being scored on. OpenAI’s disclosures prompted Anthropic to review its own testing. Last week, the Claude chatbot developer found that its models had gained unauthorized access to the computer systems of three different unnamed organizations.

So far, the AI models have caused limited damage beyond allegedly violating some services’ terms of use and pointing to security lapses on the part of organizations they have breached. But the incidents have underscored the capabilities of AI models to find vulnerabilities across the internet and the dangers that await if they are allowed to operate with few restrictions. OpenAI called the Hugging Face situation “unprecedented,” but the pileup of breaches point to what cybersecurity experts have described as a clear pattern of human negligence and recklessness by the AI developers.

Share This Article
Facebook Twitter Copy Link
Leave a comment

Leave a Reply Cancel reply

Your email address will not be published. Required fields are marked *

SpaceX made more revenue as an AI company than a space company

SpaceX made more revenue as an AI company than a space company

News Room News Room 5 August 2026
FacebookLike
InstagramFollow
YoutubeSubscribe
TiktokFollow

Trending

Naoki Hamaguchi Plans to Unite Final Fantasy VII Spin-Offs in Revelation’s Finale

Final Fantasy VII Revelation director Naoki Hamaguchi opened up on the game’s connection to the…

4 August 2026

The White House Is Keeping Its AI Cybersecurity Framework Secret

The Trump administration has finalized a plan to address the cybersecurity risks posed by increasingly…

4 August 2026

Now you can securely link multiple phones to one Signal account

You can link more devices with one phone number on Signal now, including an Android…

4 August 2026
Gaming

Leon Kennedy Gets Life-Sized Statue for Resident Evil’s 30th Anniversary and Fans Are Going Gaga

Leon Kennedy Gets Life-Sized Statue for Resident Evil’s 30th Anniversary and Fans Are Going Gaga

Leon Kennedy is coming to the real world… Well, sort of. Capcom is preparing for Resident Evil’s 30th anniversary exhibition in Shibuya, and to celebrate, it’s commissioning a life-sized statue…

News Room 5 August 2026

Your may also like!

AMD’s datacenter business is booming while gaming takes a backseat
News

AMD’s datacenter business is booming while gaming takes a backseat

News Room 4 August 2026
Humble Choice August 2026 Is Live
Gaming

Humble Choice August 2026 Is Live

News Room 4 August 2026
Telegram CEO says an extortionist planted CSAM in a chat to get it pulled from the App Store
News

Telegram CEO says an extortionist planted CSAM in a chat to get it pulled from the App Store

News Room 4 August 2026
World of Warcraft Crossover Is Up for Preorder
Gaming

World of Warcraft Crossover Is Up for Preorder

News Room 4 August 2026

Our website stores cookies on your computer. They allow us to remember you and help personalize your experience with our site.

Read our privacy policy for more information.

Quick Links

  • Subscribe
  • Privacy Policy
  • Contact
  • Terms of Use
Advertise with us

Socials

Follow US
Welcome Back!

Sign in to your account

Lost your password?