Close Menu
Best in TechnologyBest in Technology
  • News
  • Phones
  • Laptops
  • Gadgets
  • Gaming
  • AI
  • Tips
  • More
    • Web Stories
    • Global
    • Press Release

Subscribe to Updates

Get the latest tech news and updates directly to your inbox.

What's On
Google Health adds a faster way to share your medical history with doctors

Google Health adds a faster way to share your medical history with doctors

5 August 2026
Apple will finally stop making iPhone-to-Windows copy-paste such a chore

Apple will finally stop making iPhone-to-Windows copy-paste such a chore

5 August 2026
OK, Well, Rogue AI Agents Are Hacking Again

OK, Well, Rogue AI Agents Are Hacking Again

5 August 2026
Facebook X (Twitter) Instagram
Just In
  • Google Health adds a faster way to share your medical history with doctors
  • Apple will finally stop making iPhone-to-Windows copy-paste such a chore
  • OK, Well, Rogue AI Agents Are Hacking Again
  • TCL just refreshed the Tab A1 Plus with its NXTPAPER tech and a bigger battery
  • The White House Is Keeping Its AI Cybersecurity Framework Secret
  • Electronic Arts Has Been Fully Acquired By Trio Of Investor
  • ChatGPT Pro replaced Gemini Notebook as my favorite research app, and I didn’t see it coming
  • The Pixel Watch 5 just leaked in full before Google’s August 12 event
Facebook X (Twitter) Instagram Pinterest Vimeo
Best in TechnologyBest in Technology
  • News
  • Phones
  • Laptops
  • Gadgets
  • Gaming
  • AI
  • Tips
  • More
    • Web Stories
    • Global
    • Press Release
Subscribe
Best in TechnologyBest in Technology
Home » OK, Well, Rogue AI Agents Are Hacking Again
News

OK, Well, Rogue AI Agents Are Hacking Again

News RoomBy News Room5 August 20264 Mins Read
Share Facebook Twitter Pinterest LinkedIn Tumblr Reddit Telegram Email
OK, Well, Rogue AI Agents Are Hacking Again
Share
Facebook Twitter LinkedIn Pinterest Email

It’s officially getting hard to keep track of all the times and ways AI models from OpenAI and Anthropic have been involved in “security incidents,” going outside the confines of their testing and interacting with the wider internet in unintended, often unwelcome ways. Add these to the list: Agents from both AI labs went on recent, previously undisclosed hacking sprees, with one going so far as to leave instructions for future versions of itself.

The most alarming behavior disclosed on Tuesday appears to have been tied to testing conducted by the UK’s AI Security Institute, which evaluates frontier models to identify potential issues before public release. AISI tests those models in “cyber ranges,” a simulated network in which AI agents are tasked with solving cybersecurity challenges. In a recent bout of testing, models from both Anthropic and OpenAI took “autonomous, unsanctioned action on the live internet” a total of 19 times over 122 training runs.

The institute attributed 17 unsanctioned actions to Anthropic’s Mythos 5 model and two to OpenAI’s GPT-5.6-Sol. In what the institute described as “the most serious case,” an AI agent attempted to insert malicious code into an open-source project on GitHub. It went so far as to create online personas “to pressure the project’s maintainer to approve the code,” according to AISI. Despite its elaborate attempts at social engineering, a human reviewer for the project ultimately rejected the pull request.

Still, the agent went even further. “The agent tried to insert malicious instructions where it reasoned that other automated AI systems might pick them up and execute them,” AISI says, describing an attempt at prompt injection. One agent even left public messages on GitHub, offering to work with other agents to complete its task and giving a rundown of the work it had done so far. Subsequent agents found—and used—those instructions.

AISI says it’s too soon to say whether the agents in question understood they had left the testing environment, or if they believed they were still within the boundaries of the simulation. Importantly, AISI does not test in a so-called sandbox environment; it allows agents access to the open internet during testing, in part so that they can access tools to accomplish their tasks. In this case, they did much more than that.

In the other set of incidents detailed by OpenAI on Tuesday, a third-party AI security lab called Irregular mistakenly gave an unspecified OpenAI model access to the open internet. The model had been given an objective that was supposed to be completed in a sandbox environment, but thanks to a misconfiguration, it instead hacked a real website, using what OpenAI described as “a basic security vulnerability.” Not only that, but the model “found and used credentials to operate that same site.”

It’s unclear what kind of site the OpenAI agent hacked, or what “operating” it might entail. Irregular did not respond to a request for comment.

The latest discoveries follow several revelations from OpenAI last month, including the high-profile incident in which two of the company’s models hacked into servers of the AI evaluation and hosting startup Hugging Face—and four other organizations along the way—to steal the answers to a test they were being scored on. OpenAI’s disclosures prompted Anthropic to review its own testing. Last week, the Claude chatbot developer found that its models had gained unauthorized access to the computer systems of three different unnamed organizations.

So far, the AI models have caused limited damage beyond allegedly violating some services’ terms of use and pointing to security lapses on the part of organizations they have breached. But the incidents have underscored the capabilities of AI models to find vulnerabilities across the internet and the dangers that await if they are allowed to operate with few restrictions. OpenAI called the Hugging Face situation “unprecedented,” but the pileup of breaches point to what cybersecurity experts have described as a clear pattern of human negligence and recklessness by the AI developers.

Share. Facebook Twitter Pinterest LinkedIn Tumblr Email
Previous ArticleTCL just refreshed the Tab A1 Plus with its NXTPAPER tech and a bigger battery
Next Article Apple will finally stop making iPhone-to-Windows copy-paste such a chore

Related Articles

Google Health adds a faster way to share your medical history with doctors
News

Google Health adds a faster way to share your medical history with doctors

5 August 2026
Apple will finally stop making iPhone-to-Windows copy-paste such a chore
News

Apple will finally stop making iPhone-to-Windows copy-paste such a chore

5 August 2026
TCL just refreshed the Tab A1 Plus with its NXTPAPER tech and a bigger battery
News

TCL just refreshed the Tab A1 Plus with its NXTPAPER tech and a bigger battery

5 August 2026
The White House Is Keeping Its AI Cybersecurity Framework Secret
News

The White House Is Keeping Its AI Cybersecurity Framework Secret

5 August 2026
ChatGPT Pro replaced Gemini Notebook as my favorite research app, and I didn’t see it coming
News

ChatGPT Pro replaced Gemini Notebook as my favorite research app, and I didn’t see it coming

4 August 2026
The Pixel Watch 5 just leaked in full before Google’s August 12 event
News

The Pixel Watch 5 just leaked in full before Google’s August 12 event

4 August 2026
Demo
Top Articles
5 laptops to buy instead of the M4 MacBook Pro

5 laptops to buy instead of the M4 MacBook Pro

17 November 2024133 Views
ChatGPT o1 vs. o1-mini vs. 4o: Which should you use?

ChatGPT o1 vs. o1-mini vs. 4o: Which should you use?

15 December 2024111 Views
Costco partners with Electric Era to bring back EV charging in the U.S.

Costco partners with Electric Era to bring back EV charging in the U.S.

28 October 2024100 Views

Subscribe to Updates

Get the latest tech news and updates directly to your inbox.

Latest News
Electronic Arts Has Been Fully Acquired By Trio Of Investor Gaming

Electronic Arts Has Been Fully Acquired By Trio Of Investor

News Room5 August 2026
ChatGPT Pro replaced Gemini Notebook as my favorite research app, and I didn’t see it coming News

ChatGPT Pro replaced Gemini Notebook as my favorite research app, and I didn’t see it coming

News Room4 August 2026
The Pixel Watch 5 just leaked in full before Google’s August 12 event News

The Pixel Watch 5 just leaked in full before Google’s August 12 event

News Room4 August 2026
Most Popular
The Spectacular Burnout of a Solar Panel Salesman

The Spectacular Burnout of a Solar Panel Salesman

13 January 2025137 Views
5 laptops to buy instead of the M4 MacBook Pro

5 laptops to buy instead of the M4 MacBook Pro

17 November 2024133 Views
ChatGPT o1 vs. o1-mini vs. 4o: Which should you use?

ChatGPT o1 vs. o1-mini vs. 4o: Which should you use?

15 December 2024111 Views
Our Picks
TCL just refreshed the Tab A1 Plus with its NXTPAPER tech and a bigger battery

TCL just refreshed the Tab A1 Plus with its NXTPAPER tech and a bigger battery

5 August 2026
The White House Is Keeping Its AI Cybersecurity Framework Secret

The White House Is Keeping Its AI Cybersecurity Framework Secret

5 August 2026
Electronic Arts Has Been Fully Acquired By Trio Of Investor

Electronic Arts Has Been Fully Acquired By Trio Of Investor

5 August 2026

Subscribe to Updates

Get the latest tech news and updates directly to your inbox.

Facebook X (Twitter) Instagram Pinterest
  • Privacy Policy
  • Terms of use
  • Advertise
  • Contact Us
© 2026 Best in Technology. All Rights Reserved.

Type above and press Enter to search. Press Esc to cancel.