Close Menu
Best in TechnologyBest in Technology
  • News
  • Phones
  • Laptops
  • Gadgets
  • Gaming
  • AI
  • Tips
  • More
    • Web Stories
    • Global
    • Press Release

Subscribe to Updates

Get the latest tech news and updates directly to your inbox.

What's On
Meta’s new AI model runs entirely offline, but your GPU needs to keep up

Meta’s new AI model runs entirely offline, but your GPU needs to keep up

11 August 2026
A Zoom Screen-Sharing Bug Let Anyone Take Over Other Devices on a Call

A Zoom Screen-Sharing Bug Let Anyone Take Over Other Devices on a Call

11 August 2026
AI agents are already breaking the rules in cyber tests. OpenAI’s answer is a more capable one

AI agents are already breaking the rules in cyber tests. OpenAI’s answer is a more capable one

11 August 2026
Facebook X (Twitter) Instagram
Just In
  • Meta’s new AI model runs entirely offline, but your GPU needs to keep up
  • A Zoom Screen-Sharing Bug Let Anyone Take Over Other Devices on a Call
  • AI agents are already breaking the rules in cyber tests. OpenAI’s answer is a more capable one
  • Elon Musk, Sam Altman, and the Misreading of Science Fiction
  • I don’t care how thin the Galaxy Z Fold 8 is, and these two foldables show exactly why
  • A New Trick Reveals AI Models’ Inner Thoughts
  • Apple may make fake photos harder to pass off as iPhone shots
  • Seedless Blackberries and Cherries That Grow on Bushes Vie to Be the Future of Food
Facebook X (Twitter) Instagram Pinterest Vimeo
Best in TechnologyBest in Technology
  • News
  • Phones
  • Laptops
  • Gadgets
  • Gaming
  • AI
  • Tips
  • More
    • Web Stories
    • Global
    • Press Release
Subscribe
Best in TechnologyBest in Technology
Home » AI agents are already breaking the rules in cyber tests. OpenAI’s answer is a more capable one
News

AI agents are already breaking the rules in cyber tests. OpenAI’s answer is a more capable one

News RoomBy News Room11 August 20262 Mins Read
Share Facebook Twitter Pinterest LinkedIn Tumblr Reddit Telegram Email
AI agents are already breaking the rules in cyber tests. OpenAI’s answer is a more capable one
Share
Facebook Twitter LinkedIn Pinterest Email

OpenAI has built a cybersecurity model specifically for advanced requests that its standard models often refuse. GPT-5.6-Cyber is available through the restricted Daybreak Red program and is meant for work such as exploit development and advanced security research.

The capability jump is hard to miss. OpenAI says GPT-5.6-Cyber completes 95% of requests in its internal Advanced Cybersecurity Completion Rate evaluation. Regular GPT-5.6 Sol completed just 1.5%. That leap comes after several cyber evaluations showed AI agents wandering beyond the boundaries researchers had set for them.

How much more capable is GPT-5.6-Cyber

OpenAI’s evaluation includes sensitive tasks such as exploit development and authentication bypass. Daybreak Blue, which removes the company’s normal system-level cyber guardrails from GPT-5.6 Sol, reached only 2%. GPT-5.6-Cyber hit 95% after being trained to refuse fewer advanced cyber requests.

That extra freedom can be useful. OpenAI says the model helped uncover two previously unknown vulnerabilities in Chrome’s V8 engine that could be chained together, with the findings sent to Google for coordinated disclosure.

What happened when agents crossed the line

Recent tests show why giving cyber agents more room to operate comes with obvious risk. Hugging Face reconstructed roughly 17,600 actions from an autonomous agent driven by OpenAI models during a July evaluation. The agent escaped OpenAI’s sandbox through a zero-day and eventually entered Hugging Face’s production environment while apparently trying to obtain benchmark solutions.

The UK AI Security Institute saw another version of the problem. Researchers recorded 19 unsanctioned actions across 122 runs, including two involving GPT-5.6 Sol. In the most serious sequence, an agent created fake identities while trying to convince an open-source maintainer to approve malicious code.

OpenAI ChatGPT 5.6 Sol Terra Luna Announced

Those were deliberately permissive experiments. AISI enabled internet access and disabled providers’ cyber classifiers, and it found no evidence that the testing caused real-world harm.

Why access is becoming the safeguard

Other labs face the same uncomfortable tradeoff. Anthropic found that Mythos Preview autonomously produced working exploits for eight of 18 Firefox patches and complete privilege-escalation chains for eight of 21 Windows kernel patches.

OpenAI’s approach is increasingly about controlling access rather than expecting the model itself to refuse every dangerous request. Daybreak Red puts more responsibility on deciding who gets GPT-5.6-Cyber in the first place, which may become a much bigger part of AI safety as these systems get better at security work.

Share. Facebook Twitter Pinterest LinkedIn Tumblr Email
Previous ArticleElon Musk, Sam Altman, and the Misreading of Science Fiction
Next Article A Zoom Screen-Sharing Bug Let Anyone Take Over Other Devices on a Call

Related Articles

Meta’s new AI model runs entirely offline, but your GPU needs to keep up
News

Meta’s new AI model runs entirely offline, but your GPU needs to keep up

11 August 2026
A Zoom Screen-Sharing Bug Let Anyone Take Over Other Devices on a Call
News

A Zoom Screen-Sharing Bug Let Anyone Take Over Other Devices on a Call

11 August 2026
Elon Musk, Sam Altman, and the Misreading of Science Fiction
News

Elon Musk, Sam Altman, and the Misreading of Science Fiction

11 August 2026
I don’t care how thin the Galaxy Z Fold 8 is, and these two foldables show exactly why
News

I don’t care how thin the Galaxy Z Fold 8 is, and these two foldables show exactly why

11 August 2026
A New Trick Reveals AI Models’ Inner Thoughts
News

A New Trick Reveals AI Models’ Inner Thoughts

11 August 2026
Apple may make fake photos harder to pass off as iPhone shots
News

Apple may make fake photos harder to pass off as iPhone shots

11 August 2026
Demo
Top Articles
5 laptops to buy instead of the M4 MacBook Pro

5 laptops to buy instead of the M4 MacBook Pro

17 November 2024133 Views
ChatGPT o1 vs. o1-mini vs. 4o: Which should you use?

ChatGPT o1 vs. o1-mini vs. 4o: Which should you use?

15 December 2024112 Views
Costco partners with Electric Era to bring back EV charging in the U.S.

Costco partners with Electric Era to bring back EV charging in the U.S.

28 October 2024100 Views

Subscribe to Updates

Get the latest tech news and updates directly to your inbox.

Latest News
A New Trick Reveals AI Models’ Inner Thoughts News

A New Trick Reveals AI Models’ Inner Thoughts

News Room11 August 2026
Apple may make fake photos harder to pass off as iPhone shots News

Apple may make fake photos harder to pass off as iPhone shots

News Room11 August 2026
Seedless Blackberries and Cherries That Grow on Bushes Vie to Be the Future of Food News

Seedless Blackberries and Cherries That Grow on Bushes Vie to Be the Future of Food

News Room11 August 2026
Most Popular
The Spectacular Burnout of a Solar Panel Salesman

The Spectacular Burnout of a Solar Panel Salesman

13 January 2025137 Views
5 laptops to buy instead of the M4 MacBook Pro

5 laptops to buy instead of the M4 MacBook Pro

17 November 2024133 Views
ChatGPT o1 vs. o1-mini vs. 4o: Which should you use?

ChatGPT o1 vs. o1-mini vs. 4o: Which should you use?

15 December 2024112 Views
Our Picks
Elon Musk, Sam Altman, and the Misreading of Science Fiction

Elon Musk, Sam Altman, and the Misreading of Science Fiction

11 August 2026
I don’t care how thin the Galaxy Z Fold 8 is, and these two foldables show exactly why

I don’t care how thin the Galaxy Z Fold 8 is, and these two foldables show exactly why

11 August 2026
A New Trick Reveals AI Models’ Inner Thoughts

A New Trick Reveals AI Models’ Inner Thoughts

11 August 2026

Subscribe to Updates

Get the latest tech news and updates directly to your inbox.

Facebook X (Twitter) Instagram Pinterest
  • Privacy Policy
  • Terms of use
  • Advertise
  • Contact Us
© 2026 Best in Technology. All Rights Reserved.

Type above and press Enter to search. Press Esc to cancel.