Close Menu
Best in TechnologyBest in Technology
  • News
  • Phones
  • Laptops
  • Gadgets
  • Gaming
  • AI
  • Tips
  • More
    • Web Stories
    • Global
    • Press Release

Subscribe to Updates

Get the latest tech news and updates directly to your inbox.

What's On
Researchers expose a worryingly simple trick to make AI bots go rogue and skip safety

Researchers expose a worryingly simple trick to make AI bots go rogue and skip safety

19 August 2026
Review: Google Pixel 11 Pro and Pixel 11 Pro XL

Review: Google Pixel 11 Pro and Pixel 11 Pro XL

19 August 2026
I underestimated the Pixel 11, and now I’m eating my words

I underestimated the Pixel 11, and now I’m eating my words

19 August 2026
Facebook X (Twitter) Instagram
Just In
  • Researchers expose a worryingly simple trick to make AI bots go rogue and skip safety
  • Review: Google Pixel 11 Pro and Pixel 11 Pro XL
  • I underestimated the Pixel 11, and now I’m eating my words
  • Review: Google Pixel Watch 5
  • Amazon is adding Alexa Plus to Fire TV devices without charging extra
  • Coders Say They Already Found Workarounds to Claude’s Invisible Watermarks
  • We Played Tides Of Annihilation For Two Hours And – Aug 19, 2026
  • [Update] Googlebook finally gets a launch date, and it’s sooner than you would expect
Facebook X (Twitter) Instagram Pinterest Vimeo
Best in TechnologyBest in Technology
  • News
  • Phones
  • Laptops
  • Gadgets
  • Gaming
  • AI
  • Tips
  • More
    • Web Stories
    • Global
    • Press Release
Subscribe
Best in TechnologyBest in Technology
Home » Researchers expose a worryingly simple trick to make AI bots go rogue and skip safety
News

Researchers expose a worryingly simple trick to make AI bots go rogue and skip safety

News RoomBy News Room19 August 20262 Mins Read
Share Facebook Twitter Pinterest LinkedIn Tumblr Reddit Telegram Email
Researchers expose a worryingly simple trick to make AI bots go rogue and skip safety
Share
Facebook Twitter LinkedIn Pinterest Email

If you ask an AI agent to hack an account, it will most certainly refuse, but researchers at EPFL just proved there is an easier way in, and it involves patience rather than technical skill. Their new study shows that breaking a harmful goal into small, harmless-sounding requests can trick AI agents into completing tasks they would normally reject outright (via TechXplore).

It echoes the recent ‘Bioshocking’ exploit in which AI browsers were manipulated into treating credential theft as part of a harmless game.

How researchers exposed this weakness

The team built an automated testing tool called STING, short for Sequential Testing of Illicit N-step Goal execution, designed to mimic how a real attacker would actually operate. Instead of stating a harmful goal directly, STING plans ahead and breaks that goal into a sequence of smaller, seemingly innocent steps that build toward it over multiple conversation turns.

Researchers tested this approach across 176 harmful scenarios against leading AI models, including ChatGPT, Gemini, and Claude. Each AI agent was tested as a tool-using agent, capable of browsing the web, sending emails, and completing multistep tasks.

Gradual, multistep manipulation succeeded far more often than blunt, single-prompt attempts. In some cases, models were twice as likely to complete a harmful task once the request was broken down into smaller steps. That finding tracks with separate research showing that even average users can talk their way past AI safety guardrails using nothing more than carefully worded prompts.

Why does this matter?

The concerns raised by researchers is not a hypothetical risk. Meta admitted in June that attackers used simple social engineering, not malware or hacking tools, to trick its AI support assistant into granting unauthorized access to Instagram accounts.

AI Chatbot

The researchers also expected attacks to be more effective in languages with less available training data. However, they found that completion rates stayed roughly consistent across all seven languages tested. They found one exception, though: switching languages midway through a multi-step attack made success rates jump significantly.

Lead researcher Ayush Kumar Tarun argues that safety testing needs to happen much earlier, built into an agent’s design from the start. Bolting it on after something goes wrong is no longer good enough, especially as these systems keep gaining more real-world capabilities.

Share. Facebook Twitter Pinterest LinkedIn Tumblr Email
Previous ArticleReview: Google Pixel 11 Pro and Pixel 11 Pro XL

Related Articles

Review: Google Pixel 11 Pro and Pixel 11 Pro XL
News

Review: Google Pixel 11 Pro and Pixel 11 Pro XL

19 August 2026
I underestimated the Pixel 11, and now I’m eating my words
News

I underestimated the Pixel 11, and now I’m eating my words

19 August 2026
Review: Google Pixel Watch 5
News

Review: Google Pixel Watch 5

19 August 2026
Amazon is adding Alexa Plus to Fire TV devices without charging extra
News

Amazon is adding Alexa Plus to Fire TV devices without charging extra

19 August 2026
Coders Say They Already Found Workarounds to Claude’s Invisible Watermarks
News

Coders Say They Already Found Workarounds to Claude’s Invisible Watermarks

19 August 2026
[Update] Googlebook finally gets a launch date, and it’s sooner than you would expect
News

[Update] Googlebook finally gets a launch date, and it’s sooner than you would expect

19 August 2026
Demo
Top Articles
5 laptops to buy instead of the M4 MacBook Pro

5 laptops to buy instead of the M4 MacBook Pro

17 November 2024133 Views
ChatGPT o1 vs. o1-mini vs. 4o: Which should you use?

ChatGPT o1 vs. o1-mini vs. 4o: Which should you use?

15 December 2024112 Views
Costco partners with Electric Era to bring back EV charging in the U.S.

Costco partners with Electric Era to bring back EV charging in the U.S.

28 October 2024100 Views

Subscribe to Updates

Get the latest tech news and updates directly to your inbox.

Latest News
Coders Say They Already Found Workarounds to Claude’s Invisible Watermarks News

Coders Say They Already Found Workarounds to Claude’s Invisible Watermarks

News Room19 August 2026
We Played Tides Of Annihilation For Two Hours And – Aug 19, 2026 Gaming

We Played Tides Of Annihilation For Two Hours And – Aug 19, 2026

News Room19 August 2026
[Update] Googlebook finally gets a launch date, and it’s sooner than you would expect News

[Update] Googlebook finally gets a launch date, and it’s sooner than you would expect

News Room19 August 2026
Most Popular
The Spectacular Burnout of a Solar Panel Salesman

The Spectacular Burnout of a Solar Panel Salesman

13 January 2025137 Views
5 laptops to buy instead of the M4 MacBook Pro

5 laptops to buy instead of the M4 MacBook Pro

17 November 2024133 Views
ChatGPT o1 vs. o1-mini vs. 4o: Which should you use?

ChatGPT o1 vs. o1-mini vs. 4o: Which should you use?

15 December 2024112 Views
Our Picks
Review: Google Pixel Watch 5

Review: Google Pixel Watch 5

19 August 2026
Amazon is adding Alexa Plus to Fire TV devices without charging extra

Amazon is adding Alexa Plus to Fire TV devices without charging extra

19 August 2026
Coders Say They Already Found Workarounds to Claude’s Invisible Watermarks

Coders Say They Already Found Workarounds to Claude’s Invisible Watermarks

19 August 2026

Subscribe to Updates

Get the latest tech news and updates directly to your inbox.

Facebook X (Twitter) Instagram Pinterest
  • Privacy Policy
  • Terms of use
  • Advertise
  • Contact Us
© 2026 Best in Technology. All Rights Reserved.

Type above and press Enter to search. Press Esc to cancel.