Close Menu
Best in TechnologyBest in Technology
  • News
  • Phones
  • Laptops
  • Gadgets
  • Gaming
  • AI
  • Tips
  • More
    • Web Stories
    • Global
    • Press Release

Subscribe to Updates

Get the latest tech news and updates directly to your inbox.

What's On
X Says Australia’s Under-16 Social Media Ban Risks Interfering With Foreign Law

X Says Australia’s Under-16 Social Media Ban Risks Interfering With Foreign Law

30 July 2026
I didn’t think the budget Alienware 15 would work. I was wrong.

I didn’t think the budget Alienware 15 would work. I was wrong.

30 July 2026
Tropical Diseases Like Dengue Fever and Chikungunya Are on the Rise in Europe

Tropical Diseases Like Dengue Fever and Chikungunya Are on the Rise in Europe

30 July 2026
Facebook X (Twitter) Instagram
Just In
  • X Says Australia’s Under-16 Social Media Ban Risks Interfering With Foreign Law
  • I didn’t think the budget Alienware 15 would work. I was wrong.
  • Tropical Diseases Like Dengue Fever and Chikungunya Are on the Rise in Europe
  • Former Perplexity engineer launches new Polar AI browser which aims to automate repetitive tasks
  • It’s Frighteningly Easy to Jailbreak Some Frontier AI Models
  • I use ChatGPT Pro everyday, and these 3 features keep me subscribed
  • Save Up to $100 on Apple’s Newest AirPods (2026)
  • Pokémon Pokopia Bubbly Basin Expansion Gets Extended Gameplay Trailer And August 5 Release Date
Facebook X (Twitter) Instagram Pinterest Vimeo
Best in TechnologyBest in Technology
  • News
  • Phones
  • Laptops
  • Gadgets
  • Gaming
  • AI
  • Tips
  • More
    • Web Stories
    • Global
    • Press Release
Subscribe
Best in TechnologyBest in Technology
Home » It’s Frighteningly Easy to Jailbreak Some Frontier AI Models
News

It’s Frighteningly Easy to Jailbreak Some Frontier AI Models

News RoomBy News Room30 July 20263 Mins Read
Share Facebook Twitter Pinterest LinkedIn Tumblr Reddit Telegram Email
It’s Frighteningly Easy to Jailbreak Some Frontier AI Models
Share
Facebook Twitter LinkedIn Pinterest Email

I recently got to watch what happens when you jailbreak some of the world’s most powerful artificial intelligence models.

Don’t worry—this AI manipulation wasn’t used to hack anyone or build a nuclear bomb. I simply got to see firsthand how vulnerable some frontier models are to ditching their safety guardrails.

FAR.AI, an AI safety nonprofit based in California, built a tool that takes a range of problematic prompts and generates more than a thousand different versions in an attempt to identify functioning jailbreaks. I saw some models generate a detailed plan for launching a cyberattack on an imaginary hydroelectric dam, among other things. Often, it involved trying dozens of prompts, with models rejecting many of them out of hand.

I chatted with FAR.AI in advance of a new report, which saw the group test the safety guardrails of models from four popular US companies: Anthropic’s Claude Opus 4.8 and Fable 5; OpenAI’s GPT 5.5 and 5.6; Google’s Gemini 3.1 Pro; and Grok 4.3 and 4.5, from Elon Musk’s newly combined SpaceXAI. It auto-generated prompts designed to trick the models into doing potentially harmful things, like generating software exploits and providing details for developing chemical or biological weapons.

The report found that Grok was most vulnerable to jailbreaks, with 448 jailbreaks found, followed by Gemini, with 249 found, while Claude, Fable, and GPT were impervious to the attacks. However, that doesn’t mean those models are immune to more sophisticated jailbreaks, which may involve interacting with a model in more complex ways, according to FAR.AI and other experts.

The report also calculated the cost of getting models to misbehave by using another AI model to automatically generate different jailbreaks. The results are dirt cheap, all things considered—$58 to jailbreak Grok and $278 to jailbreak Gemini.

“AI models right now are less regulated than restaurants,” says Adam Gleave, the CEO of FAR.AI and an expert on AI safety and alignment.

Gleave says that the findings demonstrate the need for externally imposed standards and regulations. “Talk of relying on voluntary commitments, that AI companies are going to be able to self-regulate, is nonsense,” he says.

But Gleave also believes that the findings show that models can be systematically tested for safety. “There’s an optimistic angle here,” he says. “Defense and safety really are possible.”

Rohin Shah, the director of AGI safety and alignment at Google DeepMind, says the results of the report “should not be interpreted as a comprehensive assessment of Gemini’s safety and security,” because not all jailbreaks are equally severe.

“We are constantly working to improve our safeguards,” Shah says. “We conduct extensive red teaming and evaluations across severe misuse risks and apply multiple layers of protection throughout development and deployment.”

“These findings reflect the sustained investment we’ve made in our safeguards,” Anthropic spokesperson Michael Aciman tells WIRED. “We continue to evolve our safety systems as these attacks become more sophisticated.”

“Jailbreaks are an ongoing challenge across the industry, and we continuously strengthen our safeguards as attack techniques evolve. We rigorously test our models against new threats and use those findings to improve our protections,” OpenAI spokesperson Gaby Raila said in a statement to WIRED.

SpaceXAI did not respond to WIRED’s request for comment.

Recently passed state laws in California and New York require frontier AI developers to publish safety reports, and soon, an Illinois law will require those companies to have their safety practices evaluated by third-party auditors. But the federal government hasn’t yet passed any specific safety requirements, and chaos has ensued as the industry—and officials—try to figure it out.

Share. Facebook Twitter Pinterest LinkedIn Tumblr Email
Previous ArticleI use ChatGPT Pro everyday, and these 3 features keep me subscribed
Next Article Former Perplexity engineer launches new Polar AI browser which aims to automate repetitive tasks

Related Articles

X Says Australia’s Under-16 Social Media Ban Risks Interfering With Foreign Law
News

X Says Australia’s Under-16 Social Media Ban Risks Interfering With Foreign Law

30 July 2026
I didn’t think the budget Alienware 15 would work. I was wrong.
News

I didn’t think the budget Alienware 15 would work. I was wrong.

30 July 2026
Tropical Diseases Like Dengue Fever and Chikungunya Are on the Rise in Europe
News

Tropical Diseases Like Dengue Fever and Chikungunya Are on the Rise in Europe

30 July 2026
Former Perplexity engineer launches new Polar AI browser which aims to automate repetitive tasks
News

Former Perplexity engineer launches new Polar AI browser which aims to automate repetitive tasks

30 July 2026
I use ChatGPT Pro everyday, and these 3 features keep me subscribed
News

I use ChatGPT Pro everyday, and these 3 features keep me subscribed

30 July 2026
Save Up to 0 on Apple’s Newest AirPods (2026)
News

Save Up to $100 on Apple’s Newest AirPods (2026)

30 July 2026
Demo
Top Articles
5 laptops to buy instead of the M4 MacBook Pro

5 laptops to buy instead of the M4 MacBook Pro

17 November 2024133 Views
ChatGPT o1 vs. o1-mini vs. 4o: Which should you use?

ChatGPT o1 vs. o1-mini vs. 4o: Which should you use?

15 December 2024111 Views
Costco partners with Electric Era to bring back EV charging in the U.S.

Costco partners with Electric Era to bring back EV charging in the U.S.

28 October 2024100 Views

Subscribe to Updates

Get the latest tech news and updates directly to your inbox.

Latest News
I use ChatGPT Pro everyday, and these 3 features keep me subscribed News

I use ChatGPT Pro everyday, and these 3 features keep me subscribed

News Room30 July 2026
Save Up to 0 on Apple’s Newest AirPods (2026) News

Save Up to $100 on Apple’s Newest AirPods (2026)

News Room30 July 2026
Pokémon Pokopia Bubbly Basin Expansion Gets Extended Gameplay Trailer And August 5 Release Date Gaming

Pokémon Pokopia Bubbly Basin Expansion Gets Extended Gameplay Trailer And August 5 Release Date

News Room30 July 2026
Most Popular
The Spectacular Burnout of a Solar Panel Salesman

The Spectacular Burnout of a Solar Panel Salesman

13 January 2025137 Views
5 laptops to buy instead of the M4 MacBook Pro

5 laptops to buy instead of the M4 MacBook Pro

17 November 2024133 Views
ChatGPT o1 vs. o1-mini vs. 4o: Which should you use?

ChatGPT o1 vs. o1-mini vs. 4o: Which should you use?

15 December 2024111 Views
Our Picks
Former Perplexity engineer launches new Polar AI browser which aims to automate repetitive tasks

Former Perplexity engineer launches new Polar AI browser which aims to automate repetitive tasks

30 July 2026
It’s Frighteningly Easy to Jailbreak Some Frontier AI Models

It’s Frighteningly Easy to Jailbreak Some Frontier AI Models

30 July 2026
I use ChatGPT Pro everyday, and these 3 features keep me subscribed

I use ChatGPT Pro everyday, and these 3 features keep me subscribed

30 July 2026

Subscribe to Updates

Get the latest tech news and updates directly to your inbox.

Facebook X (Twitter) Instagram Pinterest
  • Privacy Policy
  • Terms of use
  • Advertise
  • Contact Us
© 2026 Best in Technology. All Rights Reserved.

Type above and press Enter to search. Press Esc to cancel.