By using this site, you agree to the Privacy Policy and Terms of Use.
Accept
Online Tech Guru
  • News
  • PC/Windows
  • Mobile
  • Apps
  • Gadgets
  • More
    • Gaming
    • Accessories
    • Editor’s Choice
    • Press Release
Reading: It’s Frighteningly Easy to Jailbreak Some Frontier AI Models
Best Deal
Font ResizerAa
Online Tech GuruOnline Tech Guru
  • News
  • Mobile
  • PC/Windows
  • Gaming
  • Apps
  • Gadgets
  • Accessories
Search
  • News
  • PC/Windows
  • Mobile
  • Apps
  • Gadgets
  • More
    • Gaming
    • Accessories
    • Editor’s Choice
    • Press Release
UK Games Fund announces largest cohort of indie studios to receive grants of up to £100,000

UK Games Fund announces largest cohort of indie studios to receive grants of up to £100,000

News Room News Room 29 July 2026
FacebookLike
InstagramFollow
YoutubeSubscribe
TiktokFollow
  • Subscribe
  • Privacy Policy
  • Contact
  • Terms of Use
© Foxiz News Network. Ruby Design Company. All Rights Reserved.
Online Tech Guru > News > It’s Frighteningly Easy to Jailbreak Some Frontier AI Models
News

It’s Frighteningly Easy to Jailbreak Some Frontier AI Models

News Room
Last updated: 29 July 2026 19:41
By News Room 5 Min Read
Share
It’s Frighteningly Easy to Jailbreak Some Frontier AI Models
SHARE

I recently got to watch what happens when you jailbreak some of the world’s most powerful artificial intelligence models.

Don’t worry—this AI manipulation wasn’t used to hack anyone or build a nuclear bomb. I simply got to see firsthand how vulnerable some frontier models are to ditching their safety guardrails.

FAR.AI, an AI safety nonprofit based in California, built a tool that takes a range of problematic prompts, and generates more than a thousand different versions in an attempt to identify functioning jailbreaks. I saw some models generate a detailed plan for launching a cyberattack on an imaginary hydroelectric dam, among other things. Often, it involved trying dozens of prompts, with models rejecting many of them out of hand.

I chatted with FAR.AI in advance of a new report, which saw the group test the safety guardrails of models from four popular US companies: Anthropic’s Claude Opus 4.8 and Fable 5; OpenAI’s GPT 5.5 and 5.6; Google’s Gemini 3.1 Pro; and Grok 4.3 and 4.5, from Elon Musk’s newly combined SpaceXAI. It auto-generated prompts designed to trick the models into doing potentially harmful things, like generating software exploits and providing details for developing chemical or biological weapons.

The report found that Grok was most vulnerable to jailbreaks, with 448 jailbreaks found, followed by Gemini, with 249 found, while Claude, Fable, and GPT were impervious to the attacks. However, that doesn’t mean those models are immune to more sophisticated jailbreaks, which may involve interacting with a model in more complex ways, according to FAR.AI and other experts.

The report also calculated the cost of getting models to misbehave by using another AI model to automatically generate different jailbreaks. The results are dirt cheap, all things considered—$58 to jailbreak Grok and $278 to jailbreak Gemini.

“AI models right now are less regulated than restaurants,” says Adam Gleave, the CEO of FAR.AI and an expert on AI safety and alignment.

Gleave says that the findings demonstrate the need for externally imposed standards and regulations. “Talk of relying on voluntary commitments, that AI companies are going to be able to self-regulate, is nonsense,” he says.

But Gleave also believes that the findings show that models can be systematically tested for safety. “There’s an optimistic angle here,” he says. “Defense and safety really are possible.”

Rohin Shah, the director of AGI safety and alignment at Google DeepMind, says the results of the report “should not be interpreted as a comprehensive assessment of Gemini’s safety and security,” because not all jailbreaks are equally severe.

“We are constantly working to improve our safeguards,” Shah says. “We conduct extensive red teaming and evaluations across severe misuse risks and apply multiple layers of protection throughout development and deployment.”

“These findings reflect the sustained investment we’ve made in our safeguards,” Anthropic spokesperson Michael Aciman tells WIRED. “We continue to evolve our safety systems as these attacks become more sophisticated.”

OpenAI and SpaceXAI did not respond to WIRED’s request for comment.

Recently passed state laws in California and New York require frontier AI developers to publish safety reports, and soon, an Illinois law will require those companies to have their safety practices evaluated by third-party auditors. But the federal government hasn’t yet passed any specific safety requirements, and chaos has ensued as the industry—and officials—try to figure it out.

In June, the Trump administration imposed export controls on Anthropic’s Fable 5 and Mythos 5 models, citing national security concerns, and the company took them offline for several weeks. The White House has also asked both Anthropic and OpenAI to delay recent model releases over fears they could introduce new cybersecurity risks.

Share This Article
Facebook Twitter Copy Link
Leave a comment

Leave a Reply Cancel reply

Your email address will not be published. Required fields are marked *

What’s the catch with the Apple Upgrade program?

What’s the catch with the Apple Upgrade program?

News Room News Room 29 July 2026
FacebookLike
InstagramFollow
YoutubeSubscribe
TiktokFollow

Trending

OpenAI president says it’s ‘building a family of devices’ for its AI chatbots

In an interview with our friend Joanna Stern on her YouTube channel, OpenAI president Greg…

29 July 2026

Pokémon Pokopia Update Adds Dive Ability Next Week, Unlocking Underwater Gameplay

Pokémon Pokopia will unlock underwater gameplay next week as part of a free update, alongside…

29 July 2026

Apple Upgrade Isn’t the Best Way to Buy an iPhone

Apple has spent more than a decade tending to its thriving walled garden of products.…

29 July 2026
News

Which of Dyson’s 2026 Vacuum Models Is the Best?

Which of Dyson’s 2026 Vacuum Models Is the Best?

It's one of the most expensive for a reason, and comes with more attachments than any other model. I do find myself regularly grabbing the Fluffy Optic head for vacuuming…

News Room 29 July 2026

Your may also like!

Samsung’s Galaxy Z Fold 8 feels like the future
News

Samsung’s Galaxy Z Fold 8 feels like the future

News Room 29 July 2026
“People want to play arcade racers again” – How the developer of Wreckreation went from a company-wide redundancy notice to working on a sequel
Gaming

“People want to play arcade racers again” – How the developer of Wreckreation went from a company-wide redundancy notice to working on a sequel

News Room 29 July 2026
Tropical Diseases Like Dengue Fever and Chikungunya Are on the Rise in Europe
News

Tropical Diseases Like Dengue Fever and Chikungunya Are on the Rise in Europe

News Room 29 July 2026
The Ferrari Luce reportedly already hit its 2026 sales target
News

The Ferrari Luce reportedly already hit its 2026 sales target

News Room 29 July 2026

Our website stores cookies on your computer. They allow us to remember you and help personalize your experience with our site.

Read our privacy policy for more information.

Quick Links

  • Subscribe
  • Privacy Policy
  • Contact
  • Terms of Use
Advertise with us

Socials

Follow US
Welcome Back!

Sign in to your account

Lost your password?