Close Menu
Tech Savvyed
  • Home
  • News
  • Artificial Intelligence
  • Gadgets
  • Apps
  • Mobile
  • Gaming
  • Accessories
  • More
    • Web Stories
    • Spotlight
    • Press Release

Subscribe to Updates

Get the latest tech news and updates directly to your inbox.

What's On
Heroes Of The Storm Gets First New Character In – Sep 12, 2026

Heroes Of The Storm Gets First New Character In – Sep 12, 2026

13 September 2026
Blizzard Announces Diablo V, Arriving Spring 2029

Blizzard Announces Diablo V, Arriving Spring 2029

12 September 2026
Forever – Sep 12, 2026

Forever – Sep 12, 2026

12 September 2026
Facebook X (Twitter) Instagram
Facebook X (Twitter) Instagram
Tech Savvyed
SUBSCRIBE
  • Home
  • News
  • Artificial Intelligence
  • Gadgets
  • Apps
  • Mobile
  • Gaming
  • Accessories
  • More
    • Web Stories
    • Spotlight
    • Press Release
Tech Savvyed
Home»News»Researchers expose a worryingly simple trick to make AI bots go rogue and skip safety
News

Researchers expose a worryingly simple trick to make AI bots go rogue and skip safety

News RoomBy News Room19 August 20262 Mins Read
Researchers expose a worryingly simple trick to make AI bots go rogue and skip safety
Share
Facebook Twitter Reddit Telegram Pinterest Email

If you ask an AI agent to hack an account, it will most certainly refuse, but researchers at EPFL just proved there is an easier way in, and it involves patience rather than technical skill. Their new study shows that breaking a harmful goal into small, harmless-sounding requests can trick AI agents into completing tasks they would normally reject outright (via TechXplore).

It echoes the recent ‘Bioshocking’ exploit in which AI browsers were manipulated into treating credential theft as part of a harmless game.

How researchers exposed this weakness

The team built an automated testing tool called STING, short for Sequential Testing of Illicit N-step Goal execution, designed to mimic how a real attacker would actually operate. Instead of stating a harmful goal directly, STING plans ahead and breaks that goal into a sequence of smaller, seemingly innocent steps that build toward it over multiple conversation turns.

Researchers tested this approach across 176 harmful scenarios against leading AI models, including ChatGPT, Gemini, and Claude. Each AI agent was tested as a tool-using agent, capable of browsing the web, sending emails, and completing multistep tasks.

Gradual, multistep manipulation succeeded far more often than blunt, single-prompt attempts. In some cases, models were twice as likely to complete a harmful task once the request was broken down into smaller steps. That finding tracks with separate research showing that even average users can talk their way past AI safety guardrails using nothing more than carefully worded prompts.

Why does this matter?

The concerns raised by researchers is not a hypothetical risk. Meta admitted in June that attackers used simple social engineering, not malware or hacking tools, to trick its AI support assistant into granting unauthorized access to Instagram accounts.

AI Chatbot

The researchers also expected attacks to be more effective in languages with less available training data. However, they found that completion rates stayed roughly consistent across all seven languages tested. They found one exception, though: switching languages midway through a multi-step attack made success rates jump significantly.

Lead researcher Ayush Kumar Tarun argues that safety testing needs to happen much earlier, built into an agent’s design from the start. Bolting it on after something goes wrong is no longer good enough, especially as these systems keep gaining more real-world capabilities.

Share. Facebook Twitter Pinterest LinkedIn Telegram Reddit Email
Previous ArticleI underestimated the Pixel 11, and now I’m eating my words
Next Article The Pixel 11 Pro made me fall in love with Night Sight once again

Related Articles

Soundcore Space 2 Pro headphones promise crystal-clear calls with Anker’s new AI chip

Soundcore Space 2 Pro headphones promise crystal-clear calls with Anker’s new AI chip

3 September 2026
Xbox Cloud Gaming is about to become a lot more appealing to casual gamers with a new pay-as-you-go option

Xbox Cloud Gaming is about to become a lot more appealing to casual gamers with a new pay-as-you-go option

3 September 2026
Lenovo reveals iMac rival that upstages Apple with more goodies at a lower price

Lenovo reveals iMac rival that upstages Apple with more goodies at a lower price

3 September 2026
Motorola’s new Edge 70 Plus packs a 200MP camera and a massive battery

Motorola’s new Edge 70 Plus packs a 200MP camera and a massive battery

3 September 2026
We got GTA VI limited-edition DualSense controllers before GTA VI

We got GTA VI limited-edition DualSense controllers before GTA VI

3 September 2026
Anker unveils smarter chargers and power banks built to fight heat and battery degradation

Anker unveils smarter chargers and power banks built to fight heat and battery degradation

3 September 2026
Demo
Stay In Touch
  • Facebook
  • Twitter
  • Pinterest
  • Instagram
  • YouTube
  • Vimeo
Don't Miss
Blizzard Announces Diablo V, Arriving Spring 2029

Blizzard Announces Diablo V, Arriving Spring 2029

By News Room12 September 2026

This morning during BlizzCon’s opening ceremony, Blizzard announced Diablo V, set a half-century after the…

Forever – Sep 12, 2026

Forever – Sep 12, 2026

12 September 2026
Everything Announced At The BlizzCon 2026 Opening Ceremony

Everything Announced At The BlizzCon 2026 Opening Ceremony

12 September 2026
Reforged Gets A New Campaign Today – Sep 12, 2026

Reforged Gets A New Campaign Today – Sep 12, 2026

12 September 2026
Tech Savvyed
Facebook X (Twitter) Instagram Pinterest
  • Privacy Policy
  • Terms of use
  • Advertise
  • Contact
© 2026 Tech Savvyed. All Rights Reserved.

Type above and press Enter to search. Press Esc to cancel.