Close Menu
Tech Savvyed
  • Home
  • News
  • Artificial Intelligence
  • Gadgets
  • Apps
  • Mobile
  • Gaming
  • Accessories
  • More
    • Web Stories
    • Spotlight
    • Press Release

Subscribe to Updates

Get the latest tech news and updates directly to your inbox.

What's On
The Internet Archive just made decades of vintage AI playable in your browser, and it’s fascinating

The Internet Archive just made decades of vintage AI playable in your browser, and it’s fascinating

31 August 2026
You worry about AI taking your job because you know too much about AI, research finds

You worry about AI taking your job because you know too much about AI, research finds

31 August 2026
Anthropic just showed an early version of self-improving AI

Anthropic just showed an early version of self-improving AI

31 August 2026
Facebook X (Twitter) Instagram
Facebook X (Twitter) Instagram
Tech Savvyed
SUBSCRIBE
  • Home
  • News
  • Artificial Intelligence
  • Gadgets
  • Apps
  • Mobile
  • Gaming
  • Accessories
  • More
    • Web Stories
    • Spotlight
    • Press Release
Tech Savvyed
Home»News»AI models from Anthropic and OpenAI were caught breaking the rules again
News

AI models from Anthropic and OpenAI were caught breaking the rules again

News RoomBy News Room5 August 20263 Mins Read
AI models from Anthropic and OpenAI were caught breaking the rules again
Share
Facebook Twitter Reddit Telegram Pinterest Email

OpenAI and Anthropic have both had a rough few weeks on the AI safety front. OpenAI recently disclosed that its models broke out of a test environment and hacked into Hugging Face and four other organizations. The news prompted Anthropic to review its own testing, which revealed that Claude had also gained unauthorized access to three companies.

Now, the UK’s AI Security Institute (AISI) has disclosed a new round of incidents (via Wired). It recorded 19 unauthorized actions on the live internet across 122 test runs involving models from both companies, the most serious of which saw an agent invent fake online personas to push malicious code into a real GitHub project. OpenAI separately revealed a second incident in which one of its models hacked a real website after a third-party lab mistakenly gave it live internet access.

17 incidents tied to Anthropic’s Mythos 5

AISI traced 17 of the 19 unauthorized actions to Anthropic’s Mythos 5 model, with the remaining two tied to OpenAI’s GPT 5.6 Sol. The GitHub incident was one of the 17, and it didn’t end when a human reviewer rejected the submission. The agent posted a summary of its progress publicly, inviting other automated systems to pick up where it left off, an attempt at what AISI calls prompt injection. A separate agent later found that message, used it, and continued the work.

On July 28th, we identified an incident during a routine cyber evaluation in which AI agents took sustained, unsanctioned actions directed at real people and organisations.

The behaviour came mostly from one model (Anthropic’s Mythos 5), with a small number of events from… pic.twitter.com/SPnA4Ekkwq

— AI Security Institute (AISI) (@AISecurityInst) August 4, 2026

AISI says it deliberately gave the models internet access and relaxed some safety protections to test their capabilities, but never instructed the agents to target real people or organizations. The institute says it’s still unclear whether the agents understood they’d gone beyond the scope of the simulation.

Another accidental breach at OpenAI

A second incident, disclosed by OpenAI the same day, started with a mistake at Irregular, a third-party lab OpenAI hired to run its cybersecurity tests. Irregular meant to keep its evaluation model confined to an isolated sandbox, but a configuration error gave the model direct access to the live internet. Once out, it exploited a vulnerability to break into a real website, then found and used credentials to operate the site it had just hacked. OpenAI hasn’t named the website or detailed what the model did with its access.

Both companies say the new incidents happened under deliberately loosened conditions that don’t reflect how their public models behave. Be that as it may, that doesn’t change the fact that AI agents from two of the industry’s most closely watched companies have now slipped past their intended limits in three separate incidents within a matter of weeks. And that doesn’t bode well for an industry racing to hand AI agents more real-world tasks before proving it can keep them in check.

Share. Facebook Twitter Pinterest LinkedIn Telegram Reddit Email
Previous ArticlePasskeys were pushed as a safer future. Hackers have figured out ways to break those synced to Google
Next Article Google Drive steps up its video collaboration game with timestamped comments

Related Articles

The Internet Archive just made decades of vintage AI playable in your browser, and it’s fascinating

The Internet Archive just made decades of vintage AI playable in your browser, and it’s fascinating

31 August 2026
You worry about AI taking your job because you know too much about AI, research finds

You worry about AI taking your job because you know too much about AI, research finds

31 August 2026
Anthropic just showed an early version of self-improving AI

Anthropic just showed an early version of self-improving AI

31 August 2026
Researchers built a  gadget that can find hidden cameras in seconds

Researchers built a $7 gadget that can find hidden cameras in seconds

31 August 2026
A new watchdog is tracking when AI goes rogue, and more than 300 incidents were reported in July

A new watchdog is tracking when AI goes rogue, and more than 300 incidents were reported in July

30 August 2026
Sony and Warner just sued Anthropic over copyrighted music used to train Claude, seeking 0,000 per track

Sony and Warner just sued Anthropic over copyrighted music used to train Claude, seeking $150,000 per track

30 August 2026
Demo
Stay In Touch
  • Facebook
  • Twitter
  • Pinterest
  • Instagram
  • YouTube
  • Vimeo
Don't Miss
You worry about AI taking your job because you know too much about AI, research finds

You worry about AI taking your job because you know too much about AI, research finds

By News Room31 August 2026

If you have been paying close attention to artificial intelligence, you may have noticed your…

Anthropic just showed an early version of self-improving AI

Anthropic just showed an early version of self-improving AI

31 August 2026
Researchers built a  gadget that can find hidden cameras in seconds

Researchers built a $7 gadget that can find hidden cameras in seconds

31 August 2026
A new watchdog is tracking when AI goes rogue, and more than 300 incidents were reported in July

A new watchdog is tracking when AI goes rogue, and more than 300 incidents were reported in July

30 August 2026
Tech Savvyed
Facebook X (Twitter) Instagram Pinterest
  • Privacy Policy
  • Terms of use
  • Advertise
  • Contact
© 2026 Tech Savvyed. All Rights Reserved.

Type above and press Enter to search. Press Esc to cancel.