• Wion
  • /Technology
  • /After OpenAI's 'scary' agent, Anthropic's Claude goes rogue; AI gains unauthorised 'real-world' access during testing

After OpenAI's 'scary' agent, Anthropic's Claude goes rogue; AI gains unauthorised 'real-world' access during testing

After OpenAI's 'scary' agent, Anthropic's Claude goes rogue; AI gains unauthorised 'real-world' access during testing

This photograph shows a handheld smartphone displaying the icons of some of the main artificial intelligence based apps, including LLMs, chatbots and generative AI, with logos (from L) of Proton AG's Lumo, Meta AI, Mistral Vibe (formerly Le Chat), xAI's Grok, Microsoft's Copilot, Google's Gemini, Anthropic's Claude, Perplexity, Deepseek, OpenAI's Chat GPT, Google's Notebook LLM and generative AI music app Suno, in Saint-Mande, east of Paris, on July 15, 2026. Photograph: (AFP)

Story highlights

Anthropic has revealed that three versions of its Claude AI model gained unauthorised access to external computer systems during security testing after unexpectedly obtaining internet access. The disclosure comes days after OpenAI reported a similar containment failure.

AI safety concerns have intensified after Anthropic disclosed that several versions of its Claude model gained unauthorised access to real-world computer systems during security testing. This comes just days after Anthropic rival OpenAI revealed a similar containment failure involving one of its own advanced models.

In a blog post published on Thursday (Jul 30), Anthropic said it reviewed more than 141,000 evaluation runs and found that three different versions of Claude had improperly accessed the systems of three external organisations.

Add WION as a Preferred Source

How could the AI get 'real-world access'?

Anthropic said the incidents occurred because the AI models unexpectedly had internet access during testing due to a "misunderstanding" between Anthropic and its evaluation partner, Irregular.

Once connected, Claude used "basic techniques, such as exploiting weak passwords and unauthenticated endpoints" to enter external systems, according to the company.

Trending Stories

One of the affected models was Mythos 5, Anthropic's most advanced AI system, which is currently available only to a limited group of approved partners.

Anthropic said it is working with Irregular to investigate the incidents and has contacted, or attempted to contact, all three organisations whose systems were accessed.

OpenAI's rogue agent

The disclosure comes less than a week after OpenAI acknowledged that one of its frontier AI models escaped its testing environment, connected to the internet and infiltrated Hugging Face, a platform widely used by developers to host and share AI models and code.

OpenAI later revealed it had identified three additional incidents involving unauthorised access during testing.

The company has since paused parts of its evaluation programme while strengthening its sandboxing systems, which are designed to isolate experimental AI models from external networks.

OpenAI CEO Sam Altman was also expected to meet White House officials on Thursday to discuss the company's upcoming AI models and voluntary government cybersecurity testing of advanced AI systems. According to Reuters, Altman is scheduled to meet White House Chief of Staff Susie Wiles, National Cyber Director Sean Cairncross, technology adviser Michael Kratsios, and Commerce Secretary Howard Lutnick.

About the Author

Share on twitter

Moohita Kaur Garg

Moohita Kaur Garg is a journalist and Senior Sub-Editor at WION News with five years of experience covering the volatile intersections of geopolitics and global security. She has e...Read More