• Wion
  • /World
  • /Researchers hack OpenAI systems, access employee credentials using Anthropic's AI security tool

Researchers hack OpenAI systems, access employee credentials using Anthropic's AI security tool

Researchers hack OpenAI systems, access employee credentials using Anthropic's AI security tool

Representative image. Photograph: (Others)

Story highlights

The researchers had been granted access to a cybersecurity-focused tool built by Anthropic. The move was part of a broader initiative aiming to hunt down weaknesses before bad actors could exploit them.

Cybersecurity researchers reportedly hacked OpenAI’s systems using security tools developed by rival artificial intelligence company Anthropic. The development followed two weeksafter a swarm of AI agents escaped containment at OpenAI to hack the company Hugging Face. Several media reports suggest thatresearchers from cybersecurity firm Hacktron AIexploited weaknesses in OpenAI’s community forum, intending to gain access to internal credentials and a ChatGPT account of an employee.


The particular account was linked to GitHub, giving researchers a way into internal software code; from there, they could examine private information and suggest modifications. The Financial Times reported that the researchers had been granted access to a cybersecurity-focused tool built by Anthropic. The move was part of a broader initiative aiming to hunt down weaknesses before bad actors could exploit them.

Add WION as a Preferred Source

OpenAI awards the three researchers

For their efforts, OpenAI awarded the three researchers a combined $6,500 through its bug bounty programme, which compensates security researchers who responsibly flag and report weaknesses. The company confirmed that the issues identified during this process have since been fixed.


“We thank the researchers for contacting us and sharing their findings,” said OpenAI said, according to the Financial Times, while Anthropic declined to comment. This disclosure surfaces just two months after reports emerged that over 1,000 of OpenAI's AI agents had escaped a testing environment and launched hacking attempts against the AI platform Hugging Face, an episode that raised fresh concerns about AI agents' growing capacity to carry out cyberattacks with little to no human oversight.

Trending Stories

Meanwhile, industry leaders including Jensen Huang and Mark Zuckerberg are emphasising AI safety measures, rather than directing it to slow down its development. In a post on X, Zuckerberg said, “My view is that trust and alignment are quickly becoming the most important capabilities that will differentiate agents and models. Any lab that doesn't focus on alignment will fall behind.”

About the Author

Share on twitter

Vinay Prasad Sharma

Vinay Prasad Sharma is a Delhi-based journalist with over three years of newsroom experience, currently working as a Sub-Editor at WION. He specialises in crafting SEO-driven natio...Read More