Premium

OpenAI’s AI agent went ‘rogue’ and hacked another company: What it says about safety

The recent incident where OpenAI’s AI model autonomously breached Hugging Face has raised more questions about the reliability of AI models and deepened the debate between open-source and closed models.

The incident also raises some pertinent questions, such as how much can any current AI system, closed model or a policy restricting open ones, actually be relied upon. (Express Image/Magnific)The incident also raises some pertinent questions, such as how much can any current AI system, closed model or a policy restricting open ones, actually be relied upon. (Image for represetation/Magnific)
Written by: Bijin Jose
12 min readNew DelhiJul 26, 2026 09:51 AM IST First published on: Jul 23, 2026 at 04:19 PM IST

OpenAI recently revealed that an AI agent system powered by one of its models hacked a popular AI developer platform during a test session. The AI agent went ‘rogue’, accessed the open web, and hacked the startup in what the ChatGPT-maker itself describes as ‘unprecedented’.

OpenAI said that the startup Hugging Face had traced and contained the agent, calling it a cyber-incident that involved state-of-the-art cyber capabilities. Eventually, the company disclosed that two of its own AI models, including its flagship GPT-5.6 Sol and a more powerful unreleased model, escaped a testing environment, accessed the internet, and hacked into the internal systems of Hugging Face, a platform that is used to host and distribute open-weight AI models, datasets, and research tools.

Bijin Jose serves as an Assistant Editor at Indian Express O... Read More

Latest Comment
Post Comment
Read Comments