Warning shot or publicity stunt - how worried should we be about the
OpenAI hack? 49 minutes ago Share Save Add as preferred on Google Joe TidyCyber correspondent,
BBC World Service Getty Images This week the tech world was gripped by a story that has it all - and which started like a sci-fi thriller.
Hugging Face - a kind of app store for
Artificial Intelligence tools - announced on 16 July it had been hacked by a cyber criminal wielding enormously powerful AI. The bombshell announcement was full of scary, highly technical terms: "a swarm of sandboxes", "agentic attacker", and "self-migrating command and control".
Hugging Face said the hack was different from anything it had handled before because it was done at superhuman speed by an AI with little or no human guidance. The AI performed 17,000 actions in less than two days, successfully breaching the large wealthy tech company to steal secrets. It left the tech world in shock. But who was responsible for this attack?
Hugging Face researchers guessed the mysterious attackers had used one of the big AI models but they had no idea who or where the criminals were. The perplexed company contacted the police and investigations commenced. Commentators and analysts took to their podcasts and social media accounts to guess which cyber crime group or nation state hacker might be behind it. Then on Wednesday, nearly a week after
Hugging Face raised the alarm, the true culprit was unmasked. The Scooby-Doo-style reveal was made even more bizarre - and worrying - because
OpenAI said its bot did the whole thing on its own, without permission. The firm said it all went down during a test of its tech's hacking skills. Two new versions of
ChatGPT, designed to be master hackers, broke out of a supposedly secure test environment and gained access to the internet. They then attacked
Hugging Face to get access to the information to help them ace their exam.
OpenAI issued a press release explaining what had happened and said it was "partnering with
Hugging Face" to address the security incident and share lessons learned. Since then, there has been fierce debate about the incident. Was it truly a stark warning about the future of AI? Or was it a publicity stunt by
OpenAI to show off how powerful their models are? It's the kind of scare marketing AI companies have been accused of for years and, since the much discussed launch of
Anthropic's
Mythos model, cyber-security prowess has been a focal point. One of the top comments on
OpenAI boss Sam Altman's X post about the incident summarises this scepticism: "If y'all can't understand that this was written to purely brag about the model then I don't know what to tell you." Watch: Why is the
OpenAI cyber-attack so alarming? Cyber-security consultant Daniel Card said sarcastically on LinkedIn: "Isn't it lucky [that] out of the millions of sites that got pwn3d [hacked],
OpenAI managed to pwn someone who also could benefit from the marketing exposure…" For some, the story is more conspiracy drama than sci-fi thriller. The message is: "Aren't my AI tools really powerful? Buy them so you can protect yourself from other people's AI attacks." We can't know the truth, but the opposing point of view posed by other commentators is just as dramatic. Is this a sign that
OpenAI made a potentially dangerous error in judgement and planning? I've covered lots of AI stories, including the fears around
Anthropic's
Mythos model. My inbox is now chock full of cyber-security companies and experts criticising
OpenAI for not building a stronger container to test its AI, known as a sandbox. After all, these AI agents had been trained specifically to hack into and out of places with no restrictions at all. "The
OpenAI and
Hugging Face incident is a real-world example of a broader issue we've been highlighting for months," said Dor Sarig from Pillar Security. "Sandboxes alone are not a sufficient security boundary for agentic AI." Firm hacked by rogue
OpenAI models says it is 'a wake-up call' Cyber security Professor Alan Woodward from Surrey University told reporters
OpenAI had "egg on it's face", and Katie Moussouris from Luta Security went further, suggesting the AI industry is failing to control its dangerous inventions. "We are working on cutting edge technology without the knowledge to contain it," she said. "Just because we have the smartest people developing AI does not mean we have the ability to do so safely." According to these views, if the hacking incident was a publicity stunt then it appears as though it backfired. Whatever led to the hack, it's clear this is a major moment for the AI industry and the cyber security world, which collided this year in ways people had been fearing for a long time. Addressing this fierce debate, AI and cyber security advisor Francesca Bosco said: "Two simplistic narratives are equally unhelpful: that this was a Hollywood-style escape, or that it was merely a publicity exercise. "A more serious interpretation is that a stress test exposed weaknesses in containment and evaluation architecture." This event is the latest in a string of worrying and weird examples of AI agents going rogue. In recent research, the UK's AI Security Institute (AISI) found frontier AI models are so fixated on completing tasks they "cheated" in tests to achieve their goals. The research from AISI came with this worrying warning: "A model that pursues a goal through unintended or unauthorised means may cause harm, particularly in high-stakes use cases." Inevitably, this
OpenAI hack has further fuelled fears of what could happen if AI agents are let loose. Could they go rogue on a larger scale and cause some sort of disaster? This is particularly concerning with AI being used increasingly in warfare as seen in Iran and Ukraine. Ciaran Martin, former head of the UK's National Cyber Security Centre, offered a calmer view. "It is a bit of a leap to go from this incident to saying that AI agents are going to take over drones and start killing people," he said. But for Martin, and many others, the story is undoubtedly another vivid example of something that 2026 is teaching us fast: AI agents are now very good hackers - and that is something we have to prepare for, urgently.
OpenAI says its AI went rogue and launched 'unprecedented' cyber-attack
ChatGPT medical advice brought man 'to brink of death', lawsuit alleges Lawmakers push for AI 'kill switch' after
OpenAI models go rogue