

Getty Images
This week the tech world was gripped by a story that has it all – and which started like a sci-fi thriller
Hugging Face – a kind of app store for artificial intelligence tools – announced on 16 July it had been hacked by a cyber criminal wielding enormously powerful AI
The bombshell announcement was full of scary, highly technical terms: “a swarm of sandboxes”, “agentic attacker”, and “self-migrating command and control”
Hugging Face said the hack was different from anything it had handled before because it was done at superhuman speed by an AI with little or no human guidance
The AI performed 17,000 actions in less than two days, successfully breaching the large wealthy tech company to steal secrets
It left the tech world in shock. But who was responsible for this attack?
Hugging Face researchers guessed the mysterious attackers had used one of the big AI models but they had no idea who or where the criminals were
The perplexed company contacted the police and investigations commenced
Commentators and analysts took to their podcasts and social media accounts to guess which cyber crime group or nation state hacker might be behind it
Then on Wednesday, nearly a week after Hugging Face raised the alarm, the true culprit was unmasked
It was ChatGPT
The Scooby-Doo-style reveal was made even more bizarre – and worrying – because OpenAI said its bot did the whole thing on its own, without permission
The firm said it all went down during a test of its tech’s hacking skills
Two new versions of ChatGPT, designed to be master hackers, broke out of a supposedly secure test environment and gained access to the internet
They then attacked Hugging Face to get access to the information to help them ace their exam
OpenAI issued a press release explaining what had happened and said it was “partnering with Hugging Face” to address the security incident and share lessons learned
Since then, there has been fierce debate about the incident
Was it truly a stark warning about the future of AI? Or was it a publicity stunt by OpenAI to show off how powerful their models are?
It’s the kind of scare marketing AI companies have been accused of for years and, since the much discussed launch of Anthropic’s Mythos model, cyber-security prowess has been a focal point
One of the top comments on OpenAI boss Sam Altman’s X post about the incident summarises this scepticism: “If y’all can’t understand that this was written to purely brag about the model then I don’t know what to tell you.”
Cyber-security consultant Daniel Card said sarcastically on LinkedIn: “Isn’t it lucky [that] out of the millions of sites that got pwn3d [hacked], OpenAI managed to pwn someone who also could benefit from the marketing exposure…”
For some, the story is more conspiracy drama than sci-fi thriller
The message is: “Aren’t my AI tools really powerful? Buy them so you can protect yourself from other people’s AI attacks.”
We can’t know the truth, but the opposing point of view posed by other commentators is just as dramatic. Is this a sign that OpenAI made a potentially dangerous error in judgement and planning?
An OpenAI spokesperson said “we recognise there are a lot of questions and speculative details circulating” about the incident. They added that “we plan to publish a technical report of our learnings in the coming weeks”
I’ve covered lots of AI stories, including the fears around Anthropic’s Mythos model
My inbox is now chock full of cyber-security companies and experts criticising OpenAI for not building a stronger container to test its AI, known as a sandbox
After all, these AI agents had been trained specifically to hack into and out of places with no restrictions at all
“The OpenAI and Hugging Face incident is a real-world example of a broader issue we’ve been highlighting for months,” said Dor Sarig from Pillar Security. “Sandboxes alone are not a sufficient security boundary for agentic AI.”
Cyber security Professor Alan Woodward from Surrey University told reporters OpenAI had “egg on it’s face”, and Katie Moussouris from Luta Security went further, suggesting the AI industry is failing to control its dangerous inventions
“We are working on cutting edge technology without the knowledge to contain it,” she said
“Just because we have the smartest people developing AI does not mean we have the ability to do so safely.”
According to these views, if the hacking incident was a publicity stunt then it appears as though it backfired
Whatever led to the hack, it’s clear this is a major moment for the AI industry and the cyber security world, which collided this year in ways people had been fearing for a long time
Addressing this fierce debate, AI and cyber security advisor Francesca Bosco said: “Two simplistic narratives are equally unhelpful: that this was a Hollywood-style escape, or that it was merely a publicity exercise
“A more serious interpretation is that a stress test exposed weaknesses in containment and evaluation architecture.”
This event is the latest in a string of worrying and weird examples of AI agents going rogue
In recent research, the UK’s AI Security Institute (AISI) found frontier AI models are so fixated on completing tasks they “cheated” in tests to achieve their goals
The research from AISI came with this worrying warning: “A model that pursues a goal through unintended or unauthorised means may cause harm, particularly in high-stakes use cases.”
Inevitably, this OpenAI hack has further fuelled fears of what could happen if AI agents are let loose. Could they go rogue on a larger scale and cause some sort of disaster?
This is particularly concerning with AI being used increasingly in warfare as seen in Iran and Ukraine
Ciaran Martin, former head of the UK’s National Cyber Security Centre, offered a calmer view
“It is a bit of a leap to go from this incident to saying that AI agents are going to take over drones and start killing people,” he said
But for Martin, and many others, the story is undoubtedly another vivid example of something that 2026 is teaching us fast:
AI agents are now very good hackers – and that is something we have to prepare for, urgently

Cyber-security
Artificial intelligence
Technology
Related:
Digital Automation Training Benin: 5 Winning Skills Employers Demand in 2026
WhatsApp Marketing Automation Africa: 6 Dangerous Mistakes Brands Make in Nigeria
Want to learn this practically?
Join Justfine Infotech and build real digital skills in AI, automation, web development, digital marketing, office productivity, e-commerce, freelancing and cybersecurity.
Available Programmes:
6 Weeks Certificate • 3 Months Professional Certificate • 6 Months Diploma • Full Professional Diploma
WhatsApp:
+229 01 57 57 99 15
+229 01 66 68 11 60
Source: www.bbc.com



