Saturday, July 25, 2026
Home / World / Warning shot or publicity stunt - how worried shou...
World

Warning shot or publicity stunt - how worried should we be about the OpenAI hack?

CN
CitrixNews Staff
·
Warning shot or publicity stunt - how worried should we be about the OpenAI hack?
OpenAI logoImage source, Getty ImagesByJoe TidyCyber correspondent, BBC World Service
  • Published22 minutes ago

This week the tech world was gripped by a story that has it all - and which started like a sci-fi thriller.

Hugging Face - a kind of app store for artificial intelligence tools - announced on 16 July it had been hacked by a cyber criminal wielding enormously powerful AI.

The bombshell announcement was full of scary, highly technical terms: "a swarm of sandboxes", "agentic attacker", and "self-migrating command and control".

Hugging Face said the hack was different from anything it had handled before because it was done at superhuman speed by an AI with little or no human guidance.

The AI performed 17,000 actions in less than two days, successfully breaching the large wealthy tech company to steal secrets.

It left the tech world in shock. But who was responsible for this attack?

Hugging Face researchers guessed the mysterious attackers had used one of the big AI models but they had no idea who or where the criminals were.

The perplexed company contacted the police and investigations commenced.

Who did it?

Commentators and analysts took to their podcasts and social media accounts to guess which cyber crime group or nation state hacker might be behind it.

Then on Wednesday, nearly a week after Hugging Face raised the alarm, the true culprit was unmasked.

It was ChatGPT.

The Scooby-Doo-style reveal was made even more bizarre - and worrying - because OpenAI said its bot did the whole thing on its own, without permission.

The firm said it all went down during a test of its tech's hacking skills.

Two new versions of ChatGPT, designed to be master hackers, broke out of a supposedly secure test environment and gained access to the internet.

They then attacked Hugging Face to get access to the information to help them ace their exam.

OpenAI issued a press release explaining what had happened and said it was "partnering with Hugging Face" to address the security incident and share lessons learned.

Conspiracy drama

Since then, there has been fierce debate about the incident.

Was it truly a stark warning about the future of AI? Or was it a publicity stunt by OpenAI to show off how powerful their models are?

It's the kind of scare marketing AI companies have been accused of for years and, since the much discussed launch of Anthropic's Mythos model, cyber-security prowess has been a focal point.

One of the top comments on OpenAI boss Sam Altman's X post about the incident summarises this scepticism: "If y'all can't understand that this was written to purely brag about the model then I don't know what to tell you."

To play this video you need to enable JavaScript in your browser.

This video can not be played

Figure caption,

Watch: Why is the OpenAI cyber-attack so alarming?

Cyber-security consultant Daniel Card said sarcastically on LinkedIn: "Isn't it lucky [that] out of the millions of sites that got pwn3d [hacked], OpenAI managed to pwn someone who also could benefit from the marketing exposure…"

For some, the story is more conspiracy drama than sci-fi thriller.

The message is: "Aren't my AI tools really powerful? Buy them so you can protect yourself from other people's AI attacks."

We can't know the truth, but the opposing point of view posed by other commentators is just as dramatic. Is this a sign that OpenAI made a potentially dangerous error in judgement and planning?

Comedy of errors

I've covered lots of AI stories, including the fears around Anthropic's Mythos model.

My inbox is now chock full of cyber-security companies and experts criticising OpenAI for not building a stronger container to test its AI, known as a sandbox.

After all, these AI agents had been trained specifically to hack into and out of places with no restrictions at all.

"The OpenAI and Hugging Face incident is a real-world example of a broader issue we've been highlighting for months," said Dor Sarig from Pillar Security. "Sandboxes alone are not a sufficient security boundary for agentic AI."

Cyber security Professor Alan Woodward from Surrey University told reporters OpenAI had "egg on it's face", and Katie Moussouris from Luta Security went further, suggesting the AI industry is failing to control its dangerous inventions.

"We are working on cutting edge technology without the knowledge to contain it," she said.

"Just because we have the smartest people developing AI does not mean we have the ability to do so safely."

According to these views, if the hacking incident was a publicity stunt then it appears as though it backfired.

Whatever led to the hack, it's clear this is a major moment for the AI industry and the cyber security world, which collided this year in ways people had been fearing for a long time.

Addressing this fierce debate, AI and cyber security advisor Francesca Bosco said: "Two simplistic narratives are equally unhelpful: that this was a Hollywood-style escape, or that it was merely a publicity exercise.

"A more serious interpretation is that a stress test exposed weaknesses in containment and evaluation architecture."

Disaster movie

This event is the latest in a string of worrying and weird examples of AI agents going rogue.

In recent research, the UK's AI Security Institute (AISI) found frontier AI models are so fixated on completing tasks they "cheated" in tests to achieve their goals.

The research from AISI came with this worrying warning: "A model that pursues a goal through unintended or unauthorised means may cause harm, particularly in high-stakes use cases."

Inevitably, this OpenAI hack has further fuelled fears of what could happen if AI agents are let loose. Could they go rogue on a larger scale and cause some sort of disaster?

This is particularly concerning with AI being used increasingly in warfare as seen in Iran and Ukraine.

Ciaran Martin, former head of the UK's National Cyber Security Centre, offered a calmer view.

"It is a bit of a leap to go from this incident to saying that AI agents are going to take over drones and start killing people," he said.

But for Martin, and many others, the story is undoubtedly another vivid example of something that 2026 is teaching us fast:

AI agents are now very good hackers - and that is something we have to prepare for, urgently.

A green promotional banner with black squares and rectangles forming pixels, moving in from the right. The text says: “Tech Decoded: The world’s biggest tech news in your inbox every Monday.”

Sign up for our Tech Decoded newsletter to follow the world's top tech stories and trends. Outside the UK? Sign up here.

Related topics

Originally reported by BBC News. Read the full story at the original source.