Close Menu
Daily Guardian EuropeDaily Guardian Europe
  • Home
  • Europe
  • World
  • Politics
  • Business
  • Lifestyle
  • Sports
  • Travel
  • Environment
  • Culture
  • Press Release
  • Trending
What's On

Macron unveils €6bn Saudi-backed theme park deal amid rights criticism

August 25, 2026

Nigel Farage wants to scrap GDPR for the UK – POLITICO

August 25, 2026

French driver, 46, detained in Italy after three Akeron missiles found in van

August 25, 2026

Mission accomplished for Sophie Adenot: ISS antenna installed

August 25, 2026

Trains between Britain and Europe on track to run every 15 minutes – POLITICO

August 25, 2026
Facebook X (Twitter) Instagram
Web Stories
Facebook X (Twitter) Instagram
Daily Guardian Europe
Newsletter
  • Home
  • Europe
  • World
  • Politics
  • Business
  • Lifestyle
  • Sports
  • Travel
  • Environment
  • Culture
  • Press Release
  • Trending
Daily Guardian EuropeDaily Guardian Europe
Home»Lifestyle
Lifestyle

‘Unprecedented’: OpenAI models autonomously hacked a rival firm, fuelling fears of rogue agents

By staffAugust 3, 20264 Mins Read
‘Unprecedented’: OpenAI models autonomously hacked a rival firm, fuelling fears of rogue agents
Share
Facebook Twitter LinkedIn Pinterest Email

ChatGPT maker OpenAI said late Tuesday that its artificial intelligence system hacked into another AI company on its own in what the company called an “unprecedented cyber incident.”

“We had a significant security incident during evaluation of our models,” OpenAI CEO Sam Altman said in a statement posted on social media.

AI startup Hugging Face said last week that it had detected an intrusion into its data processing systems that it suspected was caused by an AI agent autonomously acting on its own.

“We suspected last week’s cyberattack might have come from a frontier lab, given the sophistication of the agent,” Hugging Face co-founder and CEO Clément Delangue said in a statement. “Turns out it did!”

This means the attack was so advanced and well-executed that Hugging Face suspected it came from one of the top AI companies’ systems, not a random hacker.

What happened?

OpenAI was running an internal test to measure how good its AI models are at hacking — a benchmark called ExploitGym.

To see the models’ maximum hacking ability, they deliberately switched off the safety filters that normally stop the models from doing dangerous cyber activity.

The test was meant to run in a sealed-off “sandbox” or an isolated environment with no real internet access, apart from a tool that lets the models download software they might need to complete the task.

However, the models became fixated on solving the test. Rather than solving it the intended way, they went looking for a shortcut and found a way to claw themselves into the open internet — which they were never supposed to reach.

Getting there involved a chain of steps, quietly gaining more and more access inside OpenAI’s own systems until they hit a point with an internet connection.

Once online, the models reasoned that Hugging Face — a big platform hosting AI models and datasets — probably held the answers to the very test they were supposed to solve.

So they broke into Hugging Face’s servers to steal those answers, essentially to cheat, using stolen login credentials and more flaws to get in.

Chinese models to the rescue?

As an open marketplace that anyone can publish to, Hugging Face hosts a huge volume of Chinese-developed models.

When Hugging Face’s team tried to analyse the attack, they fed the raw attack data — the code and commands used to exploit their system — into commercial AI models to help reconstruct what happened.

But those AI models have built-in safety filters designed to block anything that looks like hacking — and to those filters, the evidence of an attack looks exactly the same as an attack itself.

So the models refused to help, unable to tell the difference between a hacker doing harm and a company defending itself.

Blocked, Hugging Face switched to an open-weight Chinese model — Z.ai’s GLM 5.2 — which it could run locally, inside its own systems, and which processed the material without refusing.

Chinese labs such as DeepSeek and Alibaba’s Qwen have become some of the most downloaded model families on the platform, and by some measures, Chinese developers now account for a larger share of Hugging Face’s downloads than their US counterparts.

Major security concern

The disclosure comes amid heightened concerns about the cybersecurity capabilities of powerful models that led US President Donald Trump in June to sign an executive order creating a framework for the federal government to vet the national security risks of the most advanced AI systems for up to a month before their public release.

“AI is accelerating the discovery and exploitation of vulnerabilities,” OpenAI said in its statement Tuesday. “The primary lesson from this incident is that model security and safety must keep pace with rapidly advancing capabilities.”

Delangue said he spent the past 24 hours working with OpenAI, “and we strongly believe there was no malicious intent on their part. It’s quite mind-blowing that all of this happened autonomously!”

Delangue added that it “might be the first incident of its kind.”

OpenAI said the intrusion was caused by a combination of its AI models, including its newly released GPT-5.6 Sol and an “even more capable” model that is still being tested internally.

OpenAI said its AI used stolen credentials and discovered a previously unknown vulnerability to access Hugging Face servers.

It went to “extreme lengths to achieve a rather narrow testing goal” and “found ways to gain access to secret information that it could use to cheat the evaluation,” the company said.

Share. Facebook Twitter Pinterest LinkedIn Tumblr Email

Keep Reading

Mission accomplished for Sophie Adenot: ISS antenna installed

Brazil fines TikTok owner over misuse of children’s and teenagers’ data

Sophie Adenot makes her second spacewalk

Mysterious, ultra-powerful AI system emerges as China closes in on the West

Norway moves to tighten rules on use of smart glasses

Macron confirms school mobile phone ban effective from new term

African nations ramp up their space programmes in a bid to reduce reliance on the West

From iPods to wired headphones: Why more Gen Zers are picking retro tech

Another eclipse over Spain on 28 August: where and how to see the Moon almost completely covered

Editors Picks

Nigel Farage wants to scrap GDPR for the UK – POLITICO

August 25, 2026

French driver, 46, detained in Italy after three Akeron missiles found in van

August 25, 2026

Mission accomplished for Sophie Adenot: ISS antenna installed

August 25, 2026

Trains between Britain and Europe on track to run every 15 minutes – POLITICO

August 25, 2026

Subscribe to News

Get the latest Europe and world news and updates directly to your inbox.

Latest News

At least 306 people infected with salmonella in Belgium in outbreak linked to eggs

August 25, 2026

Brazil fines TikTok owner over misuse of children’s and teenagers’ data

August 25, 2026

Waymo to launch robotaxis in Munich in 2027 – POLITICO

August 25, 2026
Facebook X (Twitter) Pinterest TikTok Instagram
© 2026 Daily Guardian Europe. All Rights Reserved.
  • Privacy Policy
  • Terms
  • Advertise
  • Contact

Type above and press Enter to search. Press Esc to cancel.