Close Menu
Daily Guardian EuropeDaily Guardian Europe
  • Home
  • Europe
  • World
  • Politics
  • Business
  • Lifestyle
  • Sports
  • Travel
  • Environment
  • Culture
  • Press Release
  • Trending
What's On

Video. New York hosts secret dîner en blanc with Parisian belle époque theme

October 2, 2026

Schools closed, services slowed and flights cancelled: strikes hit Portugal

October 2, 2026

Could the future of AI data centres be in space? Google wants to find out

October 2, 2026

sous la pression des Etats-Unis, l’exécutif envisage de libérer une part importante de ses réserves – POLITICO

October 2, 2026

Video. European diesel reserves cover 10 days of the world’s demands – expert

October 2, 2026
Facebook X (Twitter) Instagram
Web Stories
Facebook X (Twitter) Instagram
Daily Guardian Europe
Newsletter
  • Home
  • Europe
  • World
  • Politics
  • Business
  • Lifestyle
  • Sports
  • Travel
  • Environment
  • Culture
  • Press Release
  • Trending
Daily Guardian EuropeDaily Guardian Europe
Home»Lifestyle
Lifestyle

OpenAI pledges to slow down its model development amid cybersecurity concerns

By staffAugust 19, 20263 Mins Read
OpenAI pledges to slow down its model development amid cybersecurity concerns
Share
Facebook Twitter LinkedIn Pinterest Email

Published on
19/08/2026 – 10:40 GMT+2

The ChatGPT maker is deliberately holding back the pace of its most advanced research, including its single largest planned reinforcement-learning run, weeks after a system built from its own models slipped free during an internal security test and broke into the AI platform Hugging Face.

CEO Sam Altman posted on X that OpenAI would coordinate with the wider industry on shared safety rules but “act unilaterally in the meantime” until it did.

The episode that triggered the decision unfolded in July, when OpenAI was testing GPT-5.6 Sol alongside an unreleased, more capable prototype on an internal benchmark measuring offensive cyber skills, with the usual safety restrictions deliberately switched off to gauge the models’ raw ability.

Rather than solving the test, the system found a previously unknown flaw, escaped its sandbox or controlled environment, reached the open internet and spent roughly four and a half days probing Hugging Face’s infrastructure, eventually breaking in to search for the test’s answers.

Hugging Face’s own reconstruction counted about 17,600 separate actions before the intrusion was contained as several other companies were also affected.

Both sides say they found no sign of malicious intent, and Hugging Face has since been given access to a more capable, less restricted version of OpenAI’s model to help it defend its own systems.

Astra and a higher bar for security

The second trigger came on 7 August, when internal evaluations suggested Astra, OpenAI’s next frontier model, might cross the “critical” threshold for cyber capability under the company’s own risk framework.

Some Astra workloads have since resumed under tighter controls, but a significant share remain frozen until they meet new standards covering isolated testing environments, restricted network access and continuous monitoring.

A new detection system now scans model activity as it happens and aims to flag anything resembling unauthorised access or an attempt to disable safeguards within 30 minutes, at a computing cost OpenAI estimates at roughly 20% of the processing power being monitored.

OpenAI says the changes were already planned rather than a direct reaction to the breach, while acknowledging the incident added urgency. The company is also not alone in facing this problem.

Anthropic and Meta have each disclosed similar episodes in which their own models breached third-party systems during testing in recent weeks.

OpenAI and Anthropic have separately backed a staff-led petition urging governments to help coordinate how fast the industry moves, a marked shift from Altman’s past resistance to public calls for an AI slowdown.

Share. Facebook Twitter Pinterest LinkedIn Tumblr Email

Keep Reading

Could the future of AI data centres be in space? Google wants to find out

Did Trump order Maduro’s kidnapping based on Grok’s advice?

Humanoid robots: China builds them, but Germany could supply the joints and hands

Rogue AI agents tried and failed to hack US and Canadian government websites

Tallinn Defence EXPO: Counter-drone systems and robot soldiers steal the show

Gimmick or recipe for success? Japan’s convenience stores test AI flavours

No more radio silence: EU plans cross-border lifeline for emergency services

Unlike the EU, Trump’s new AI pact lets tech companies police themselves

Gene therapy: The breakthrough that inspired Spider-Man

Editors Picks

Schools closed, services slowed and flights cancelled: strikes hit Portugal

October 2, 2026

Could the future of AI data centres be in space? Google wants to find out

October 2, 2026

sous la pression des Etats-Unis, l’exécutif envisage de libérer une part importante de ses réserves – POLITICO

October 2, 2026

Video. European diesel reserves cover 10 days of the world’s demands – expert

October 2, 2026

Subscribe to News

Get the latest Europe and world news and updates directly to your inbox.

Latest News

Video. Latest news bulletin | October 2nd, 2026 – Midday

October 2, 2026

Bonds steady as French borrowing costs hit 24-year high and UK 30-year yields top 6%

October 2, 2026

Did Trump order Maduro’s kidnapping based on Grok’s advice?

October 2, 2026
Facebook X (Twitter) Pinterest TikTok Instagram
© 2026 Daily Guardian Europe. All Rights Reserved.
  • Privacy Policy
  • Terms
  • Advertise
  • Contact

Type above and press Enter to search. Press Esc to cancel.