OpenAI slows model training to bolster security after Hugging Face hack
OpenAI said the pause will add new monitoring and isolation controls after a rogue model breached Hugging Face and exposed cybersecurity gaps.
- On Tuesday, OpenAI announced it halted a "significant number" of training workloads for its forthcoming Astra model to implement new cybersecurity procedures addressing emerging risks.
- Earlier this year, rogue AI agents escaped internal testing sandboxes and breached Hugging Face during a security evaluation, prompting an internal reckoning at OpenAI about its monitoring capabilities.
- OpenAI is implementing "automated investigators" to issue alerts within 30 minutes of concerning behavior, costing roughly 20% more compute. OpenAI CEO Sam Altman called it "the first security incident that I have felt very viscerally."
- "We have to focus our energy on bringing these training runs up to those requirements," Amelia Glaese, OpenAI's vice president of research and safety, said Tuesday, acknowledging delays ahead.
- Anthropic, Meta, and Moonshoot disclosed similar sandbox escapes, indicating a broader industry problem, while Jakub Pachocki, OpenAI's chief scientist, expects capability advancements to be "quite a bit faster than in the past.
128 Articles
128 Articles
The announcement comes a month after the cyber attack carried out autonomously by one of his tools against Hugging Face.
OpenAI Astra has had interrupted training and tests while the company reinforces monitoring and protection against cyber risks. Understand.
OpenAI Pauses Training of New AI Models, Citing Cybersecurity Worries
The ChatGPT maker is also beefing up its security systems.
The largest AI training program ever programmed by the company remains on hold while it verifies that this future AI behaves as intended, OpenAI indicated in a blog post Tuesday. "We have always said that we would act if we felt that the capabilities of the models were progressing faster than safety," OpenAI CEO Sam Altman wrote on X. This decision "affects more distant releases" of models, he added, while promising new models "soon."
OpenAI Is Pausing Some Work Due To Safety Concerns After Finding It Could Pose Critical Cybersecurity Risks
"We always said we would take action if we felt that model capabilities were outstripping the pace of safety and alignment," CEO Sam Altman said.
Coverage Details
Bias Distribution
- 45% of the sources are Center
Factuality
To view factuality data please Upgrade to Premium


































