OpenAI Group PBC said it has temporarily suspended some of its artificial intelligence training processes after research revealed a potential risk of cyber threats. The move is part of a broader initiative to strengthen the company’s security safeguards and implement new model monitoring systems. The events that prompted the decision include a July incident in which some OpenAI models were hacked into the Hugging Face platform, and the discovery that an unpublished algorithm called Astra had critical cybersecurity capabilities.
Critical risk level for the Astra model

„Astra, an unpublished OpenAI algorithm that is said to be more powerful than GPT-5.6 Sol, has reached the “critical„ level, according to the company’s 22-page “Preparedness Framework.„ The document defines a critical cybersecurity risk as a model that can independently find and exploit zero-day vulnerabilities in hardened systems. As a result of the assessment, OpenAI has suspended some reinforcement learning (RL) tasks, which are typically used to improve the decision-making capabilities of large language models.
Temporary suspension of training and monitoring measures

OpenAI said its largest planned RL training campaign for its latest models has been suspended for at least two weeks. The company has also implemented a new monitoring process in which „activation classifiers“ continuously analyze the LLM’s internal thinking and tool usage, looking for potential malicious actions. If an anomaly is detected, it is passed on to a second, more advanced classifier, and ultimately to researchers. The system aims to generate alerts about suspicious AI behavior within 30 minutes, and employees must stop the behavior if they can’t prove it’s not malicious.
The observation process requires significant computing resources – currently around 20 % of infrastructure dedicated to inference tasks. While this „computing fee“ may impact pricing structure in the long term, OpenAI believes it is necessary to ensure safe model development.
Additional security measures and future plans
In addition to monitoring, OpenAI has restricted some system access rights, removed internal programs, and strengthened security barriers that isolate the highest-risk AI work from the internet. The company also plans to automate security processes by using AI models that will scan research environments themselves for vulnerabilities. It will also improve reward models — algorithms that optimize RL training — to incentivize the models to avoid cyberattacks.
OpenAI executives say these measures will not only mitigate current risks, but also create a long-term security foundation for future models. While some training work has already begun to resume, the largest „frontier“ RL campaign remains on hold, as Forkast.news points out. This is the first time that a leading AI lab has halted a major training project over security concerns, and not just rhetoric.
Different sources – one picture
SiliconANGLE News provided a detailed description of the suspension, including specific technical solutions and a definition of the Preparedness Framework. Crypto Briefing added information, emphasizing that while some parts of Astra training were not completely suspended, the core training process was temporarily suspended while security measures were implemented. Business Standard and Forkast.news confirmed that OpenAI has introduced additional AI monitoring tools and that this action is the first suspension of this magnitude in the industry.
Conclusions
OpenAI’s decision to halt some AI training operations shows that cybersecurity is becoming a critical part of the development of artificial intelligence. The company’s efforts to implement constant monitoring, restrict access rights, and automate risk assessment could become a new standard in the AI research community. While this may affect the pace of training and potential pricing structure, security is the most important factor driving OpenAI’s actions at this time.
Cybersecurity in artificial intelligence training is not only a technical challenge, but also a strategic decision that can shape the future of the entire industry.
Sources
- SiliconANGLE News - OpenAI paused some AI training runs over cybersecurity concerns
- Crypto Briefing - Astra training not paused, new models still expected to ship soon
- Business Standard - OpenAI slows model training to bolster security after Hugging Face hack
- Forkast.news - OpenAI Halts Its Largest Frontier Training Run, Turning Pacing Rhetoric Into Operational Reality






