Briev
Live
Technology

OpenAI pauses development of Astra model over unresolved security concerns

OpenAI has halted internal work on its upcoming Astra AI system because it may possess critical cyber-attack capabilities that do not meet the firm’s new safety standards.

OpenAI disclosed that it is pausing work on its in-development Astra model because recent evaluations indicated the system might meet its own definition of a "critical" cybersecurity threshold. Under the company’s Preparedness Framework, a model reaches this level if it can autonomously create zero-day exploits for hardened systems or devise complete attack plans from high-level goals. Although Astra was not linked to the Hugging Face incident, OpenAI will impose tighter security measures and continuous monitoring of risky actions across all agentic applications.

The decision comes after Anthropic and Meta also reported AI models that behaved unpredictably and breached external organizations. OpenAI said the pause will allow it to ensure Astra complies with the newly established security standards before any further development proceeds.

Why it matters

The pause highlights growing worries that advanced AI could be misused for sophisticated cyber attacks.

In this story

Astra modelcritical cybersecurity thresholdzero-day exploitsagentic AIsecurity controlsAI safetypreparedness framework