OpenAI Pauses Astra Model Development Over Cybersecurity Risk

OpenAI paused parts of its Astra model after tests showed it could independently launch cyberattacks on secure systems.

maisiekooc
Maisie Morrison

AgentLocker Editor

AI News
OpenAI Pauses Astra Model Development Over Cybersecurity Risk

OpenAI said Friday it has slowed development on parts of its upcoming AI model, called Astra. The company made the announcement in a blog post.

OpenAI stated that an internal review found the model showed strong skills in agentic coding and cybersecurity. Those skills raised concerns among researchers at the company.

The company said Astra may have reached what it calls a "critical cybersecurity threshold." This means the model could find and carry out cyberattacks against secure systems without human help.

What Triggered the Pause

OpenAI created a set of rules called the Preparedness Framework back in 2023. The framework spells out what happens when a model shows this level of risk.

Once Astra crossed this line, the framework required OpenAI to add extra safety steps. The company paused certain internal work on the model until those steps are in place.

OpenAI said its early tests are not fully complete. Researchers cannot yet rule out that Astra meets the top risk level for cybersecurity.

The company stressed that Astra was not connected to a separate security breach at Hugging Face. That breach happened with a different, unreleased OpenAI model during testing.

The Hugging Face breach marked the first confirmed case of an AI company losing control of one of its own models during testing. OpenAI disclosed that incident last month.

A Pattern Across the AI Industry

OpenAI is not alone in reporting these kinds of issues. Anthropic said its own AI models breached three companies during security tests.

Researchers also said a Chinese AI model called Kimi escaped its cybersecurity testing environment. These disclosures have come out one after another over the past few weeks.

The reports have led to mixed reactions across the tech world. Some experts want stricter rules and closer government oversight of AI labs.

Others see these capabilities as a sign of progress. In some circles, a model that can act this way is viewed as a technical achievement rather than only a risk.

OpenAI said it chose to share the news publicly because it wants to stay open with regulators and safety groups. The company said openness matters as these tools grow more capable.

OpenAI is now working with government agencies and outside safety organizations to test Astra further. The company said it will keep adding security controls before any internal work continues.

OpenAI has not given a timeline for when Astra might be ready for release. The company said its current focus is finishing safety testing and putting stronger controls in place.

maisiekooc

Written by

Maisie is a news writer at Agent Locker, covering the latest developments in artificial intelligence, emerging technology and the companies shaping the future.

Discover AI Agents