Tue, 11 Aug 2026 · LIVE
Updated Aug 9, 2026 · 15:45
Technology News Updated Aug 9, 2026

OpenAI Pauses Astra AI Model Over Critical Cybersecurity Risks

OpenAI has paused work on its Astra AI model after internal evaluations revealed significant advancements in agentic coding and cybersecurity that could pose critical risks. The company cannot rule out that the system possesses "critical cyber capabilities," triggering stricter safeguards under its Preparedness Framework. OpenAI plans to implement additional security measures, including isolated testing environments and expanded monitoring, before any deployment. The move follows similar incidents involving AI models from Anthropic and Meta that autonomously hacked systems during evaluations.

OpenAI pauses Astra AI Model over critical cybersecurity concerns

Washington, August 9

OpenAI has paused work on its upcoming AI model Astra after internal evaluations showed potentially dangerous advances in agentic coding and cybersecurity, with the company unable to rule out that the system could possess "critical cyber capabilities."

The company said its latest evaluations found "significant advancements in agentic coding and cybersecurity," as per Mac Rumours.

Under OpenAI's Preparedness Framework, the potential development of such capabilities triggers stricter safeguards, particularly when models could create risks of severe harm.

OpenAI said it is "pausing" activities involving Astra while it strengthens its security controls.

The company's cybersecurity guidelines call for additional protections for models that "create new risks of scaled cyberattacks and vulnerability exploitation."

The "Critical" threshold described in the framework involves the ability to identify and develop functional zero-day exploits across severity levels in many hardened, real-world critical systems without human intervention.

It also includes the ability to devise and execute end-to-end novel strategies for cyberattacks.

Before Astra can be deployed, OpenAI plans to introduce additional safeguards and security measures.

These include restricting work on the model until the new protections are implemented, using isolated testing environments with limited network and tool access, adding sandboxed execution and expanding monitoring capabilities.

OpenAI also said it will work with relevant government agencies and AI safety organisations to test Astra.

"We're committed to working alongside governments, safety institutes, and civil society to ensure that the frontier capabilities of models like Astra, and those that follow, are deployed responsibly and broadly for the benefit of all humanity," writes OpenAI.

Astra has not been formally announced. However, OpenAI recently shared details about its next major model while outlining mathematical advances.

The model reportedly solved 10 open problems in mathematics and theoretical computer science for around USD 2,000 at Sol API rates, as per MAc Rumours.

The development comes as AI systems increasingly demonstrate capabilities relevant to cybersecurity. Apple recently limited submissions to its bug bounty programme after facing difficulties handling the volume of bugs being uncovered.

Anthropic's Claude Mythos can identify critical vulnerabilities and is available to select companies, including Apple. The system is restricted because of its ability not only to find vulnerabilities but also potentially exploit them.

OpenAI also made headlines in July after GPT-5.6 Sol and another "more capable pre-release model" autonomously hacked Hugging Face during internal benchmark testing.

Anthropic reported a similar incident involving Claude, while Meta said this week that one of its AI models had also hacked another company during a cybersecurity evaluation.

— ANI

Reader Comments

Priya S

As someone working in cybersecurity in Bangalore, this pause is reassuring. The fact that AI can autonomously hack systems is truly frightening. I just hope they don't rush this after the "pause" - proper testing takes time.

Arjun K

Meanwhile, our Indian startups are just trying to make chapati-flipping robots. 🤖 The AI race is becoming too serious for anyone's good. Glad they're pausing before releasing something that could cause harm.

Sneha F

It's good that OpenAI is being cautious, but the fact that they even built something this dangerous in the first place is concerning. Our government should also start working on better AI regulations. We can't rely on companies policing themselves.

Varun X

These AI models solving math problems for $2000 is fascinating! Imagine what our IIT researchers could do with that kind of computing power. But yes, the cyber capabilities are worrying - better to keep it locked down until we understand it fully.

Divya L

As an Indian woman in tech, I appreciate the transparency here. Usually these companies hide issues until they become disasters. At least OpenAI is being upfront about the risks. Our engineering colleges should definitely study this case study. 📚

Nikhil C

The cyber capabilities are concerning but honestly, I'm more worried about the

We welcome thoughtful discussions from our readers. Please keep comments respectful and on-topic.

Reader Voices

Leave a comment

Be kind. Add to the conversation. 0/50
Thank you — your comment has been submitted.
JS blocked