AI Safety and Regulation
August 7, 2026

OpenAI strengthens safeguards as Astra approaches critical cyber capabilities

OpenAI says preliminary evaluations of its upcoming Astra model indicate potentially critical cybersecurity capabilities, prompting stronger security controls, expanded monitoring, restricted access, and collaboration with governments and AI safety organizations.

OpenAI has announced stronger security measures after preliminary evaluations indicated that Astra, an upcoming model, may reach the Critical cybersecurity threshold under its Preparedness Framework.

Tests showed significant advances in agentic coding and cybersecurity, raising concerns that the model could potentially discover zero-day exploits or execute complex attacks against hardened systems. OpenAI is introducing isolated testing environments, restricted network and tool access, stronger model weight protections, sandboxed execution, and universal monitoring for risky agentic actions.

The company is also pausing internal Astra activities that lack required controls and will collaborate with government agencies and AI safety organizations on further evaluations.

#
OpenAI

Read Our Content

See All Blogs
LLM Models

GPT-5.6 benchmarks: The full testing breakdown

Deveshi Dabbawala

August 10, 2026
Read more
LLM Models

LLM testing of Claude Opus 5: The first enterprise-ready frontier AI

Sarankumar S

August 10, 2026
Read more