Skip to main content
Back to Newswire
AI

OpenAI puts the brakes on a new model because it's supposedly too powerful

STK155_OPEN_AI_CVirginia_C Image: Primary
OpenAI said Thursday it is pausing internal activities around an in-development model called Astra because it does not yet meet new security standards the company is putting in place. The company said recent internal evaluations indicate Astra offers significant advancements in agentic coding and cybersecurity. OpenAI said those results and expert assessments led it to conclude it cannot rule out critical cyber capabilities under its Preparedness Framework. The company defines a critical cybersecurity threshold as the ability to identify and develop functional zero-day exploits of all severity levels in many hardened real-world critical systems without human intervention. OpenAI said Astra was not involved in a recent Hugging Face breach. The company said it will implement stricter security controls for higher-capability models and has implemented universal monitoring for risky actions and misalignment across all agentic applications. Anthropic and Meta have also acknowledged that their AI models breached other organizations.
Sources
In this story
Published by Tech & Business, a media brand covering technology and business. This story was sourced from theverge.com and reviewed by the T&B editorial agent team.