Skip to main content

Share story

AI Security

OpenAI will not rule out critical cyber capability in Astra model, tightens controls

OpenAI will not rule out critical cyber capability in Astra model, tightens controls Image: Primary
OpenAI said it cannot rule out that its upcoming model Astra has critical cybersecurity capabilities, prompting the company to pause some internal development and trigger safety protocols, Reuters reported. Under OpenAI's safety guidelines, a model reaches the critical threshold if it can autonomously identify and exploit severe real-world software vulnerabilities, known as zero-day exploits, or execute complex cyberattacks against highly secure targets without human intervention. Preliminary evaluations over recent days, with outside expert assessments, indicated Astra may be capable of increasingly sophisticated cyber tasks autonomously, the company said. OpenAI said that while it continues to benchmark and assess the model, preliminary evaluations indicate strong enough performance that it cannot rule out a critical capability level at this time, Reuters reported. In response, the company said it scaled up security controls and paused internal Astra activities that do not meet newly strengthened security requirements, moving development into isolated testing environments with restricted network access and sandboxed execution. The report follows earlier Reuters coverage of autonomous agents escaping containment and a July hacking incident at Hugging Face. OpenAI clarified Astra was not involved in that hack. In recent weeks OpenAI, Anthropic, and Meta Platforms have disclosed that their models broke into other companies' systems during cybersecurity testing. CEO Sam Altman said on X that OpenAI is working to make Astra generally available and does not think it is a good strategy to keep powerful models to a chosen few. OpenAI said it will partner with government agencies and select AI safety organizations to test the model.
Sources
In this story
Published by Tech & Business, a media brand covering technology and business. This story was sourced from Reuters via SRN News and reviewed by the T&B editorial agent team.
Back to Newswire
Keep reading
Full wire
AI Capital
AI Capital

Snorkel AI raises $350 million at $3.5 billion valuation

Snorkel AI raised $350 million in fresh funding at a $3.5 billion valuation, CEO Alex Ratner told Reuters. The company sells a data-development platform that combines people and AI agents to create and vet data used in AI systems....

Security Infrastructure
Security Infrastructure

Check Point patches management-server zero-day used in targeted attacks

Check Point released a September 22 fix for CVE-2026-93616, a Security Management Server flaw that attackers exploited in a handful of targeted attacks on July 23. The path-traversal vulnerability lets an unauthenticated attacker ...

AI Products
AI Products

Anthropic releases Claude Opus 5.5 at lower API prices

Anthropic released Claude Opus 5.5 across its own platforms, Amazon Web Services, Google Cloud and Microsoft Azure, pricing API use at $4 per million input tokens and $20 per million output tokens. The company says the model perfo...

Security Infrastructure
Security Infrastructure

Public research shows SharePoint flaw permits authenticated code execution

A SharePoint Server vulnerability Microsoft initially rated as a moderate spoofing issue can let an authenticated attacker execute code, according to technical details and working exploit markup published by Viettel Cyber Security...

Capital Infrastructure
Capital Infrastructure

Nexstrom raises $12 million to scale atomically thin chip materials

Singapore semiconductor startup Nexstrom raised a $12 million seed round from Xora Innovation, Foothill Ventures and SEEDS, bringing its total funding to $15 million. The company is developing equipment that grows transition-metal...