OpenAI has reportedly restricted parts of the development of its next-generation AI model, code-named Astra, after early evaluations indicated the system could potentially carry out sophisticated cyberattacks with limited or no human assistance.
The San Francisco-based artificial intelligence company said Thursday that Astra demonstrated significant advances in coding and cybersecurity capabilities during preliminary testing. The results raised concerns that the unreleased OpenAI model could approach the company’s internal “Critical” risk threshold.
Under OpenAI’s safety framework, an AI model may receive a Critical classification if it becomes capable of independently identifying previously unknown, or zero-day, software vulnerabilities or conducting end-to-end cyberattacks against secured networks without human guidance.
In response to the findings, OpenAI is limiting Astra development to highly controlled and isolated environments. Internal workflows that do not meet stricter security requirements have reportedly been paused as the company introduces additional safeguards.
Among the planned security measures are automated monitoring systems designed to detect and interrupt potentially dangerous or misaligned model behavior in real time. OpenAI also plans to involve government organizations and independent AI safety institutes in external testing designed to identify vulnerabilities and assess Astra’s cybersecurity risks.
The move highlights a growing challenge for the AI industry as developers race to build increasingly capable models while attempting to prevent those systems from being misused for hacking and other cyber threats. Earlier OpenAI technology, including GPT-5.6-Sol, reportedly reached a maximum internal risk classification of “High,” making Astra’s potential capabilities particularly significant.
OpenAI emphasized that Astra has not been publicly released and was not connected to recent prominent cybersecurity incidents, including the Hugging Face exploit. The company presented the development restrictions as evidence that its safety systems are identifying potentially dangerous capabilities before advanced AI technology reaches consumers or enterprise customers.
Astra’s testing could become an important case for how leading AI companies manage increasingly powerful cybersecurity capabilities while balancing innovation, AI safety and responsible deployment.


UOB Q2 Net Profit Rises 10% as Wealth Management Growth Boosts Earnings
T-Mobile Executives Reportedly Oppose $300 Billion Deutsche Telekom Merger
Western Digital Q4 Earnings Beat Estimates as FY2027 Outlook Tops Expectations
Meta AI Model Exploits Security Flaw During Cybersecurity Test, Raising AI Safety Concerns
Apple Restores Telegram to App Store After Content Policy Violation
SpaceX Targets Starship Flight 14 With First V3 Starlink Satellite Launch
SoftBank Q1 Profit Beats Forecast as Intel Rally and OpenAI Investments Boost Returns
AMP Shares Surge 13% After Strong Profit and A$150 Million Buyback
Cloudflare Stock Jumps 15% as Earnings Beat Estimates, 2026 Outlook Raised
Telegram Restored on Apple App Store After Temporary Removal
SanDisk Q4 Earnings Beat Estimates as Q1 Revenue Outlook Meets Expectations
Nvidia to Invest Up to $3 Billion in Blackstone-Backed Lancium
Sony Raises Full-Year Outlook After Q1 Profit Jumps 40% on Gaming and Chip Growth
Samsung, SK Hynix Test AMEC Chipmaking Tools for China Backup Plan 



