OpenAI is strengthening oversight of its most advanced artificial intelligence models as the company works to prevent unexpected or potentially dangerous behavior from increasingly autonomous AI systems.
The ChatGPT developer announced Tuesday that it is expanding real-time AI safety monitoring for unreleased models. The enhanced system will track how advanced models handle complex tasks, make decisions and interact with online tools. Potentially risky behavior could be flagged to OpenAI’s safety teams within approximately 30 minutes, allowing researchers to respond more quickly to emerging threats.
OpenAI is also introducing stricter controls for AI models performing higher-risk activities online. Certain systems will operate with tighter internet access restrictions, while training and testing that involves AI-generated or untrusted code will require stronger isolation in secure, sandboxed environments.
The new OpenAI security measures come after the company and rival AI developer Anthropic disclosed incidents in which AI systems unintentionally accessed computer systems belonging to several organizations during security evaluations. Hugging Face was among the organizations affected.
These incidents have highlighted a growing challenge for the AI industry: predicting how highly capable AI agents will behave as they gain greater autonomy and access to external tools and digital environments.
Mia Glaese, OpenAI’s vice president of research, said the company is working to ensure that AI safeguards develop alongside rapidly advancing model capabilities. OpenAI has previously paused work on an upcoming AI model to improve its safety protections. The company also confirmed Tuesday that a major model training run remains suspended.
OpenAI plans to publish a detailed assessment of the Hugging Face security incident as part of its continuing investigation into how its AI models interacted with external computer systems.
The expanded monitoring and sandboxing requirements reflect growing efforts across the AI industry to address security risks before more powerful autonomous AI agents are widely deployed. As AI models become increasingly capable of operating independently online, developers face mounting pressure to ensure safety systems keep pace with technological advances.


Anthropic Eyes $10B-Plus Credit Line Ahead of Potential IPO
Shein Targets $26B-$27B Valuation for Hong Kong IPO
Ferrari Luce EV Sells for Record $40 Million at Charity Auction
Google to Move All Pixel Production Out of China by 2027
SMIC Shares Rally as Q2 Profit Surges 262% on Strong Chip Demand
Google to Buy Spirit Airlines Business Data for $10 Million to Train AI
LG, Nvidia Expand AI Partnership With Humanoid Robots, AI Factories
Samsung, SK Hynix Shares Surge on Report of Potential Temasek Investment
Micron Stock Upgraded to Buy as New Street Sees $2 Trillion Valuation Potential
OpenAI Executive Brad Lightcap to Leave for New AI Venture
Nvidia Eyes $3 Billion Investment in SoftBank-Backed AI Data Center
NAB Q3 Cash Earnings Rise 2% as Lower Credit Charges Boost Profit
Stripe, Advent Reportedly Pursue $53 Billion PayPal Takeover
Nvidia’s $500 Billion AI Infrastructure Push Wins Morgan Stanley Support
Apple Develops China-Specific AI Model With Alibaba as Apple Intelligence Launch Nears 



