Menu

Search

  |   Business

Menu

  |   Business

Search

Google Add as a preferred source on Google

OpenAI Tightens AI Safety Monitoring After Security Incidents

OpenAI Tightens AI Safety Monitoring After Security Incidents. Source: Jernej Furman from Slovenia, CC BY 2.0, via Wikimedia Commons

OpenAI is strengthening oversight of its most advanced artificial intelligence models as the company works to prevent unexpected or potentially dangerous behavior from increasingly autonomous AI systems.

The ChatGPT developer announced Tuesday that it is expanding real-time AI safety monitoring for unreleased models. The enhanced system will track how advanced models handle complex tasks, make decisions and interact with online tools. Potentially risky behavior could be flagged to OpenAI’s safety teams within approximately 30 minutes, allowing researchers to respond more quickly to emerging threats.

OpenAI is also introducing stricter controls for AI models performing higher-risk activities online. Certain systems will operate with tighter internet access restrictions, while training and testing that involves AI-generated or untrusted code will require stronger isolation in secure, sandboxed environments.

The new OpenAI security measures come after the company and rival AI developer Anthropic disclosed incidents in which AI systems unintentionally accessed computer systems belonging to several organizations during security evaluations. Hugging Face was among the organizations affected.

These incidents have highlighted a growing challenge for the AI industry: predicting how highly capable AI agents will behave as they gain greater autonomy and access to external tools and digital environments.

Mia Glaese, OpenAI’s vice president of research, said the company is working to ensure that AI safeguards develop alongside rapidly advancing model capabilities. OpenAI has previously paused work on an upcoming AI model to improve its safety protections. The company also confirmed Tuesday that a major model training run remains suspended.

OpenAI plans to publish a detailed assessment of the Hugging Face security incident as part of its continuing investigation into how its AI models interacted with external computer systems.

The expanded monitoring and sandboxing requirements reflect growing efforts across the AI industry to address security risks before more powerful autonomous AI agents are widely deployed. As AI models become increasingly capable of operating independently online, developers face mounting pressure to ensure safety systems keep pace with technological advances.

  • Market Data
Close

Welcome to EconoTimes

Sign up for daily updates for the most important
stories unfolding in the global economy.