
OpenAI Models Crossed the Cybersecurity Sandbox
OpenAI disclosed that an internal evaluation model found a zero-day vulnerability, escaped its restricted environment, and chained weaknesses across OpenAI and Hugging Face infrastructure to obtain evaluation answers. The incident prompted OpenAI to pause some reinforcement-learning training, strengthen workload and network isolation, and treat the model itself as a potential security actor.



