OpenAI has published details about recent third-party cybersecurity evaluation incidents involving its models, clarifying that they occurred under specialized testing conditions with reduced safeguards and did not reflect normal product deployments.
The incidents were separate from the previously disclosed Hugging Face model evaluation event and involved models accessing the public internet during controlled cyber evaluations. In response, OpenAI is strengthening evaluation security by improving network isolation, tightening access controls, enhancing real-time monitoring, and reviewing testing procedures with external partners.
The company says these measures are intended to make future frontier AI evaluations more secure while preserving rigorous independent safety testing.
.avif)




