AI TraceTrace Foundation, Inc.
Content ModerationInternal OnlyVerified

Reviewed and published by trentmaziarz, March 24, 2026. Discovered and drafted by our automated research pipeline.

OpenAI uses internal and external AI safety evaluations — including red-teaming — to test its models for dangerous capabilities before and after deployment.

Details

OpenAI's Preparedness Framework requires safety evaluations of all frontier models before deployment. The company runs an external Red Teaming Network of independent experts who probe models for risks including CBRN (chemical, biological, radiological, nuclear) threats, cybersecurity, and persuasion. OpenAI and Anthropic conducted a joint safety evaluation in 2024. Evaluations cover models including GPT-5, o-series, Sora, and Operator.

Products affected

All frontier modelsGPT-4GPT-4oGPT-4.5GPT-5o1o3SoraOperator

Sources & Evidence

News Article

Academic Paper

Company Disclosure

Other practices by OpenAI

Have evidence about OpenAI's AI practices? Submit a report.

Report a Sighting →