AI TraceTrace Foundation, Inc.
Content ModerationInternal OnlyVerified

Reviewed and published by trentmaziarz, March 24, 2026. Discovered and drafted by our automated research pipeline.

OpenAI offers a Moderation API that automatically classifies text and images for potentially harmful content, available to all API users and used internally across OpenAI products.

Details

The Moderation API uses AI classifiers to detect categories of potentially policy-violating content including hate speech, harassment, self-harm, sexual content, and violence. OpenAI upgraded the API to a multimodal moderation model capable of analyzing both text and images. The API is free to use for developers building on OpenAI's platform and is also applied internally to ChatGPT and DALL-E.

Products affected

OpenAI API (all models)ChatGPTDALL-E

Sources & Evidence

Other practices by OpenAI

Have evidence about OpenAI's AI practices? Submit a report.

Report a Sighting →