AI TraceTrace Foundation, Inc.
Content ModerationVerified

Reviewed and published by trentmaziarz, July 20, 2026. Discovered and drafted by our automated research pipeline.

Red Hat offers AI Guardrails within Red Hat OpenShift AI — available as a technology preview since March 2025 — that monitor and filter both user inputs and AI model outputs to detect and block hateful, abusive, or policy-violating content in enterprise AI deployments.

Details

Red Hat OpenShift AI 2.18 introduced AI Guardrails as a technology preview. The feature helps improve LLM accuracy, performance, latency, and transparency by detecting potentially hateful, abusive, or profane speech, personally identifiable information (PII), competitive information, and content limited by corporate policies. Red Hat AI 3.4 extended safety capabilities with automated adversarial scanning using the Garak tool to screen models for jailbreaks, prompt injections, and bias, paired with NVIDIA NeMo Guardrails for runtime safety enforcement. These features are provided as tools for enterprise customers building or deploying AI applications, not as moderation of Red Hat's own user community.

Products affected

Red Hat OpenShift AIRed Hat AI

Sources & Evidence

Cite this record

Trace Foundation. (2026). Red Hat: Red Hat offers AI Guardrails within Red Hat OpenShift AI — available as a technology preview since March 2025 — that monitor and filter both user inputs and AI model outputs to detect and block hateful, abusive, or policy-violating content in enterprise AI deployments (data as of 2026-07-20) [Data set record]. AI Trace. https://www.aitrace.org/r/practice/74581a75-4e91-46f1-a577-5254cc24e132. Accessed September 4, 2026.

Stable link
https://www.aitrace.org/r/practice/74581a75-4e91-46f1-a577-5254cc24e132
Data as of
July 20, 2026
Last verified
Not recorded

How to cite AI Trace

Other practices by Red Hat

Have evidence about Red Hat's AI practices? Submit a report.

Report a Sighting →