Skip to main content
RAI (Responsible AI) guardrails protect your agents by filtering harmful content, detecting prompt injection, handling PII, and enforcing topic boundaries.

Quick Start


Key Features

Toxicity Detection

Block or filter toxic, harmful, or offensive content:

Prompt Injection Protection

Prevent malicious prompt manipulation:

PII Detection & Handling

Detect and handle personally identifiable information:

Secrets Detection

Prevent API keys, passwords, and other secrets from leaking:

Topic Control

Restrict agent to specific topics:

NSFW Detection

Filter inappropriate content:

How RAI Works


RAI Components


PII Types


Actions

PII Actions

Secrets Actions


Use Cases

Customer Support

Healthcare

Children’s Content


Next Steps

Creating Policies

Learn how to create and configure RAI policies

RAI Features

Detailed guide to each RAI feature