Response Sanitization
The process of filtering or removing sensitive, unsafe, or non-compliant content from model outputs before they reach users.
Explore more about Guardrails
Related terms
The process of checking AI-generated outputs against expected formats, schemas, rules, or safety requirements before they are accepted or used.
Content FilteringThe process of detecting, blocking, or modifying content that violates safety policies, compliance requirements, or application rules.
ModerationThe process of detecting, classifying, and handling harmful, unsafe, or policy-violating content in AI inputs and outputs.
Safety FilterA mechanism designed to detect and block toxic, harmful, or policy-violating content in model inputs and outputs.
PII DetectionPIIA guardrail that identifies personally identifiable information in inputs or outputs so it can be protected, redacted, or blocked.
Sensitive Data DetectionThe automated identification and scanning of confidential information or personally identifiable data within system inputs and outputs.