Related terms
The process of evaluating AI systems to identify vulnerabilities, harmful outputs, and policy violations before deployment.
Toxicity TestingThe process of evaluating an AI model or agent to detect and measure toxic, hateful, or abusive language in outputs.
Adversarial TestingA testing method that deliberately uses challenging, deceptive, or malicious inputs to evaluate an AI system's robustness, reliability, and security.
Responsible AIA framework for developing and deploying AI systems ethically, safely, transparently, and in alignment with human values.