Related terms
A testing method that deliberately uses challenging, deceptive, or malicious inputs to evaluate an AI system's robustness, reliability, and security.
Jailbreak TestingA security testing method that evaluates whether an AI system can be manipulated into bypassing its safety policies, behavioral constraints, or security guardrails through adversarial prompts or other attack techniques.
Safety TestingThe process of evaluating AI systems to identify vulnerabilities, harmful outputs, and policy violations before deployment.
Bias TestingThe process of evaluating an AI system for unfair, systematic, or discriminatory behavior across different users, groups, or scenarios.
Toxicity TestingThe process of evaluating an AI model or agent to detect and measure toxic, hateful, or abusive language in outputs.