Related terms
A performance metric measuring the degree to which a system consistently performs its intended function without failure over time.
Adversarial TestingA testing method that deliberately uses challenging, deceptive, or malicious inputs to evaluate an AI system's robustness, reliability, and security.
Red TeamingA structured testing methodology where adversarial tactics are simulated to discover vulnerabilities, safety flaws, and failure modes in a system.
Safety TestingThe process of evaluating AI systems to identify vulnerabilities, harmful outputs, and policy violations before deployment.
AccuracyA metric that measures the proportion of correct predictions or outputs produced by a model or AI system out of all evaluated cases.