Related terms
An evaluation metric that measures whether an AI system's output is fully supported by the provided input, retrieved context, or source material without introducing unsupported or fabricated information.
ConsistencyA metric that measures how reliably an AI system produces stable, coherent, and similar outputs for equivalent inputs or repeated evaluations.
CompletenessA metric that measures how fully an AI response covers the required information, tasks, or expected outputs for a given request.
GroundednessAn evaluation metric that measures whether an AI system's output is supported by the provided context, retrieved information, or source material without introducing unsupported claims or hallucinations.
Reference-based EvaluationAn evaluation method that measures model output quality by comparing generated responses against ground-truth reference data or golden datasets.