Pointwise Evaluation
An evaluation method that assigns an independent score to each individual model response using defined criteria, without comparing it directly to another response.
Explore more about Evaluation Methods
Related terms
An evaluation method that compares two outputs directly to determine which performs better against a defined criterion or preference.
Reference-based EvaluationAn evaluation method that measures model output quality by comparing generated responses against ground-truth reference data or golden datasets.
Reference-free EvaluationAn evaluation approach that assesses model output quality without relying on ground-truth reference answers or golden datasets.
Rubric-based EvaluationAn evaluation approach that assesses AI outputs against structured, criteria-specific scoring rubrics.
Human EvaluationAn evaluation method in which human reviewers assess the quality of AI system outputs against defined criteria such as correctness, relevance, helpfulness, safety, or fluency, providing judgments that complement or validate automated evaluation metrics.