Reference-free Evaluation
An evaluation approach that assesses model output quality without relying on ground-truth reference answers or golden datasets.
Explore more about Evaluation Methods
Related terms
An evaluation method that measures model output quality by comparing generated responses against ground-truth reference data or golden datasets.
LLM-as-a-JudgeAn evaluation method that uses a large language model to assess the quality, correctness, or other attributes of AI-generated outputs.
Automated EvaluationAn evaluation method that uses software, benchmarks, metrics, or models to assess the quality, correctness, or performance of AI systems without manual review.
Model-based EvaluationAn evaluation approach that uses another trained model to assess the quality, correctness, safety, or other properties of an AI system's outputs.
Rubric-based EvaluationAn evaluation approach that assesses AI outputs against structured, criteria-specific scoring rubrics.