Prompt Experiment
A structured experiment that compares prompt variants to measure their effects on AI system behavior, quality, or performance.
Explore more about Experimentation
Related terms
A testing practice that evaluates prompts across defined inputs and criteria to assess the quality, consistency, and reliability of AI model outputs.
A/B TestingAn experimentation method that compares two variants under similar conditions to determine which performs better against predefined metrics.
Multivariate TestingAn experimentation method that tests multiple variables and their combinations simultaneously to measure their effects on an AI system or product.
Model ComparisonAn evaluation method that compares AI models on the same tasks, datasets, or metrics to identify differences in capabilities and performance.
Experiment TrackingThe process of recording and managing information about AI experiments, including datasets, models, hyperparameters, code versions, metrics, and outcomes to enable reproducibility and comparison.