MMBench
MMBenchAlso called: MM-Bench
A benchmark for evaluating multimodal large language models across a broad range of vision-language tasks and capabilities.
Explore more about Benchmarks
Related terms
A multimodal benchmark for evaluating models on complex tasks spanning diverse academic disciplines and requiring both visual and textual reasoning.
MMMU-ProA challenging multimodal benchmark designed to evaluate advanced reasoning across diverse academic and professional tasks.
DocVQADocVQAA benchmark for evaluating an AI model's ability to answer questions by understanding and extracting information from document images.
Benchmark RunA single execution of a benchmark used to measure and record a model or system's performance under defined conditions.