MMMU
MMMUAlso called: Massive Multi-discipline Multimodal Understanding
A multimodal benchmark for evaluating models on complex tasks spanning diverse academic disciplines and requiring both visual and textual reasoning.
Explore more about Benchmarks
Related terms
A challenging multimodal benchmark designed to evaluate advanced reasoning across diverse academic and professional tasks.
MMBenchMMBenchA benchmark for evaluating multimodal large language models across a broad range of vision-language tasks and capabilities.
DocVQADocVQAA benchmark for evaluating an AI model's ability to answer questions by understanding and extracting information from document images.
MMLUMMLUA benchmark for evaluating language models across diverse subjects spanning humanities, social sciences, STEM, and professional knowledge.