Skip to content
AI Atlas
BenchmarkActivecategory · multimodal

MMMU

mmmu-benchmark.github.io

college-level multimodal understanding

quality57

Updated 46 min ago · first seen 11 Sept 2026

bench_01M293SPFEJ607A0J5NZNEE082

Metric
accuracy · %
Direction
Results
0
Leader

As of

Rewind the record: see this entity's attributes exactly as AI Atlas knew them on a given day.

Claim history · Category

1 claims · 1 propertiesShow all properties

Categorycategory1

Claim history for Category
ValueValid from → toStatusSourceConfidenceExtractor
multimodalcurrentcurrentAI Atlas curated registry (YAML, versioned in git, every entry carries its source URL)T2mediumcurated

Claims are temporal and append-only: a new observation closes the previous claim (valid_to) instead of overwriting it. Conflicting claims from different sources are kept side by side and flagged — never averaged. Methodology →