Skip to content
AI Atlas
BenchmarkActivecategory · knowledge

MMLU

Hendrycks et al.github.com/hendrycks/test

57-subject multiple-choice questions

quality57

Updated 24 min ago · first seen 11 Sept 2026

bench_01M293SPDZHWVW43W1K6V407BF

Metric
accuracy · %
Direction
Results
0
Leader

Leaderboard 0 current results

Select models with +, then open Compare.

No benchmark results recorded

Results appear when a tier 1–3 source publishes them; we never copy scores without a source.