Skip to content
AI Atlas
BenchmarkActivecategory · coding

HumanEval

github.com/openai/human-eval

Python function synthesis from docstrings

quality57

Updated 24 min ago · first seen 11 Sept 2026

bench_01M293SPECN73H13J38KYN4DNN

Metric
pass@1 · %
Direction
Results
0
Leader

Leaderboard 0 current results

Select models with +, then open Compare.

No benchmark results recorded

Results appear when a tier 1–3 source publishes them; we never copy scores without a source.