Skip to content
AI Atlas
FrameworkActive

lm-evaluation-harness

EleutherAIgithub.com/EleutherAI/lm-evaluation-harness

A framework for few-shot evaluation of language models.

Updated 39 min ago · first seen 11 Sept 2026

framework_01M294NC1A8C5XABTXDFE0X1BJ

Version
0.4.13
T2 · 44 min ago
Released
31 Aug 2026
T2 · 44 min ago
License
MIT
T2 · 44 min ago
Stars
13,954
T2 · 39 min ago

As of

Rewind the record: see this entity's attributes exactly as AI Atlas knew them on a given day.

Claim history · Latest version

1 claims · 1 propertiesShow all properties

Latest versionlatest_version1

Claim history for Latest version
ValueValid from → toStatusSourceConfidenceExtractor
0.4.13currentcurrentGitHub public repositories (HTML, releases.atom, raw files)T2mediumdeterministic

Claims are temporal and append-only: a new observation closes the previous claim (valid_to) instead of overwriting it. Conflicting claims from different sources are kept side by side and flagged — never averaged. Methodology →