Skip to content
AI Atlas
FrameworkActive

lm-evaluation-harness

EleutherAIgithub.com/EleutherAI/lm-evaluation-harness

A framework for few-shot evaluation of language models.

Updated 40 min ago · first seen 11 Sept 2026

framework_01M294NC1A8C5XABTXDFE0X1BJ

Version
0.4.13
T2 · 45 min ago
Released
31 Aug 2026
T2 · 45 min ago
License
MIT
T2 · 45 min ago
Stars
13,954
T2 · 40 min ago

As of

Rewind the record: see this entity's attributes exactly as AI Atlas knew them on a given day.

Claim history · Releases

1 claims · 1 propertiesShow all properties

Releasesmetric.releases1

Claim history for Releases
ValueValid from → toStatusSourceConfidenceExtractor
19currentcurrentGitHub public repositories (HTML, releases.atom, raw files)T2mediumdeterministic

Claims are temporal and append-only: a new observation closes the previous claim (valid_to) instead of overwriting it. Conflicting claims from different sources are kept side by side and flagged — never averaged. Methodology →