Skip to content
AI Atlas
ModelDeprecatedOpen weights
data quality57

Updated 28 min ago · first seen 11 Sept 2026

model_01M2950KPSQD8AH04S24H4CFHQ

Overview

Identity

Canonical model
Yesidentity confidence: highOne row per real model release. Artifacts (checkpoints, quantisations, conversions) and folded evaluation variants point here.
Official checkpoints
official_checkpoints = hf_repo identifiers carried by the model itself; artifacts are separate entities pointing here through canonical_id.
Artifacts
None recordedSeparate entities (checkpoint · quantization · conversion · packaging) pointing to this model through canonical_id.
Provider deployments
None recorded
API aliases
grok-2-1212Identifiers under which providers and evaluators refer to this model.
Folded evaluation variants
0Effort / thinking variants (…-high, …-non-reasoning) are result configurations of this model, not separate models. Their old URLs redirect here.

Openness

Open weightsweights downloadable; 7 dimensions unknown.

Weights downloadable under a permissive or Creative Commons licence allowing commercial use; code or data may be missing.

  • Weights

    Yes

  • Inference code

  • Training code

  • Training data

  • Dataset

  • Commercial use

  • Redistribution

  • Derivatives

dimensions marked null are unknown, not false

Key facts

Release date

Source:Hugging Face Hub (public pages, model cards, papers)T2observed 15 h agomedium

Status

Source:Artificial AnalysisT2observed 14 h agomedium

Model card

Source:Hugging Face Hub (public pages, model cards, papers)T2observed 16 h agomedium

Architecture

Architecture

Source:Hugging Face Hub (public pages, model cards, papers)T2observed 15 h agomedium

Model type

Source:Hugging Face Hub (public pages, model cards, papers)T2observed 15 h agomedium

Hugging Face repo

Source:Hugging Face Hub (public pages, model cards, papers)T2observed 16 h agomedium

Capabilities

Modalities

Modalities unavailable.

Capabilities

  • Tool calling

    Unavailable

  • Structured output

    Unavailable

  • Reasoning

    No

    Artificial Analysis · T2

  • Vision

    Unavailable

  • Audio

    Unavailable

  • Fine-tuning available

    Unavailable

Context window

Source:Artificial AnalysisT2observed 14 h agomedium

Comparable same task and conditions · Partially comparable same task, conditions differ (effort, temperature, judge) · Not comparable different variant or metric

Benchmark results grouped by comparability group
Benchmark · groupBest scoreTrustConfigurationResultsvs leaderEvaluatedSource
Artificial Analysis Intelligence Indexcomposite · indexIndependentreasoningoffversion4.3conditions differ across rows → partially comparable2−46.3vs Claude Fable 5.1obs. 12 Sept 2026artificialanalysis.aiT2
Humanity's Last Examknowledge · accuracy · evaluator=Artificial AnalysisIndependentevaluatorArtificial Analysisreasoningoffconditions differ across rows → partially comparable2−56.1 ptvs Claude Fable 5.1obs. 12 Sept 2026artificialanalysis.aiT2
GPQA Diamondreasoning · accuracy · variant=Diamond · evaluator=Artificial AnalysisIndependentevaluatorArtificial AnalysisvariantDiamondreasoningoffconditions differ across rows → partially comparable1non-primary groupobs. 12 Sept 2026artificialanalysis.aiT2
GPQA Diamondreasoning · accuracy · variant=GPQA Diamond · evaluator=Artificial AnalysisIndependentevaluatorArtificial AnalysisvariantGPQA Diamond1−45.3 ptvs gpt-6-astraobs. 11 Sept 2026artificialanalysis.aiT2

Current rows only, grouped by benchmark → canonical metric → comparability group (task configuration). Effort variants folded into this model appear as rows of the same group. 6 current rows in total. “vs leader” compares with the current leader of the benchmark's primary group only; other groups are not directly comparable. Comparability rules →

Versions & Artifacts0

Version history

Context windowfirst observation only

11 Sept 2026current

Opennessfirst observation only

11 Sept 2026current

Statusfirst observation only

11 Sept 2026current

Each hop is a claim: click a value for its source, tier and observation time. Nothing is overwritten — a new observation closes the previous claim.

Artifacts 0

No artifact (checkpoint, quantisation, conversion or packaging) points to this model yet.

Change history

Temporal, append-only claims: a new observation closes the previous claim instead of overwriting it. Rewind the record with the as-of picker.
1 claims · 1 propertiesShow all properties

Reasoningreasoning1

Claim history for Reasoning
ValueValid from → toStatusSourceConfidenceExtractor
NocurrentcurrentArtificial AnalysisT2mediumdeterministic

Claims are temporal and append-only: a new observation closes the previous claim (valid_to) instead of overwriting it. Conflicting claims from different sources are kept side by side and flagged — never averaged. Methodology →

Provenance

Attributed facts

21

Source tiers

T221

Freshest observation

59 min ago

Conflicts

None

Source documents 4

Source documents
SourceDocumentTypeTierLast observedSnapshots
Hugging Face Hub (public pages, model cards, papers)huggingface.co/xai-org/grok-2 model_pageT2· Quality secondary59 min ago5
Hugging Face Hub (public pages, model cards, papers)huggingface.co/models?author=xai-org&p=0&sort=downloads listingT2· Quality secondary2 h ago8
Artificial Analysisartificialanalysis.ai/leaderboards/models leaderboardT2· Quality secondary6 h ago2
Hugging Face Hub (public pages, model cards, papers)huggingface.co/xai-org/grok-2/raw/main/README.md model_cardT2· Quality secondary9 h ago1

Tier 1 = official/primary, 2 = quality secondary, 3 = community, 4 = unverified. Every snapshot is archived; see all sources and the methodology.

Data quality (57/100) measures how well AI Atlas knows this entity — completeness, primary-source ratio, freshness, conflicts — never how good the model is. Methodology →