Updated 8 h ago · first seen 11 Sept 2026
model_01M29A5B9BGQQ3R4RATQT7G50Z
Overview
Identity
Identity block not returned by the API for this entity.
Openness
Openness not classified yet — no sourced evidence to place this model in the ontology.
Capabilities
Modalities
Modalities unavailable.
Capabilities
Tool calling
Unavailable
Structured output
Unavailable
Reasoning
Yes
Artificial Analysis · T2
Vision
Unavailable
Audio
Unavailable
Fine-tuning available
Unavailable
- Context window
Source:Artificial AnalysisT2observed 8 h agomedium
Benchmarks7
Compare with another model →Comparable same task and conditions · Partially comparable same task, conditions differ (effort, temperature, judge) · Not comparable different variant or metric
No benchmark results recorded
Timeline8
Full timeline →gpt-oss-120b-low scores 5.3% on Terminal-Bench
artificial_analysisgpt-oss-120b-low scores 13.86% on Terminal-Bench
artificial_analysisgpt-oss-120b-low scores 45.03% on τ²-bench
artificial_analysisgpt-oss-120b-low scores 58.3% on IFBench
artificial_analysisgpt-oss-120b-low scores 5.89% on Humanity's Last Exam
artificial_analysisgpt-oss-120b-low scores 67.17% on GPQA
artificial_analysisgpt-oss-120b-low scores 10.21 on Artificial Analysis Intelligence Index
artificial_analysis
Change history4
Opennessopenness1
Context windowcontext_length1
Aa median output tokens per secondmetric.aa_median_output_tokens_per_second1
Reasoningreasoning1
Claims are temporal and append-only: a new observation closes the previous claim (valid_to) instead of overwriting it. Conflicting claims from different sources are kept side by side and flagged — never averaged. Methodology →
Provenance
Attributed facts
4
Source tiers
T24
Freshest observation
8 h ago
Conflicts
None
Source documents 1
Tier 1 = official/primary, 2 = quality secondary, 3 = community, 4 = unverified. Every snapshot is archived; see all sources and the methodology.
Data quality (44/100) measures how well AI Atlas knows this entity — completeness, primary-source ratio, freshness, conflicts — never how good the model is. Methodology →