gpt-oss-120b
OpenAIdevelopers.openai.com/api/docs/models/gpt-oss-12
Most powerful open-weight model, fits into an H100 GPU
Updated 6 h ago · first seen 11 Sept 2026
model_01M293VC69K5D58EP5QCDWXVWY
Overview
Identity
Identity block not returned by the API for this entity.
Openness
Openness not classified yet — no sourced evidence to place this model in the ontology.
Key facts
- Release date
Source:Hugging Face Hub (public pages, model cards, papers)T2observed 8 h agomedium
- Knowledge cutoff
Source:OpenRouter public model & pricing listingT2observed 13 h agomedium
- Official page
Source:OpenAI Platform docsT1observed 13 h agohigh
- Model card
Source:Hugging Face Hub (public pages, model cards, papers)T2observed 13 h agomedium
- API model id
Source:OpenAI Platform docsT1observed 13 h agohigh
- Openrouter id
Source:OpenRouter public model & pricing listingT2observed 13 h agomedium
Architecture
- Architecture
Source:Hugging Face Hub (public pages, model cards, papers)T2observed 12 h agomedium
- Model type
Source:Hugging Face Hub (public pages, model cards, papers)T2observed 12 h agomedium
- Parameters
Source:Hugging Face Hub (public pages, model cards, papers)T2observed 13 h agomedium
- Tokenizer
Source:OpenRouter public model & pricing listingT2observed 13 h agomedium
- Weights dtype
Source:Hugging Face Hub (public pages, model cards, papers)T2observed 12 h agomedium
- File size
Source:Hugging Face Hub (public pages, model cards, papers)T2observed 12 h agomedium
- Library name
Source:Hugging Face Hub (public pages, model cards, papers)T2observed 12 h agomedium
- Pipeline tag
Source:Hugging Face Hub (public pages, model cards, papers)T2observed 13 h agomedium
- Hugging Face repo
Source:OpenRouter public model & pricing listingT2observed 13 h agomedium
Capabilities
Modalities
- Modalities
- text
- Input
- text
- Output
- text
Capabilities
Tool calling
Yes
OpenRouter public model & pricing listing · T2
Structured output
Yes
OpenRouter public model & pricing listing · T2
Reasoning
Yes
OpenRouter public model & pricing listing · T2
Vision
Unavailable
Audio
Unavailable
Fine-tuning available
Unavailable
- Context window
Source:OpenRouter public model & pricing listingT2observed 13 h agomedium
- Max output
Source:OpenRouter public model & pricing listingT2observed 13 h agomedium
- Knowledge cutoff
Source:OpenRouter public model & pricing listingT2observed 13 h agomedium
- Tokenizer
Source:OpenRouter public model & pricing listingT2observed 13 h agomedium
Benchmarks10
Compare with another model →Comparable same task and conditions · Partially comparable same task, conditions differ (effort, temperature, judge) · Not comparable different variant or metric
No benchmark results recorded
Providers & Pricing5
All offers in the price terminal →USD per 1M tokens as published by each provider (USD). Rows are append-only: every change is kept in the history below.
Price history
Output price · USD / 1M tokens 4 providers
- Fireworks AI
- OpenAI API
- GroqCloud
- Together AI
- Together AIfirst observed $0.6011 Sept 2026
- GroqCloudfirst observed $0.6011 Sept 2026
- OpenAI API$0.17 → $0.6011 Sept 2026
- OpenAI APIfirst observed $0.1711 Sept 2026
- Fireworks AIfirst observed $0.6011 Sept 2026
Input price · USD / 1M tokens 4 providers
- Fireworks AI
- OpenAI API
- GroqCloud
- Together AI
- Together AIfirst observed $0.1511 Sept 2026
- GroqCloudfirst observed $0.1511 Sept 2026
- OpenAI API$0.037 → $0.1511 Sept 2026
- OpenAI APIfirst observed $0.03711 Sept 2026
- Fireworks AIfirst observed $0.1511 Sept 2026
Hardware fit37
Assumptions (6)
- Estimated, not measured: weights = parameters × bytes/param × 1.15 runtime overhead.
- bytes/param: 4bit = 0.5, 8bit = 1.0, fp16 = 2.0 (uniform quantization, no per-layer exceptions).
- KV cache approximated at 0.5 GB per 8 192 tokens of context, independent of architecture (GQA/MLA models need less).
- A model 'fits' when the estimate is at most the device memory minus 2 GB reserved for the OS and framework.
- Mixture-of-experts models are estimated on total parameters (all experts must be resident); active parameters are ignored.
- Device memory uses the largest configuration when several are listed (e.g. Apple silicon tiers).
Lineage
Open in Graph →Papers1
- arXiv:2508.10925Active35
Timeline18
Full timeline →gpt-oss-120b scores 26% on SWE-bench Verified
swebench_leaderboardgpt-oss-120b scores 23.48% on Terminal-Bench
artificial_analysisgpt-oss-120b scores 26.22% on Terminal-Bench
artificial_analysisgpt-oss-120b scores 19.6% on Humanity's Last Exam
artificial_analysisgpt-oss-120b scores 12.35 on Artificial Analysis Intelligence Index
artificial_analysisTogether AI lists gpt-oss-120b at $0.15 in / $0.6 out per 1M tokens
together_pricinggpt-oss-120b: release date changed from 2025-08-04 to 2025-08-05
Release date4 Aug 2025→5 Aug 2025openroutergpt-oss-120b: release date changed from 2025-08-05 to 2025-08-04
Release date5 Aug 2025→4 Aug 2025huggingfaceGroqCloud lists gpt-oss-120b at $0.15 in / $0.6 out per 1M tokens
groq_pricingOpenAI API lists gpt-oss-120b at $0.15 in / $0.6 out per 1M tokens
openrouterOpenAI API lists gpt-oss-120b at $0.037 in / $0.17 out per 1M tokens
openrouterFireworks AI lists gpt-oss-120b at $0.15 in / $0.6 out per 1M tokens
fireworks_pricing
Change history41
Viewing AI Atlas as of 12 Mar 2026 — attributes exactly as the atlas knew them on that day; later corrections are not shown.
Back to today →gpt-oss-120b was not yet in AI Atlas on 12 Mar 2026
Release daterelease_date6
Opennessopenness1
Licenselicense1
Architecturearchitecture1
Parametersparameter_count1
Context windowcontext_length1
Max outputmax_output_tokens1
Knowledge cutoffknowledge_cutoff1
Modalitiesmodalities1
Input modalitiesmodalities_input1
Output modalitiesmodalities_output1
Tokenizertokenizer1
API model idapi_model_id1
File sizefile_size_gb1
Hugging Face repohf_repo1
Pipeline tagpipeline_tag1
Official pageofficial_url1
Model cardmodel_card_url1
Downloadsmetric.downloads1
Likesmetric.likes1
Descriptiondescription1
Gatedgated1
Groq model idgroq_model_id1
Hf inference providershf_inference_providers1
Last modifiedlast_modified1
Library namelibrary_name1
Aa median output tokens per secondmetric.aa_median_output_tokens_per_second1
Downloads all timemetric.downloads_all_time1
Model typemodel_type1
Openrouter idopenrouter_id1
Reasoningreasoning1
Structured outputstructured_output1
Supported parameterssupported_parameters1
Tool callingtool_calling1
Weights dtypeweights_dtype1
Claims are temporal and append-only: a new observation closes the previous claim (valid_to) instead of overwriting it. Conflicting claims from different sources are kept side by side and flagged — never averaged. Methodology →
Provenance
Attributed facts
37
Source tiers
T1T25 / 32
Freshest observation
6 h ago
Conflicts
None
Source documents 11
Tier 1 = official/primary, 2 = quality secondary, 3 = community, 4 = unverified. Every snapshot is archived; see all sources and the methodology.
Data quality (78/100) measures how well AI Atlas knows this entity — completeness, primary-source ratio, freshness, conflicts — never how good the model is. Methodology →