Nemotron 3 Ultra
NVIDIA Nemotron 3 Ultra is an open frontier-reasoning and orchestration model from NVIDIA, with 55B active parameters out of 550B total (MoE). Built on a hybrid Transformer-Mamba mixture-of-experts architecture, it...
Updated 5 h ago · first seen 11 Sept 2026
model_01M294WVXG0ZTDNK5X5S7V5438
Overview
Identity
Identity block not returned by the API for this entity.
Openness
Openness not classified yet — no sourced evidence to place this model in the ontology.
Key facts
- Release date
Source:OpenRouter public model & pricing listingT2observed 12 h agomedium
- Openrouter id
Source:OpenRouter public model & pricing listingT2observed 12 h agomedium
Architecture
- Hugging Face repo
Source:OpenRouter public model & pricing listingT2observed 12 h agomedium
Capabilities
Modalities
- Modalities
- text
- Input
- text
- Output
- text
Capabilities
Tool calling
Yes
OpenRouter public model & pricing listing · T2
Structured output
Yes
OpenRouter public model & pricing listing · T2
Reasoning
Yes
OpenRouter public model & pricing listing · T2
Vision
Unavailable
Audio
Unavailable
Fine-tuning available
Unavailable
- Context window
Source:OpenRouter public model & pricing listingT2observed 12 h agomedium
- Max output
Source:OpenRouter public model & pricing listingT2observed 12 h agomedium
Benchmarks17
Compare with another model →Comparable same task and conditions · Partially comparable same task, conditions differ (effort, temperature, judge) · Not comparable different variant or metric
No benchmark results recorded
Providers & Pricing2
All offers in the price terminal →USD per 1M tokens as published by each provider (USD). Rows are append-only: every change is kept in the history below.
Price history
Output price · USD / 1M tokens 1 provider
- NVIDIA NIM / build.nvidia.com
- NVIDIA NIM / build.nvidia.com$3.13 → $011 Sept 2026
- NVIDIA NIM / build.nvidia.comfirst observed $3.1311 Sept 2026
Input price · USD / 1M tokens 1 provider
- NVIDIA NIM / build.nvidia.com
- NVIDIA NIM / build.nvidia.com$0.625 → $011 Sept 2026
- NVIDIA NIM / build.nvidia.comfirst observed $0.62511 Sept 2026
Timeline20
Full timeline →Nemotron 3 Ultra scores 73.445% on LiveBench
livebench_leaderboardNemotron 3 Ultra scores 70.811% on LiveBench
livebench_leaderboardNemotron 3 Ultra scores 54.457% on LiveBench
livebench_leaderboardNemotron 3 Ultra scores 88.663% on LiveBench
livebench_leaderboardNemotron 3 Ultra scores 38.737% on LiveBench
livebench_leaderboardNemotron 3 Ultra scores 70.698% on LiveBench
livebench_leaderboardNemotron 3 Ultra scores 74.697% on LiveBench
livebench_leaderboardNemotron 3 Ultra scores 67.358% on LiveBench
livebench_leaderboardNemotron 3 Ultra scores 36.36% on Terminal-Bench
artificial_analysisNemotron 3 Ultra scores 53.93% on Terminal-Bench
artificial_analysisNemotron 3 Ultra scores 0.51% on Terminal-Bench
artificial_analysisNemotron 3 Ultra scores 83.33% on τ²-bench
artificial_analysisNemotron 3 Ultra scores 81.36% on IFBench
artificial_analysisNemotron 3 Ultra scores 40.28% on SciCode
artificial_analysisNemotron 3 Ultra scores 28.41% on Humanity's Last Exam
artificial_analysisNemotron 3 Ultra scores 86.67% on GPQA
artificial_analysisNemotron 3 Ultra scores 23.41 on Artificial Analysis Intelligence Index
artificial_analysisNVIDIA NIM / build.nvidia.com lists Nemotron 3 Ultra at $0 in / $0 out per 1M tokens
openrouterNVIDIA NIM / build.nvidia.com lists Nemotron 3 Ultra at $0.625 in / $3.125 out per 1M tokens
openrouter
Change history16
Viewing AI Atlas as of 12 Sept 2026 — attributes exactly as the atlas knew them on that day; later corrections are not shown.
Back to today →Attributes as of 12 Sept 2026 20 claims in force
- Release date
- 4 Jun 2026
- Openness
- open-weights
- Context window
- 262.1K tokens
- Max output
- 32.8K tokens
- Modalities
- text
- Input modalities
- text
- Output modalities
- text
- Hugging Face repo
- nvidia/NVIDIA-Nemotron-3-Ultra-550B-A55B-BF16
- Aa context window
- 262,144
- Aa openness
- open-weights
- Livebench hf link
- https://huggingface.co/nvidia/NVIDIA-Nemotron-3-Ultra-550B-A55B-NVFP4
- Aa median output tokens per second
- 165.5
- Openrouter id
- nvidia/nemotron-3-ultra-550b-a55b
- Openrouter listed at
- 4 Jun 2026
- Reasoning
- Yes
- Structured output
- Yes
- Supported parameters
- frequency_penalty, include_reasoning, logit_bias, max_tokens, min_p, presence_penalty, reasoning, reasoning_effort, repetition_penalty, response_format, seed, stop, structured_outputs, temperature, tool_choice, tools, top_k, top_p
- Tool calling
- Yes
- Weights available
- Yes
Release daterelease_date1
Opennessopenness1
Context windowcontext_length1
Max outputmax_output_tokens1
Modalitiesmodalities1
Input modalitiesmodalities_input1
Output modalitiesmodalities_output1
Hugging Face repohf_repo1
Descriptiondescription1
Livebench hf linklivebench_hf_link1
Aa median output tokens per secondmetric.aa_median_output_tokens_per_second1
Openrouter idopenrouter_id1
Reasoningreasoning1
Structured outputstructured_output1
Supported parameterssupported_parameters1
Tool callingtool_calling1
Claims are temporal and append-only: a new observation closes the previous claim (valid_to) instead of overwriting it. Conflicting claims from different sources are kept side by side and flagged — never averaged. Methodology →
Provenance
Attributed facts
17
Source tiers
T217
Freshest observation
5 h ago
Conflicts
None
Source documents 3
Tier 1 = official/primary, 2 = quality secondary, 3 = community, 4 = unverified. Every snapshot is archived; see all sources and the methodology.
Data quality (60/100) measures how well AI Atlas knows this entity — completeness, primary-source ratio, freshness, conflicts — never how good the model is. Methodology →