DeepSeek V3.2
DeepSeekfamily · DeepSeekhuggingface.co/deepseek-ai/DeepSeek-V3.2
DeepSeek-V3.2 is a large language model designed to harmonize high computational efficiency with strong reasoning and agentic tool-use performance. It introduces DeepSeek Sparse Attention (DSA), a fine-grained sparse attention mechanism...
Updated 5 h ago · first seen 11 Sept 2026
model_01M294WW8H2JQJXMSPY7QTDVWF
Overview
Identity
- Canonical model
- Yesidentity confidence: mediumOne row per real model release. Artifacts (checkpoints, quantisations, conversions) and folded evaluation variants point here.
- Official checkpoints
- official_checkpoints = hf_repo identifiers carried by the model itself; artifacts are separate entities pointing here through canonical_id.
- Artifacts
- None recordedSeparate entities (checkpoint · quantization · conversion · packaging) pointing to this model through canonical_id.
- Provider deployments
- 2
- API aliases
- deepseek-v3-2-reasoningdeepseek/deepseek-v3.2Identifiers under which providers and evaluators refer to this model.
- Folded evaluation variants
- 0Effort / thinking variants (…-high, …-non-reasoning) are result configurations of this model, not separate models. Their old URLs redirect here.
Openness
Open weights— weights downloadable under MIT; commercial use allowed; redistribution allowed; derivatives allowed; 4 dimensions unknown.
Weights downloadable under a permissive or Creative Commons licence allowing commercial use; code or data may be missing.
Weights
Yes
Inference code
—
Training code
—
Training data
—
Dataset
—
Commercial use
Yes
Redistribution
Yes
Derivatives
Yes
Licence: MIT License (permissive · SPDX MIT · stated as “mit”)
dimensions marked null are unknown, not false
Key facts
- Release date
Source:OpenRouter public model & pricing listingT2observed 15 h agomedium
- Status
Source:DeepSeek — site & API docsT2observed 15 h agomediumLLM-extracted
- Version
Source:DeepSeek — site & API docsT2observed 15 h agomediumLLM-extracted
- Model card
Source:Hugging Face Hub (public pages, model cards, papers)T2observed 15 h agomedium
- Paper
Source:DeepSeek — site & API docsT2observed 15 h agomediumLLM-extracted
- Repository
Source:DeepSeek — site & API docsT2observed 15 h agomediumLLM-extracted
- Openrouter id
Source:OpenRouter public model & pricing listingT2observed 15 h agomedium
Architecture
- Architecture
Source:Hugging Face Hub (public pages, model cards, papers)T2observed 15 h agomedium
- Model type
Source:Hugging Face Hub (public pages, model cards, papers)T2observed 15 h agomedium
- Parameters
Source:Hugging Face Hub (public pages, model cards, papers)T2observed 15 h agomedium
- Tokenizer
Source:OpenRouter public model & pricing listingT2observed 15 h agomedium
- Weights dtype
Source:Hugging Face Hub (public pages, model cards, papers)T2observed 15 h agomedium
- File size
Source:Hugging Face Hub (public pages, model cards, papers)T2observed 15 h agomedium
- Library name
Source:Hugging Face Hub (public pages, model cards, papers)T2observed 15 h agomedium
- Pipeline tag
Source:Hugging Face Hub (public pages, model cards, papers)T2observed 15 h agomedium
- Hugging Face repo
Source:OpenRouter public model & pricing listingT2observed 15 h agomedium
Capabilities
Modalities
- Modalities
- text
- Input
- text
- Output
- text
Capabilities
Tool calling
Yes
OpenRouter public model & pricing listing · T2
Structured output
Yes
OpenRouter public model & pricing listing · T2
Reasoning
Yes
OpenRouter public model & pricing listing · T2
Vision
Unavailable
Audio
Unavailable
Fine-tuning available
Unavailable
- Context window
Source:OpenRouter public model & pricing listingT2observed 12 h agomedium
- Max output
Source:OpenRouter public model & pricing listingT2observed 15 h agomedium
- Tokenizer
Source:OpenRouter public model & pricing listingT2observed 15 h agomedium
Benchmarks9
Compare with another model →Comparable same task and conditions · Partially comparable same task, conditions differ (effort, temperature, judge) · Not comparable different variant or metric
Current rows only, grouped by benchmark → canonical metric → comparability group (task configuration). Effort variants folded into this model appear as rows of the same group. 9 current rows in total. “vs leader” compares with the current leader of the benchmark's primary group only; other groups are not directly comparable. Comparability rules →
Providers & Pricing2
All offers in the price terminal →USD per 1M tokens as published by each provider; native units (per-request fees, flex/priority tiers) are kept verbatim. Rows are append-only — every price change is kept in the history below. Cost of a workload →
Price history
Output price · USD / 1M tokens 2 providers
- DeepSeek API
- OpenRouter
- OpenRouterfirst observed $0.4012 Sept 2026
- DeepSeek APIfirst observed $0.4011 Sept 2026
Input price · USD / 1M tokens 2 providers
- DeepSeek API
- OpenRouter
- OpenRouterfirst observed $0.26912 Sept 2026
- DeepSeek APIfirst observed $0.26911 Sept 2026
Hardware fit37
Assumptions (7)
- Estimated, not measured: weights = parameters × bytes/param × 1.15 runtime overhead (or the observed artifact file size when one is recorded).
- bytes/param: 4bit = 0.5, 8bit = 1.0, fp16 = 2.0 (uniform quantization, no per-layer exceptions).
- KV cache: 2 × layers × kv_heads × head_dim × 2 bytes × context × batch when the architecture is known; otherwise 0.5 GB per 8 192 tokens (× batch), independent of architecture (GQA/MLA models need less).
- A model 'fits' when the estimate is at most the device memory minus 2 GB reserved for the OS and framework.
- Mixture-of-experts models are estimated on total parameters (all experts must be resident); active parameters are ignored.
- Device memory uses the largest configuration when several are listed (e.g. Apple silicon tiers).
- Multi-GPU: device memories are summed; interconnect bandwidth, tensor-parallel replication and pipeline bubbles are not modelled.
Lineage
Open in Graph →- ancestor: DeepSeek-V3.2-Exp-Base
Versions & Artifacts0
Version history
Context window2 changes
11 Sept 2026→11 Sept 2026→12 Sept 2026current
Status1 change
11 Sept 2026→11 Sept 2026
Licensefirst observation only
11 Sept 2026current
Max outputfirst observation only
11 Sept 2026current
Opennessfirst observation only
11 Sept 2026current
Parametersfirst observation only
11 Sept 2026current
Each hop is a claim: click a value for its source, tier and observation time. Nothing is overwritten — a new observation closes the previous claim.
Artifacts 0
No artifact (checkpoint, quantisation, conversion or packaging) points to this model yet.
Timeline2
Full timeline →OpenRouter lists DeepSeek V3.2 at $0.269 in / $0.4 out per 1M tokens
openrouterDeepSeek V3.2: context length changed from 128000 to 163840
Context window128K tokens→163.8K tokensopenrouter
Change history
Statusstatus2
Claims are temporal and append-only: a new observation closes the previous claim (valid_to) instead of overwriting it. Conflicting claims from different sources are kept side by side and flagged — never averaged. Methodology →
Provenance
Attributed facts
46
Source tiers
T246
Freshest observation
5 h ago
Conflicts
None
Source documents 9
Tier 1 = official/primary, 2 = quality secondary, 3 = community, 4 = unverified. Every snapshot is archived; see all sources and the methodology.
Data quality (72/100) measures how well AI Atlas knows this entity — completeness, primary-source ratio, freshness, conflicts — never how good the model is. Methodology →