DeepSeek V4 Flash 0423
DeepSeekfamily · DeepSeekhuggingface.co/deepseek-ai/DeepSeek-V4-Flash
DeepSeek V4 Flash is an efficiency-optimized Mixture-of-Experts model from DeepSeek with 284B total parameters and 13B activated parameters, supporting a 1M-token context window. It is designed for fast inference and...
Updated 4 h ago · first seen 11 Sept 2026
model_01M294WW0M9SRBDTMYPH56GG7K
Overview
Identity
- Canonical model
- Yesidentity confidence: mediumOne row per real model release. Artifacts (checkpoints, quantisations, conversions) and folded evaluation variants point here.
- Official checkpoints
- official_checkpoints = hf_repo identifiers carried by the model itself; artifacts are separate entities pointing here through canonical_id.
- Artifacts
- None recordedSeparate entities (checkpoint · quantization · conversion · packaging) pointing to this model through canonical_id.
- Provider deployments
- 2
- API aliases
- deepseek/deepseek-v4-flashIdentifiers under which providers and evaluators refer to this model.
- Folded evaluation variants
- 0Effort / thinking variants (…-high, …-non-reasoning) are result configurations of this model, not separate models. Their old URLs redirect here.
Openness
Open weights— weights downloadable under MIT; commercial use allowed; redistribution allowed; derivatives allowed; 4 dimensions unknown.
Weights downloadable under a permissive or Creative Commons licence allowing commercial use; code or data may be missing.
Weights
Yes
Inference code
—
Training code
—
Training data
—
Dataset
—
Commercial use
Yes
Redistribution
Yes
Derivatives
Yes
Licence: MIT License (permissive · SPDX MIT · stated as “mit”)
dimensions marked null are unknown, not false
Key facts
- Release date
Source:Hugging Face Hub (public pages, model cards, papers)T2observed 5 d agomedium
- Model card
Source:Hugging Face Hub (public pages, model cards, papers)T2observed 5 d agomedium
- Openrouter id
Source:OpenRouter public model & pricing listingT2observed 5 d agomedium
Architecture
- Architecture
Source:Hugging Face Hub (public pages, model cards, papers)T2observed 5 d agomedium
- Model type
Source:Hugging Face Hub (public pages, model cards, papers)T2observed 5 d agomedium
- Parameters
Source:Hugging Face Hub (public pages, model cards, papers)T2observed 5 d agomedium
- Tokenizer
Source:OpenRouter public model & pricing listingT2observed 5 d agomedium
- Weights dtype
Source:Hugging Face Hub (public pages, model cards, papers)T2observed 5 d agomedium
- File size
Source:Hugging Face Hub (public pages, model cards, papers)T2observed 5 d agomedium
- Library name
Source:Hugging Face Hub (public pages, model cards, papers)T2observed 5 d agomedium
- Pipeline tag
Source:Hugging Face Hub (public pages, model cards, papers)T2observed 5 d agomedium
- Hugging Face repo
Source:OpenRouter public model & pricing listingT2observed 5 d agomedium
Capabilities
Modalities
- Modalities
- text
- Input
- text
- Output
- text
Capabilities
Tool calling
Yes
OpenRouter public model & pricing listing · T2
Structured output
Yes
OpenRouter public model & pricing listing · T2
Reasoning
Yes
OpenRouter public model & pricing listing · T2
Vision
Unavailable
Audio
Unavailable
Fine-tuning available
Unavailable
- Context window
Source:OpenRouter public model & pricing listingT2observed 5 d agomedium
- Max output
Source:OpenRouter public model & pricing listingT2observed 5 d agomedium
- Tokenizer
Source:OpenRouter public model & pricing listingT2observed 5 d agomedium
Providers & Pricing2
All offers in the price terminal →USD per 1M tokens as published by each provider; native units (per-request fees, flex/priority tiers) are kept verbatim. Rows are append-only — every price change is kept in the history below. Cost of a workload →
Price history
Output price · USD / 1M tokens 2 providers
- DeepSeek API
- OpenRouter
- OpenRouter$0.177 → $0.17416 Sept 2026
- OpenRouter$0.153 → $0.17715 Sept 2026
- OpenRouter$0.156 → $0.15315 Sept 2026
- OpenRouter$0.162 → $0.15615 Sept 2026
- OpenRouter$0.168 → $0.16215 Sept 2026
- OpenRouter$0.174 → $0.16815 Sept 2026
- OpenRouter$0.177 → $0.17415 Sept 2026
- OpenRouter$0.153 → $0.17715 Sept 2026
- OpenRouter$0.156 → $0.15314 Sept 2026
- OpenRouter$0.159 → $0.15614 Sept 2026
- OpenRouter$0.168 → $0.15914 Sept 2026
- OpenRouter$0.174 → $0.16814 Sept 2026
Input price · USD / 1M tokens 2 providers
- DeepSeek API
- OpenRouter
- OpenRouter$0.089 → $0.08716 Sept 2026
- OpenRouter$0.076 → $0.08915 Sept 2026
- OpenRouter$0.078 → $0.07615 Sept 2026
- OpenRouter$0.081 → $0.07815 Sept 2026
- OpenRouter$0.084 → $0.08115 Sept 2026
- OpenRouter$0.087 → $0.08415 Sept 2026
- OpenRouter$0.089 → $0.08715 Sept 2026
- OpenRouter$0.076 → $0.08915 Sept 2026
- OpenRouter$0.078 → $0.07614 Sept 2026
- OpenRouter$0.079 → $0.07814 Sept 2026
- OpenRouter$0.084 → $0.07914 Sept 2026
- OpenRouter$0.087 → $0.08414 Sept 2026
Hardware fit37
Assumptions (7)
- Estimated, not measured: weights = parameters × bytes/param × 1.15 runtime overhead (or the observed artifact file size when one is recorded).
- bytes/param: 4bit = 0.5, 8bit = 1.0, fp16 = 2.0 (uniform quantization, no per-layer exceptions).
- KV cache: 2 × layers × kv_heads × head_dim × 2 bytes × context × batch when the architecture is known; otherwise 0.5 GB per 8 192 tokens (× batch), independent of architecture (GQA/MLA models need less).
- A model 'fits' when the estimate is at most the device memory minus 2 GB reserved for the OS and framework.
- Mixture-of-experts models are estimated on total parameters (all experts must be resident); active parameters are ignored.
- Device memory uses the largest configuration when several are listed (e.g. Apple silicon tiers).
- Multi-GPU: device memories are summed; interconnect bandwidth, tensor-parallel replication and pipeline bubbles are not modelled.
Versions & Artifacts0
Version history
Context windowfirst observation only
11 Sept 2026current
Licensefirst observation only
11 Sept 2026current
Max outputfirst observation only
11 Sept 2026current
Opennessfirst observation only
11 Sept 2026current
Parametersfirst observation only
11 Sept 2026current
Each hop is a claim: click a value for its source, tier and observation time. Nothing is overwritten — a new observation closes the previous claim.
Artifacts 0
No artifact (checkpoint, quantisation, conversion or packaging) points to this model yet.
Papers1
- arXiv:2606.19348Active35
Timeline30
Full timeline →OpenRouter changed pricing for DeepSeek V4 Flash 0423: $0.08092 in / $0.16184 out per 1M tokens → $0.07784 in / $0.15568 out per 1M tokens
$0.081 in / $0.162 out→$0.078 in / $0.156 outopenrouterOpenRouter changed pricing for DeepSeek V4 Flash 0423: $0.084 in / $0.168 out per 1M tokens → $0.08092 in / $0.16184 out per 1M tokens
$0.084 in / $0.168 out→$0.081 in / $0.162 outopenrouterOpenRouter changed pricing for DeepSeek V4 Flash 0423: $0.0763 in / $0.1526 out per 1M tokens → $0.088606 in / $0.177212 out per 1M tokens
$0.076 in / $0.153 out→$0.089 in / $0.177 outopenrouterOpenRouter changed pricing for DeepSeek V4 Flash 0423: $0.07784 in / $0.15568 out per 1M tokens → $0.0763 in / $0.1526 out per 1M tokens
$0.078 in / $0.156 out→$0.076 in / $0.153 outopenrouterOpenRouter changed pricing for DeepSeek V4 Flash 0423: $0.07938 in / $0.15876 out per 1M tokens → $0.07784 in / $0.15568 out per 1M tokens
$0.079 in / $0.159 out→$0.078 in / $0.156 outopenrouterOpenRouter changed pricing for DeepSeek V4 Flash 0423: $0.084 in / $0.168 out per 1M tokens → $0.07938 in / $0.15876 out per 1M tokens
$0.084 in / $0.168 out→$0.079 in / $0.159 outopenrouterOpenRouter changed pricing for DeepSeek V4 Flash 0423: $0.08708 in / $0.17416 out per 1M tokens → $0.084 in / $0.168 out per 1M tokens
$0.087 in / $0.174 out→$0.084 in / $0.168 outopenrouterOpenRouter changed pricing for DeepSeek V4 Flash 0423: $0.088606 in / $0.177212 out per 1M tokens → $0.08708 in / $0.17416 out per 1M tokens
$0.089 in / $0.177 out→$0.087 in / $0.174 outopenrouterOpenRouter changed pricing for DeepSeek V4 Flash 0423: $0.089866 in / $0.179732 out per 1M tokens → $0.088606 in / $0.177212 out per 1M tokens
$0.09 in / $0.18 out→$0.089 in / $0.177 outopenrouterOpenRouter changed pricing for DeepSeek V4 Flash 0423: $0.04648 in / $0.09296 out per 1M tokens → $0.089866 in / $0.179732 out per 1M tokens
$0.046 in / $0.093 out→$0.09 in / $0.18 outopenrouterOpenRouter changed pricing for DeepSeek V4 Flash 0423: $0.04704 in / $0.09408 out per 1M tokens → $0.04648 in / $0.09296 out per 1M tokens
$0.047 in / $0.094 out→$0.046 in / $0.093 outopenrouterOpenRouter changed pricing for DeepSeek V4 Flash 0423: $0.04732 in / $0.09464 out per 1M tokens → $0.04704 in / $0.09408 out per 1M tokens
$0.047 in / $0.095 out→$0.047 in / $0.094 outopenrouterOpenRouter changed pricing for DeepSeek V4 Flash 0423: $0.0476 in / $0.0952 out per 1M tokens → $0.04732 in / $0.09464 out per 1M tokens
$0.048 in / $0.095 out→$0.047 in / $0.095 outopenrouterOpenRouter changed pricing for DeepSeek V4 Flash 0423: $0.04816 in / $0.09632 out per 1M tokens → $0.0476 in / $0.0952 out per 1M tokens
$0.048 in / $0.096 out→$0.048 in / $0.095 outopenrouterOpenRouter changed pricing for DeepSeek V4 Flash 0423: $0.04844 in / $0.09688 out per 1M tokens → $0.04816 in / $0.09632 out per 1M tokens
$0.048 in / $0.097 out→$0.048 in / $0.096 outopenrouterOpenRouter changed pricing for DeepSeek V4 Flash 0423: $0.04872 in / $0.09744 out per 1M tokens → $0.04844 in / $0.09688 out per 1M tokens
$0.049 in / $0.097 out→$0.048 in / $0.097 outopenrouterOpenRouter changed pricing for DeepSeek V4 Flash 0423: $0.04928 in / $0.09856 out per 1M tokens → $0.04872 in / $0.09744 out per 1M tokens
$0.049 in / $0.099 out→$0.049 in / $0.097 outopenrouterOpenRouter changed pricing for DeepSeek V4 Flash 0423: $0.049 in / $0.098 out per 1M tokens → $0.04928 in / $0.09856 out per 1M tokens
$0.049 in / $0.098 out→$0.049 in / $0.099 outopenrouterOpenRouter changed pricing for DeepSeek V4 Flash 0423: $0.04928 in / $0.09856 out per 1M tokens → $0.049 in / $0.098 out per 1M tokens
$0.049 in / $0.099 out→$0.049 in / $0.098 outopenrouterOpenRouter changed pricing for DeepSeek V4 Flash 0423: $0.04956 in / $0.09912 out per 1M tokens → $0.04928 in / $0.09856 out per 1M tokens
$0.05 in / $0.099 out→$0.049 in / $0.099 outopenrouterOpenRouter changed pricing for DeepSeek V4 Flash 0423: $0.04984 in / $0.09968 out per 1M tokens → $0.04956 in / $0.09912 out per 1M tokens
$0.05 in / $0.10 out→$0.05 in / $0.099 outopenrouterOpenRouter changed pricing for DeepSeek V4 Flash 0423: $0.0651 in / $0.1302 out per 1M tokens → $0.04984 in / $0.09968 out per 1M tokens
$0.065 in / $0.13 out→$0.05 in / $0.10 outopenrouterOpenRouter changed pricing for DeepSeek V4 Flash 0423: $0.06538 in / $0.13076 out per 1M tokens → $0.0651 in / $0.1302 out per 1M tokens
$0.065 in / $0.131 out→$0.065 in / $0.13 outopenrouterOpenRouter changed pricing for DeepSeek V4 Flash 0423: $0.06566 in / $0.13132 out per 1M tokens → $0.06538 in / $0.13076 out per 1M tokens
$0.066 in / $0.131 out→$0.065 in / $0.131 outopenrouterOpenRouter changed pricing for DeepSeek V4 Flash 0423: $0.06594 in / $0.13188 out per 1M tokens → $0.06566 in / $0.13132 out per 1M tokens
$0.066 in / $0.132 out→$0.066 in / $0.131 outopenrouterOpenRouter changed pricing for DeepSeek V4 Flash 0423: $0.06622 in / $0.13244 out per 1M tokens → $0.06594 in / $0.13188 out per 1M tokens
$0.066 in / $0.132 out→$0.066 in / $0.132 outopenrouterOpenRouter changed pricing for DeepSeek V4 Flash 0423: $0.0665 in / $0.133 out per 1M tokens → $0.06622 in / $0.13244 out per 1M tokens
$0.067 in / $0.133 out→$0.066 in / $0.132 outopenrouterOpenRouter changed pricing for DeepSeek V4 Flash 0423: $0.06678 in / $0.13356 out per 1M tokens → $0.0665 in / $0.133 out per 1M tokens
$0.067 in / $0.134 out→$0.067 in / $0.133 outopenrouterOpenRouter lists DeepSeek V4 Flash 0423 at $0.06678 in / $0.13356 out per 1M tokens
openrouterDeepSeek API changed pricing for DeepSeek V4 Flash 0423: $0.06706 in / $0.13412 out per 1M tokens → $0.06678 in / $0.13356 out per 1M tokens
$0.067 in / $0.134 out→$0.067 in / $0.134 outopenrouter
Change history69
Viewing AI Atlas as of 16 Aug 2026 — attributes exactly as the atlas knew them on that day; later corrections are not shown.
Back to today →DeepSeek V4 Flash 0423 was not yet in AI Atlas on 16 Aug 2026
Release daterelease_date6
Opennessopenness1
Licenselicense1
Architecturearchitecture1
Parametersparameter_count1
Context windowcontext_length1
Max outputmax_output_tokens1
Modalitiesmodalities1
Input modalitiesmodalities_input1
Output modalitiesmodalities_output1
Tokenizertokenizer1
Quantization formatquant_format1
File sizefile_size_gb1
Hugging Face repohf_repo1
Pipeline tagpipeline_tag1
Model cardmodel_card_url1
Downloadsmetric.downloads6
Likesmetric.likes13
Accessaccess1
Artifact kindartifact_kind1
Commercial use allowedcommercial_use_allowed1
Derivatives allowedderivatives_allowed1
Descriptiondescription1
Gatedgated1
Hf inference providershf_inference_providers1
Is quantizedis_quantized1
Last modifiedlast_modified1
Library namelibrary_name1
License keylicense_key1
License rawlicense_raw1
Downloads all timemetric.downloads_all_time6
Model typemodel_type1
Openrouter idopenrouter_id1
Openrouter listed atopenrouter_listed_at1
Reasoningreasoning1
Redistribution allowedredistribution_allowed1
Structured outputstructured_output1
Supported parameterssupported_parameters1
Tool callingtool_calling1
Weights availableweights_available1
Weights dtypeweights_dtype1
Claims are temporal and append-only: a new observation closes the previous claim (valid_to) instead of overwriting it. Conflicting claims from different sources are kept side by side and flagged — never averaged. Methodology →
Provenance
Attributed facts
43
Source tiers
T243
Freshest observation
5 h ago
Conflicts
None
Source documents 4
Tier 1 = official/primary, 2 = quality secondary, 3 = community, 4 = unverified. Every snapshot is archived; see all sources and the methodology.
Data quality (68/100) measures how well AI Atlas knows this entity — completeness, primary-source ratio, freshness, conflicts — never how good the model is. Methodology →