Updated 10 h ago · first seen 11 Sept 2026
model_01M294ZPG8HB05BERG3MQ5TMVT
Overview
Identity
Identity block not returned by the API for this entity.
Openness
Openness not classified yet — no sourced evidence to place this model in the ontology.
Key facts
- Release date
Source:Hugging Face Hub (public pages, model cards, papers)T2observed 11 h agomedium
- Model card
Source:Hugging Face Hub (public pages, model cards, papers)T2observed 12 h agomedium
Architecture
- Architecture
Source:Hugging Face Hub (public pages, model cards, papers)T2observed 11 h agomedium
- Model type
Source:Hugging Face Hub (public pages, model cards, papers)T2observed 11 h agomedium
- Parameters
Source:Hugging Face Hub (public pages, model cards, papers)T2observed 12 h agomedium
- Weights dtype
Source:Hugging Face Hub (public pages, model cards, papers)T2observed 11 h agomedium
- File size
Source:Hugging Face Hub (public pages, model cards, papers)T2observed 11 h agomedium
- Library name
Source:Hugging Face Hub (public pages, model cards, papers)T2observed 11 h agomedium
- Pipeline tag
Source:Hugging Face Hub (public pages, model cards, papers)T2observed 12 h agomedium
- Hugging Face repo
Source:Hugging Face Hub (public pages, model cards, papers)T2observed 12 h agomedium
Capabilities
Modalities
Modalities unavailable.
Capabilities
Tool calling
Unavailable
Structured output
Unavailable
Reasoning
Unavailable
Vision
Unavailable
Audio
Unavailable
Fine-tuning available
Unavailable
No capability flags have been observed from a source yet — we do not infer them.
- Languages
Source:Hugging Face Hub (public pages, model cards, papers)T2observed 11 h agomedium
Hardware fit37
Assumptions (6)
- Estimated, not measured: weights = parameters × bytes/param × 1.15 runtime overhead.
- bytes/param: 4bit = 0.5, 8bit = 1.0, fp16 = 2.0 (uniform quantization, no per-layer exceptions).
- KV cache approximated at 0.5 GB per 8 192 tokens of context, independent of architecture (GQA/MLA models need less).
- A model 'fits' when the estimate is at most the device memory minus 2 GB reserved for the OS and framework.
- Mixture-of-experts models are estimated on total parameters (all experts must be resident); active parameters are ignored.
- Device memory uses the largest configuration when several are listed (e.g. Apple silicon tiers).
Papers2
- arXiv:2312.17279Active35
- arXiv:2305.05084Active35
Datasets6
- europarlActive35
- nvidia/GranaryActive35
- fleursActive35
- multilingual_librispeechActive35
- voxpopuliActive35
Timeline1
Full timeline →New model: nemotron-3.5-asr-streaming-0.6b (NVIDIA)
huggingface
Change history25
Viewing AI Atlas as of 12 Sept 2026 — attributes exactly as the atlas knew them on that day; later corrections are not shown.
Back to today →Attributes as of 12 Sept 2026 26 claims in force
- Release date
- 15 May 2026
- Openness
- open-weights
- License
- Other
- Architecture
- Nemotron3_5AsrForRNNT
- Parameters
- 638M
- Languages
- en, es, de, fr, it, ar, ja, ko, pt, ru, hi, zh, vi, he, nl, cs, da, pl, no, sv, th, tr, bg, el, et, fi, hr, hu, lt, lv, ro, sk, uk, mt, sl
- Quantization format
- gguf
- File size
- 2.5 GB
- Hugging Face repo
- nvidia/nemotron-3.5-asr-streaming-0.6b
- Pipeline tag
- automatic-speech-recognition
- Downloads
- 722,175
- Likes
- 1,102
- Datasets
- nvidia/Granary, multilingual_librispeech, fleurs, mozilla-foundation/common_voice_8_0, voxpopuli, europarl
- Gated
- No
- Hf inference providers
- deepinfra, fal-ai, together
- Is quantized
- Yes
- Last modified
- 2026-09-10T16:49:07+00:00
- Library name
- nemo
- License name
- openmdw-1.1
- License url
- https://openmdw.ai/license/1-1/
- Downloads all time
- 2,393,449
- Model type
- nemotron3_5_asr
- Tags
- nemo, safetensors, gguf, nemotron3_5_asr, feature-extraction, transformers, speech-recognition, cache-aware ASR, automatic-speech-recognition, streaming-asr, multilingual, speech, audio, FastConformer, RNNT, Parakeet, ASR, pytorch, NeMo, en, es, de, fr, it, ar, ja, ko, pt, ru, hi
- Weights available
- Yes
- Weights dtype
- F32
Release daterelease_date1
Opennessopenness1
Licenselicense1
Architecturearchitecture1
Parametersparameter_count1
Languageslanguages1
Quantization formatquant_format1
File sizefile_size_gb1
Hugging Face repohf_repo1
Pipeline tagpipeline_tag1
Model cardmodel_card_url1
Downloadsmetric.downloads1
Likesmetric.likes1
Datasetsdatasets1
Gatedgated1
Hf inference providershf_inference_providers1
Is quantizedis_quantized1
Last modifiedlast_modified1
Library namelibrary_name1
License namelicense_name1
License urllicense_url1
Downloads all timemetric.downloads_all_time1
Model typemodel_type1
Weights dtypeweights_dtype1
Claims are temporal and append-only: a new observation closes the previous claim (valid_to) instead of overwriting it. Conflicting claims from different sources are kept side by side and flagged — never averaged. Methodology →
Provenance
Attributed facts
25
Source tiers
T225
Freshest observation
10 h ago
Conflicts
None
Source documents 3
Tier 1 = official/primary, 2 = quality secondary, 3 = community, 4 = unverified. Every snapshot is archived; see all sources and the methodology.
Data quality (53/100) measures how well AI Atlas knows this entity — completeness, primary-source ratio, freshness, conflicts — never how good the model is. Methodology →