GDN-primed-HQwen3-8B-Instruct
Amazon Web Serviceshuggingface.co/amazon/GDN-primed-HQwen3-8B-Instr
Updated 13 min ago · first seen 11 Sept 2026
model_01M294X4DAFKMTZS6BYJX0WJW4
- Parameters
- 8.5B
- T2 · 2 h ago
- Released
- 31 Mar 2026
- T2 · 2 h ago
- License
- apache-2.0
- T2 · 2 h ago
Specification
- Release date
- 31 Mar 2026
Source:Hugging Face Hub (public pages, model cards, papers)T2observed 2 h agomedium
- Openness
- open-weights
Source:Hugging Face Hub (public pages, model cards, papers)T2observed 2 h agomedium
- License
- apache-2.0
Source:Hugging Face Hub (public pages, model cards, papers)T2observed 2 h agomedium
- Architecture
- HybridQwen3ForCausalLM
Source:Hugging Face Hub (public pages, model cards, papers)T2observed 2 h agomedium
- Parameters
- 8.5B
Source:Hugging Face Hub (public pages, model cards, papers)T2observed 2 h agomedium
- Base model
- Qwen/Qwen3-8B
Source:Hugging Face Hub (public pages, model cards, papers)T2observed 2 h agomedium
- File size
- 17 GB
Source:Hugging Face Hub (public pages, model cards, papers)T2observed 2 h agomedium
- Hugging Face repo
- amazon/GDN-primed-HQwen3-8B-Instruct
Source:Hugging Face Hub (public pages, model cards, papers)T2observed 2 h agomedium
- Pipeline tag
- text-generation
Source:Hugging Face Hub (public pages, model cards, papers)T2observed 2 h agomedium
- Model card
Source:Hugging Face Hub (public pages, model cards, papers)T2observed 2 h agomedium
- Downloads
- 258
Source:Hugging Face Hub (public pages, model cards, papers)T2observed 13 min agomedium
- Likes
- 2
Source:Hugging Face Hub (public pages, model cards, papers)T2observed 13 min agomedium
- Gated
- No
Source:Hugging Face Hub (public pages, model cards, papers)T2observed 2 h agomedium
- Last modified
- 2026-04-03T03:37:18+00:00
Source:Hugging Face Hub (public pages, model cards, papers)T2observed 2 h agomedium
- Library name
- transformers
Source:Hugging Face Hub (public pages, model cards, papers)T2observed 2 h agomedium
- Downloads all time
- 89,731
Source:Hugging Face Hub (public pages, model cards, papers)T2observed 13 min agomedium
- Model type
- hybrid_qwen3
Source:Hugging Face Hub (public pages, model cards, papers)T2observed 2 h agomedium
- Tags
- transformers, safetensors, hybrid_qwen3, text-generation, hybrid, ssm, state-space-model, linear-attention, gated-deltanet, priming, long-context, instruction-tuned
Source:Hugging Face Hub (public pages, model cards, papers)T2observed 2 h agomedium
- Weights dtype
- BF16, F32
Source:Hugging Face Hub (public pages, model cards, papers)T2observed 2 h agomedium
Each value shows its source, tier and observation time. Conflicting claims are kept side by side and flagged — never averaged. How AI Atlas records facts →
Provenance
Attributed facts
19
Source tiers
T219
Freshest observation
13 min ago
Conflicts
None
Modalities
Modalities unavailable.
Capabilities
Tool calling
Unavailable
Structured output
Unavailable
Reasoning
Unavailable
Vision
Unavailable
Audio
Unavailable
Fine-tuning available
Unavailable
No capability flags have been observed from a source yet — we do not infer them.
No structured attributes yet.
No benchmark results recorded
Current prices
No current prices recorded
Price history
Memory need = bytes per parameter (4-bit ≈ 0.5 × 1.15 overhead, 8-bit 1.0, fp16 2.0) + a KV-cache allowance. Not a measurement.
| Hardware | Quantization | Memory | Est. need | Fits |
|---|---|---|---|---|
| NVIDIA DGX B200 | 4bit | 1,440 GB | 5.4 GB | Yes |
| Apple M3 Ultra | 4bit | — | 5.4 GB | Yes |
| Mac Studio (Apple M5 Ultra) | 4bit | — | 5.4 GB | Yes |
| AMD Instinct MI325X | 4bit | 256 GB | 5.4 GB | Yes |
| AMD Instinct MI300X | 4bit | 192 GB | 5.4 GB | Yes |
| Apple M2 Ultra | 4bit | — | 5.4 GB | Yes |
| NVIDIA B200 | 4bit | 180 GB | 5.4 GB | Yes |
| NVIDIA H200 | 4bit | 141 GB | 5.4 GB | Yes |
| NVIDIA H200 NVL | 4bit | 141 GB | 5.4 GB | Yes |
| Apple M1 Ultra | 4bit | — | 5.4 GB | Yes |
| Apple M3 Max | 4bit | — | 5.4 GB | Yes |
| Apple M4 Max | 4bit | — | 5.4 GB | Yes |
| Mac Studio (Apple M5 Max) | 4bit | — | 5.4 GB | Yes |
| MacBook Pro (Apple M5 Max) | 4bit | — | 5.4 GB | Yes |
| NVIDIA DGX Spark | 4bit | 128 GB | 5.4 GB | Yes |
| Apple M2 Max | 4bit | — | 5.4 GB | Yes |
| NVIDIA H100 NVL | 4bit | 94 GB | 5.4 GB | Yes |
| NVIDIA A100 80GB | 4bit | 80 GB | 5.4 GB | Yes |
| NVIDIA H100 SXM | 4bit | 80 GB | 5.4 GB | Yes |
| Apple M1 Max | 4bit | — | 5.4 GB | Yes |
| Apple M4 Pro | 4bit | — | 5.4 GB | Yes |
| Mac mini (Apple M5 Pro) | 4bit | — | 5.4 GB | Yes |
| MacBook Pro (Apple M5 Pro) | 4bit | — | 5.4 GB | Yes |
| Apple M3 Pro | 4bit | — | 5.4 GB | Yes |
| Apple M1 Pro | 4bit | — | 5.4 GB | Yes |
| Apple M2 Pro | 4bit | — | 5.4 GB | Yes |
| Apple M4 | 4bit | — | 5.4 GB | Yes |
| iMac (Apple M4) | 4bit | — | 5.4 GB | Yes |
| Mac mini (Apple M6) | 4bit | — | 5.4 GB | Yes |
| MacBook Air (Apple M5) | 4bit | — | 5.4 GB | Yes |
| MacBook Pro (Apple M5) | 4bit | — | 5.4 GB | Yes |
| NVIDIA GeForce RTX 5090 | 4bit | 32 GB | 5.4 GB | Yes |
| Apple M2 | 4bit | — | 5.4 GB | Yes |
| Apple M3 | 4bit | — | 5.4 GB | Yes |
| NVIDIA GeForce RTX 3090 | 4bit | 24 GB | 5.4 GB | Yes |
| NVIDIA GeForce RTX 4090 | 4bit | 24 GB | 5.4 GB | Yes |
| Apple M1 | 4bit | — | 5.4 GB | Yes |
Ancestors 2
- Qwen3 8B8.19B
- Qwen3-8B-BaseQwen
This model
GDN-primed-HQwen3-8B-Instruct
8.5B params
Quantizations 0
None recorded.
Descendants 0
None recorded.
Papers 2
- arXiv:2412.06464Active35
- arXiv:2502.17605Active35
Repositories 0
No repositories linked yet.
As of
Rewind the record: see this entity's attributes exactly as AI Atlas knew them on a given day.
Claim history · Openness
Opennessopenness1
| Value | Valid from → to | Status | Source | Confidence | Extractor |
|---|---|---|---|---|---|
| open-weights | → current | current | Hugging Face Hub (public pages, model cards, papers)T2 | medium | deterministic |
Claims are temporal and append-only: a new observation closes the previous claim (valid_to) instead of overwriting it. Conflicting claims from different sources are kept side by side and flagged — never averaged. Methodology →
New model: GDN-primed-HQwen3-8B-Instruct (Amazon Web Services)
huggingface
| Source | Document | Type | Tier | Last observed | Snapshots |
|---|---|---|---|---|---|
| Hugging Face Hub (public pages, model cards, papers) | huggingface.co/amazon/GDN-primed-HQwen3-8B-Instruct | model_page | T2· Quality secondary | 13 min ago | 3 |
| Hugging Face Hub (public pages, model cards, papers) | huggingface.co/models?author=amazon&p=0&sort=downloads | listing | T2· Quality secondary | 18 min ago | 3 |
| Hugging Face Hub (public pages, model cards, papers) | huggingface.co/amazon/GDN-primed-HQwen3-8B-Instruct/raw/main/README.md | model_card | T2· Quality secondary | 56 min ago | 1 |
Tier 1 = official/primary, 2 = quality secondary, 3 = community, 4 = unverified. Every snapshot is archived; see all sources and the methodology.