amd/Qwen3-VL-8B-Instruct-w8a8-llmcompressor
Updated 12 min ago · first seen 11 Sept 2026
model_01M294X79KCTC71YFGSCQ7NFFX
- Parameters
- 8.77B
- T2 · 21 min ago
- Released
- 3 Aug 2026
- T2 · 16 min ago
- License
- apache-2.0
- T2 · 16 min ago
Specification
- Release date
- 3 Aug 2026
Source:Hugging Face Hub (public pages, model cards, papers)T2observed 16 min agomedium
- Openness
- open-weights
Source:Hugging Face Hub (public pages, model cards, papers)T2observed 21 min agomedium
- License
- apache-2.0
Source:Hugging Face Hub (public pages, model cards, papers)T2observed 16 min agomedium
- Architecture
- Qwen3VLForConditionalGeneration
Source:Hugging Face Hub (public pages, model cards, papers)T2observed 16 min agomedium
- Parameters
- 8.77B
Source:Hugging Face Hub (public pages, model cards, papers)T2observed 21 min agomedium
- Languages
- en
Source:Hugging Face Hub (public pages, model cards, papers)T2observed 16 min agomedium
- Base model
- Qwen/Qwen3-VL-8B-Instruct
Source:Hugging Face Hub (public pages, model cards, papers)T2observed 16 min agomedium
- File size
- 10.6 GB
Source:Hugging Face Hub (public pages, model cards, papers)T2observed 16 min agomedium
- Hugging Face repo
- amd/Qwen3-VL-8B-Instruct-w8a8-llmcompressor
Source:Hugging Face Hub (public pages, model cards, papers)T2observed 21 min agomedium
- Pipeline tag
- image-text-to-text
Source:Hugging Face Hub (public pages, model cards, papers)T2observed 21 min agomedium
- Model card
Source:Hugging Face Hub (public pages, model cards, papers)T2observed 21 min agomedium
- Downloads
- 14,079
Source:Hugging Face Hub (public pages, model cards, papers)T2observed 12 min agomedium
- Likes
- 0
Source:Hugging Face Hub (public pages, model cards, papers)T2observed 12 min agomedium
- Gated
- No
Source:Hugging Face Hub (public pages, model cards, papers)T2observed 16 min agomedium
- Last modified
- 2026-08-07T02:01:41+00:00
Source:Hugging Face Hub (public pages, model cards, papers)T2observed 21 min agomedium
- Library name
- transformers
Source:Hugging Face Hub (public pages, model cards, papers)T2observed 16 min agomedium
- Downloads all time
- 14,186
Source:Hugging Face Hub (public pages, model cards, papers)T2observed 12 min agomedium
- Model type
- qwen3_vl
Source:Hugging Face Hub (public pages, model cards, papers)T2observed 16 min agomedium
- Tags
- transformers, safetensors, qwen3_vl, image-text-to-text, quantized, int8, w8a8, dynamic-quantization, 8-bit, llm-compressor, zendnn, compressed-tensors, amd, cpu-inference, en
Source:Hugging Face Hub (public pages, model cards, papers)T2observed 16 min agomedium
- Weights dtype
- BF16, I8
Source:Hugging Face Hub (public pages, model cards, papers)T2observed 16 min agomedium
Each value shows its source, tier and observation time. Conflicting claims are kept side by side and flagged — never averaged. How AI Atlas records facts →
Provenance
Attributed facts
20
Source tiers
T220
Freshest observation
12 min ago
Conflicts
None
Modalities
Modalities unavailable.
Capabilities
Tool calling
Unavailable
Structured output
Unavailable
Reasoning
Unavailable
Vision
Unavailable
Audio
Unavailable
Fine-tuning available
Unavailable
No capability flags have been observed from a source yet — we do not infer them.
- Languages
- en
Source:Hugging Face Hub (public pages, model cards, papers)T2observed 16 min agomedium
No benchmark results recorded
Current prices
No current prices recorded
Price history
Memory need = bytes per parameter (4-bit ≈ 0.5 × 1.15 overhead, 8-bit 1.0, fp16 2.0) + a KV-cache allowance. Not a measurement.
| Hardware | Quantization | Memory | Est. need | Fits |
|---|---|---|---|---|
| Apple M3 Ultra | 4bit | — | 5.5 GB | Yes |
| AMD Instinct MI325X | 4bit | 256 GB | 5.5 GB | Yes |
| AMD Instinct MI300X | 4bit | 192 GB | 5.5 GB | Yes |
| Apple M2 Ultra | 4bit | — | 5.5 GB | Yes |
| NVIDIA B200 | 4bit | 180 GB | 5.5 GB | Yes |
| NVIDIA H200 | 4bit | 141 GB | 5.5 GB | Yes |
| Apple M1 Ultra | 4bit | — | 5.5 GB | Yes |
| Apple M3 Max | 4bit | — | 5.5 GB | Yes |
| Apple M4 Max | 4bit | — | 5.5 GB | Yes |
| NVIDIA DGX Spark | 4bit | 128 GB | 5.5 GB | Yes |
| Apple M2 Max | 4bit | — | 5.5 GB | Yes |
| NVIDIA A100 80GB | 4bit | 80 GB | 5.5 GB | Yes |
| NVIDIA H100 SXM | 4bit | 80 GB | 5.5 GB | Yes |
| Apple M1 Max | 4bit | — | 5.5 GB | Yes |
| Apple M4 Pro | 4bit | — | 5.5 GB | Yes |
| Apple M3 Pro | 4bit | — | 5.5 GB | Yes |
| Apple M1 Pro | 4bit | — | 5.5 GB | Yes |
| Apple M2 Pro | 4bit | — | 5.5 GB | Yes |
| Apple M4 | 4bit | — | 5.5 GB | Yes |
| NVIDIA GeForce RTX 5090 | 4bit | 32 GB | 5.5 GB | Yes |
| Apple M2 | 4bit | — | 5.5 GB | Yes |
| Apple M3 | 4bit | — | 5.5 GB | Yes |
| NVIDIA GeForce RTX 3090 | 4bit | 24 GB | 5.5 GB | Yes |
| NVIDIA GeForce RTX 4090 | 4bit | 24 GB | 5.5 GB | Yes |
| Apple M1 | 4bit | — | 5.5 GB | Yes |
Ancestors 1
- Qwen3 VL 8B Instruct8.77B
This model
amd/Qwen3-VL-8B-Instruct-w8a8-llmcompressor
8.77B params
Quantizations 0
None recorded.
Descendants 0
None recorded.
Papers 0
No papers linked yet.
Repositories 0
No repositories linked yet.
As of
Rewind the record: see this entity's attributes exactly as AI Atlas knew them on a given day.
Claim history
Release daterelease_date1
| Value | Valid from → to | Status | Source | Confidence | Extractor |
|---|---|---|---|---|---|
| 3 Aug 2026 | → current | current | Hugging Face Hub (public pages, model cards, papers)T2 | medium | deterministic |
Opennessopenness1
| Value | Valid from → to | Status | Source | Confidence | Extractor |
|---|---|---|---|---|---|
| open-weights | → current | current | Hugging Face Hub (public pages, model cards, papers)T2 | medium | deterministic |
Licenselicense1
| Value | Valid from → to | Status | Source | Confidence | Extractor |
|---|---|---|---|---|---|
| apache-2.0 | → current | current | Hugging Face Hub (public pages, model cards, papers)T2 | medium | deterministic |
Architecturearchitecture1
| Value | Valid from → to | Status | Source | Confidence | Extractor |
|---|---|---|---|---|---|
| Qwen3VLForConditionalGeneration | → current | current | Hugging Face Hub (public pages, model cards, papers)T2 | medium | deterministic |
Parametersparameter_count1
| Value | Valid from → to | Status | Source | Confidence | Extractor |
|---|---|---|---|---|---|
| 8.77B | → current | current | Hugging Face Hub (public pages, model cards, papers)T2 | medium | deterministic |
Languageslanguages1
| Value | Valid from → to | Status | Source | Confidence | Extractor |
|---|---|---|---|---|---|
| en | → current | current | Hugging Face Hub (public pages, model cards, papers)T2 | medium | deterministic |
Base modelbase_model1
| Value | Valid from → to | Status | Source | Confidence | Extractor |
|---|---|---|---|---|---|
| Qwen/Qwen3-VL-8B-Instruct | → current | current | Hugging Face Hub (public pages, model cards, papers)T2 | medium | deterministic |
File sizefile_size_gb1
| Value | Valid from → to | Status | Source | Confidence | Extractor |
|---|---|---|---|---|---|
| 10.6 GB | → current | current | Hugging Face Hub (public pages, model cards, papers)T2 | medium | deterministic |
Hugging Face repohf_repo1
| Value | Valid from → to | Status | Source | Confidence | Extractor |
|---|---|---|---|---|---|
| amd/Qwen3-VL-8B-Instruct-w8a8-llmcompressor | → current | current | Hugging Face Hub (public pages, model cards, papers)T2 | medium | deterministic |
Pipeline tagpipeline_tag1
| Value | Valid from → to | Status | Source | Confidence | Extractor |
|---|---|---|---|---|---|
| image-text-to-text | → current | current | Hugging Face Hub (public pages, model cards, papers)T2 | medium | deterministic |
Model cardmodel_card_url1
| Value | Valid from → to | Status | Source | Confidence | Extractor |
|---|---|---|---|---|---|
| https://huggingface.co/amd/Qwen3-VL-8B-Instruct-w8a8-llmcompressor | → current | current | Hugging Face Hub (public pages, model cards, papers)T2 | medium | deterministic |
Downloadsmetric.downloads1
| Value | Valid from → to | Status | Source | Confidence | Extractor |
|---|---|---|---|---|---|
| 14,079 | → current | current | Hugging Face Hub (public pages, model cards, papers)T2 | medium | deterministic |
Likesmetric.likes1
| Value | Valid from → to | Status | Source | Confidence | Extractor |
|---|---|---|---|---|---|
| 0 | → current | current | Hugging Face Hub (public pages, model cards, papers)T2 | medium | deterministic |
Gatedgated1
| Value | Valid from → to | Status | Source | Confidence | Extractor |
|---|---|---|---|---|---|
| No | → current | current | Hugging Face Hub (public pages, model cards, papers)T2 | medium | deterministic |
Last modifiedlast_modified1
| Value | Valid from → to | Status | Source | Confidence | Extractor |
|---|---|---|---|---|---|
| 2026-08-07T02:01:41+00:00 | → current | current | Hugging Face Hub (public pages, model cards, papers)T2 | medium | deterministic |
Library namelibrary_name1
| Value | Valid from → to | Status | Source | Confidence | Extractor |
|---|---|---|---|---|---|
| transformers | → current | current | Hugging Face Hub (public pages, model cards, papers)T2 | medium | deterministic |
Downloads all timemetric.downloads_all_time1
| Value | Valid from → to | Status | Source | Confidence | Extractor |
|---|---|---|---|---|---|
| 14,186 | → current | current | Hugging Face Hub (public pages, model cards, papers)T2 | medium | deterministic |
Model typemodel_type1
| Value | Valid from → to | Status | Source | Confidence | Extractor |
|---|---|---|---|---|---|
| qwen3_vl | → current | current | Hugging Face Hub (public pages, model cards, papers)T2 | medium | deterministic |
Weights dtypeweights_dtype1
| Value | Valid from → to | Status | Source | Confidence | Extractor |
|---|---|---|---|---|---|
| BF16, I8 | → current | current | Hugging Face Hub (public pages, model cards, papers)T2 | medium | deterministic |
Claims are temporal and append-only: a new observation closes the previous claim (valid_to) instead of overwriting it. Conflicting claims from different sources are kept side by side and flagged — never averaged. Methodology →
New model: amd/Qwen3-VL-8B-Instruct-w8a8-llmcompressor (AMD)
huggingface
| Source | Document | Type | Tier | Last observed | Snapshots |
|---|---|---|---|---|---|
| Hugging Face Hub (public pages, model cards, papers) | huggingface.co/amd/Qwen3-VL-8B-Instruct-w8a8-llmcompressor | model_page | T2· Quality secondary | 12 min ago | 2 |
| Hugging Face Hub (public pages, model cards, papers) | huggingface.co/models?author=amd&p=0&sort=downloads | listing | T2· Quality secondary | 19 min ago | 2 |
Tier 1 = official/primary, 2 = quality secondary, 3 = community, 4 = unverified. Every snapshot is archived; see all sources and the methodology.