Run locally
Describe the machine — memory, number of devices, platform, quantization, context, batch — and the atlas lists the downloadable models estimated to fit, with the breakdown behind each number and the GGUF / MLX artifacts recorded for them. Estimates, never measurements.
EstimatedAll fit figures are estimates, not measurements.454 of 529 evaluated models fit
Every figure is an ESTIMATE: weights = params × bytes/param (× 1.15 overhead) unless an artifact's observed file size is available; KV cache uses architecture metadata when known, else 0.5 GB per 8K tokens × batch.
Assumptions (7)
- Estimated, not measured: weights = parameters × bytes/param × 1.15 runtime overhead (or the observed artifact file size when one is recorded).
- bytes/param: 4bit = 0.5, 8bit = 1.0, fp16 = 2.0 (uniform quantization, no per-layer exceptions).
- KV cache: 2 × layers × kv_heads × head_dim × 2 bytes × context × batch when the architecture is known; otherwise 0.5 GB per 8 192 tokens (× batch), independent of architecture (GQA/MLA models need less).
- A model 'fits' when the estimate is at most the device memory minus 2 GB reserved for the OS and framework.
- Mixture-of-experts models are estimated on total parameters (all experts must be resident); active parameters are ignored.
- Device memory uses the largest configuration when several are listed (e.g. Apple silicon tiers).
- Multi-GPU: device memories are summed; interconnect bandwidth, tensor-parallel replication and pipeline bubbles are not modelled.
64 GB total1 × 64 GB4-bit8K contextbatch 1Show all evaluated models
| Model | Params | Est. memory | Headroom | Fits | Breakdown | Artifacts | |
|---|---|---|---|---|---|---|---|
| Llama-3.2-90B-Vision-Instructrestricted-weightsMeta AILlama-3.2-Community | 88.6B | 51.4 GB estimated | +10.6 GB | ✓ fits |
| none recorded | |
| HunyuanImage-3.0-InstructOpen weightsTencentOther | 83B | 48.2 GB estimated | +13.8 GB | ✓ fits |
| none recorded | |
| cerebras/GLM-4.5-Air-REAP-82B-A12BOpen weightsCerebras SystemsMIT | 81.9B | 47.6 GB estimated | +14.4 GB | ✓ fits |
| none recorded | |
| Hunyuan A13B InstructOpen weightsTencentOther | 80.4B | 46.7 GB estimated | +15.3 GB | ✓ fits |
| none recorded | |
| UI-TARS-72B-DPOOpen weightsByteDanceApache-2.0 | 73.4B | 42.7 GB estimated | +19.3 GB | ✓ fits |
| none recorded | |
| Kimi-Dev-72BOpen weightsMoonshot AIMIT | 72.7B | 42.3 GB estimated | +19.7 GB | ✓ fits |
| none recorded | |
| Qwen2.5restricted-weightsQwen TeamResearch-Only | 72.7B | 42.3 GB estimated | +19.7 GB | ✓ fits |
| none recorded | |
| QVQ-72B-PreviewOpen weightsQwen Team | 72B | 41.9 GB estimated | +20.1 GB | ✓ fits |
| none recorded | |
| Hermes-4-70Brestricted-weightsNous ResearchLlama-3-Community | 70.5B | 41.1 GB estimated | +20.9 GB | ✓ fits |
| none recorded | |
| Llama 3.1 70B Instructrestricted-weightsMeta AILlama-3.1-Community | 70.5B | 41.1 GB estimated | +20.9 GB | ✓ fits |
| none recorded | |
| NousResearch/Meta-Llama-3.1-70B-Instructrestricted-weightsNous ResearchLlama-3.1-Community | 70.5B | 41.1 GB estimated | +20.9 GB | ✓ fits |
| none recorded | |
| Meta-Llama-3-70Brestricted-weightsMeta AILlama-3-Community | 70.5B | 41.1 GB estimated | +20.9 GB | ✓ fits |
| none recorded | |
| fireworks-ai/llama-3-firefunction-v2restricted-weightsFireworks AILlama-3-Community | 70.5B | 41.1 GB estimated | +20.9 GB | ✓ fits |
| none recorded | |
| Llama 3.3 70B Instructrestricted-weightsMeta AILlama-3.3-Community | 70.5B | 41.1 GB estimated | +20.9 GB | ✓ fits |
| 2 recorded | |
| ↳ amd/Llama-3.3-70B-Instruct-FP8-KVfp8 | 72.7 GB observed | 84.1 GB | −22.1 GB | ✗ too large |
| 4bit | |
| ↳ amd/Llama-3.3-70B-Instruct-FP8-KVfp8 | 72.7 GB observed | 84.1 GB | −22.1 GB | ✗ too large |
| 4bit | |
| NousResearch/Meta-Llama-3-70B-InstructOpen weightsNous ResearchOther | 70.5B | 41.1 GB estimated | +20.9 GB | ✓ fits |
| none recorded | |
| Jamba-v0.1Open weightsAI21 LabsApache-2.0 | 51.6B | 30.2 GB estimated | +31.8 GB | ✓ fits |
| none recorded | |
| AI21-Jamba-Mini-1.5Open weightsAI21 LabsOther | 51.6B | 30.2 GB estimated | +31.8 GB | ✓ fits |
| none recorded | |
| AI21-Jamba-Mini-1.7Open weightsAI21 LabsOther | 51.6B | 30.2 GB estimated | +31.8 GB | ✓ fits |
| 1 recorded | |
| ↳ ai21labs/AI21-Jamba-Mini-1.7-FP8 | 55.1 GB observed | 63.8 GB | −1.8 GB | ✗ too large |
| 4bit | |
| AI21-Jamba-Mini-1.6Open weightsAI21 LabsOther | 51.6B | 30.2 GB estimated | +31.8 GB | ✓ fits |
| none recorded | |
| AI21-Jamba2-MiniOpen weightsAI21 LabsApache-2.0 | 51.6B | 30.2 GB estimated | +31.8 GB | ✓ fits |
| 1 recorded | |
| ↳ ai21labs/AI21-Jamba2-Mini-FP8 | 55.1 GB observed | 63.8 GB | −1.8 GB | ✗ too large |
| 4bit | |
| Kimi-Linear-48B-A3B-InstructOpen weightsMoonshot AIMIT | 49.1B | 28.7 GB estimated | +33.3 GB | ✓ fits |
| none recorded | |
| Kimi-Linear-48B-A3B-BaseOpen weightsMoonshot AIMIT | 49.1B | 28.7 GB estimated | +33.3 GB | ✓ fits |
| none recorded | |
| firefunction-v1Open weightsFireworks AIApache-2.0 | 46.7B | 27.4 GB estimated | +34.6 GB | ✓ fits |
| none recorded | |
| function-calling-v1Open weightsFireworks AI | 46.7B | 27.4 GB estimated | +34.6 GB | ✓ fits |
| none recorded | |
| Nous-Hermes-2-Mixtral-8x7B-DPOOpen weightsNous ResearchApache-2.0 | 46.7B | 27.4 GB estimated | +34.6 GB | ✓ fits |
| none recorded | |
| Mixtral-8x7B-Instruct-v0.1Open weightsMistral AIApache-2.0 | 46.7B | 27.4 GB estimated | +34.6 GB | ✓ fits |
| none recorded | |
| DFN2B-CLIP-ViT-L-14-39Brestricted-weightsAppleApple-AMLR | 39B | 22.9 GB estimated | +39.1 GB | ✓ fits |
| none recorded | |
| Seed-OSS-36B-BaseOpen weightsByteDanceApache-2.0 | 36.1B | 21.3 GB estimated | +40.7 GB | ✓ fits |
| 1 recorded | |
| ↳ NousResearch/Hermes-4.3-36B-GGUFgguf | — estimated | 21.3 GB | +40.7 GB | ✓ fits |
| 4bit | |
| Seed-OSS-36B-InstructOpen weightsByteDanceApache-2.0 | 36.1B | 21.3 GB estimated | +40.7 GB | ✓ fits |
| none recorded | |
| Qwen3.6 35B A3BOpen weightsQwenApache-2.0 | 36B | 21.2 GB estimated | +40.8 GB | ✓ fits |
| 11 recorded | |
| ↳ Intel/Qwen3.6-35B-A3B-int4-mixed-AutoRound | 21.5 GB observed | 25.2 GB | +36.8 GB | ✓ fits |
| 4bit | |
| ↳ nvidia/Qwen3.6-35B-A3B-NVFP4 | 23.4 GB observed | 27.5 GB | +34.5 GB | ✓ fits |
| 4bit | |
| ↳ unsloth/Qwen3.6-35B-A3B-NVFP4 | 26.5 GB observed | 31.0 GB | +31.0 GB | ✓ fits |
| 4bit | |
| ↳ Qwen/Qwen3.6-35B-A3B-FP8fp8 | 37.5 GB observed | 43.6 GB | +18.4 GB | ✓ fits |
| 4bit | |
| +4 more artifacts on the model page. | |||||||
| cerebras/Kimi-Linear-REAP-35B-A3B-InstructOpen weightsCerebras SystemsMIT | 35.1B | 20.7 GB estimated | +41.3 GB | ✓ fits |
| none recorded | |
| c4ai-command-r-v01restricted-weightsCohereCC-BY-NC-4.0 | 35B | 20.6 GB estimated | +41.4 GB | ✓ fits |
| none recorded | |
| Nous-Hermes-2-Yi-34BOpen weightsNous ResearchApache-2.0 | 34.4B | 20.3 GB estimated | +41.7 GB | ✓ fits |
| none recorded | |
| GKA-primed-HQwen3-32B-ReasonerOpen weightsAmazon Web ServicesApache-2.0 | 34.1B | 20.1 GB estimated | +41.9 GB | ✓ fits |
| none recorded | |
| MiniMax-H3Open weightsMiniMaxOther | 33.1B | 19.5 GB estimated | +42.5 GB | ✓ fits |
| 2 recorded | |
| ↳ unsloth/MiniMax-H3-GGUFgguf | — estimated | 19.5 GB | +42.5 GB | ✓ fits |
| 4bit | |
| ↳ unsloth/MiniMax-H3-GGUFgguf | — estimated | 19.5 GB | +42.5 GB | ✓ fits |
| 4bit | |
| DeepSeek-R1-Distill-Qwen-32BOpen weightsDeepSeekMIT | 32.8B | 19.3 GB estimated | +42.7 GB | ✓ fits |
| none recorded | |
| SynLogic-32BOpen weightsMiniMaxMIT | 32.8B | 19.3 GB estimated | +42.7 GB | ✓ fits |
| none recorded | |
| SynLogic-Mix-3-32BOpen weightsMiniMaxMIT | 32.8B | 19.3 GB estimated | +42.7 GB | ✓ fits |
| none recorded | |
| Qwen3 32BOpen weightsQwenApache-2.0 | 32.8B | 19.3 GB estimated | +42.7 GB | ✓ fits |
| none recorded | |
| Qwen2.5-Coderrestricted-weightsQwen TeamResearch-Only | 32.5B | 19.2 GB estimated | +42.8 GB | ✓ fits |
| none recorded | |
| aya-expanse-32brestricted-weightsCohereCC-BY-NC-4.0 | 32.3B | 19.1 GB estimated | +42.9 GB | ✓ fits |
| none recorded | |
| c4ai-command-r-08-2024restricted-weightsCohereCC-BY-NC-4.0 | 32.3B | 19.1 GB estimated | +42.9 GB | ✓ fits |
| none recorded | |
| OLMo-2-0325-32B-InstructOpen weightsAllen Institute for AIApache-2.0 | 32.2B | 19.0 GB estimated | +43.0 GB | ✓ fits |
| none recorded | |
| FLUX.2-devOpen weightsBlack Forest LabsOther | 32.2B | 19.0 GB estimated | +43.0 GB | ✓ fits |
| 1 recorded | |
| ↳ black-forest-labs/FLUX.2-dev-NVFP4 | — estimated | 19.0 GB | +43.0 GB | ✓ fits |
| 4bit | |
| QwQ-32BOpen weightsAlibaba GroupApache-2.0 | 32B | 18.9 GB estimated | +43.1 GB | ✓ fits |
| none recorded | |
| QwQ-32B-PreviewOpen weightsAlibaba Group | 32B | 18.9 GB estimated | +43.1 GB | ✓ fits |
| none recorded | |
| Qwen2.5-VL-32B-InstructOpen weightsQwen TeamApache-2.0 | 32B | 18.9 GB estimated | +43.1 GB | ✓ fits |
| none recorded | |
| Nemotron 3 Nano 30B A3BOpen weightsNVIDIAOther | 31.6B | 18.7 GB estimated | +43.3 GB | ✓ fits |
| none recorded | |
| Gemma 4 31BOpen weightsGoogleApache-2.0 | 31.3B | 18.5 GB estimated | +43.5 GB | ✓ fits |
| 5 recorded | |
| ↳ mlx-community/gemma-4-31b-it-4bitmlx | 18.4 GB observed | 21.7 GB | +40.3 GB | ✓ fits |
| 4bit | |
| ↳ mlx-community/gemma-4-31b-it-4bitmlx | 18.4 GB observed | 21.7 GB | +40.3 GB | ✓ fits |
| 4bit | |
| ↳ Intel/gemma-4-31B-it-int4-AutoRound | 19.2 GB observed | 22.6 GB | +39.4 GB | ✓ fits |
| 4bit | |
| ↳ nvidia/Gemma-4-31B-IT-NVFP4 | 32.6 GB observed | 38.0 GB | +24.0 GB | ✓ fits |
| 4bit | |
| +1 more artifacts on the model page. | |||||||
| GLM 4.7 FlashOpen weightsZ.ai (Zhipu AI)MIT | 31.2B | 18.4 GB estimated | +43.5 GB | ✓ fits |
| none recorded | |
| browsesafeOpen weightsPerplexity AIMIT | 30.5B | 18.1 GB estimated | +43.9 GB | ✓ fits |
| none recorded | |
| North Mini Code (free)Open weightsCohereApache-2.0 | 30.5B | 18.0 GB estimated | +44.0 GB | ✓ fits |
| none recorded | |
| Hy-MT2-30B-A3BOpen weightsTencentApache-2.0 | 30.1B | 17.8 GB estimated | +44.2 GB | ✓ fits |
| none recorded | |
| ERNIE-4.5-VL-28B-A3B-ThinkingOpen weightsBaiduApache-2.0 | 29.7B | 17.6 GB estimated | +44.5 GB | ✓ fits |
| none recorded | |
| ERNIE-4.5-VL-28B-A3B-PTOpen weightsBaiduApache-2.0 | 29.4B | 17.4 GB estimated | +44.6 GB | ✓ fits |
| none recorded | |
| granite-4.1-30bOpen weightsIBMApache-2.0 | 28.9B | 17.1 GB estimated | +44.9 GB | ✓ fits |
| none recorded | |
| Qwen3.8 27BOpen weightsQwenApache-2.0 | 27.8B | 16.5 GB estimated | +45.5 GB | ✓ fits |
| 16 recorded | |
| ↳ mlx-community/Qwen3.8-27B-4bitmlx | 16.1 GB observed | 19.0 GB | +43.0 GB | ✓ fits |
| 4bit | |
| ↳ mlx-community/Qwen3.8-27B-4bitmlx | 16.1 GB observed | 19.0 GB | +43.0 GB | ✓ fits |
| 4bit | |
| ↳ amd/Qwen3.8-27B-Quark-AWQ-INT4-W4A16awq | 19.5 GB observed | 22.9 GB | +39.1 GB | ✓ fits |
| 4bit | |
| ↳ amd/Qwen3.8-27B-Quark-AWQ-INT4-W4A16awq | 19.5 GB observed | 22.9 GB | +39.1 GB | ✓ fits |
| 4bit | |
| +4 more artifacts on the model page. | |||||||
| Qwen3.6 27BOpen weightsQwenApache-2.0 | 27.8B | 16.5 GB estimated | +45.5 GB | ✓ fits |
| 8 recorded | |
| ↳ Intel/Qwen3.6-27B-int4-AutoRound | 19.0 GB observed | 22.4 GB | +39.6 GB | ✓ fits |
| 4bit | |
| ↳ unsloth/Qwen3.6-27B-NVFP4 | 23.4 GB observed | 27.4 GB | +34.6 GB | ✓ fits |
| 4bit | |
| ↳ Qwen/Qwen3.6-27B-FP8fp8 | 30.9 GB observed | 36.0 GB | +26.0 GB | ✓ fits |
| 4bit | |
| ↳ Qwen/Qwen3.6-27B-FP8fp8 | 30.9 GB observed | 36.0 GB | +26.0 GB | ✓ fits |
| 4bit | |
| +4 more artifacts on the model page. | |||||||
| Gemma 4 26B A4BOpen weightsGoogleApache-2.0 | 25.8B | 15.3 GB estimated | +46.7 GB | ✓ fits |
| 4 recorded | |
| ↳ nvidia/Gemma-4-26B-A4B-NVFP4 | 18.8 GB observed | 22.1 GB | +39.9 GB | ✓ fits |
| 4bit | |
| ↳ unsloth/gemma-4-26B-A4B-it-qat-GGUFgguf | — estimated | 15.3 GB | +46.7 GB | ✓ fits |
| 4bit | |
| ↳ unsloth/gemma-4-26B-A4B-it-GGUFgguf | — estimated | 15.3 GB | +46.7 GB | ✓ fits |
| 4bit | |
| ↳ unsloth/gemma-4-26B-A4B-it-GGUFgguf | — estimated | 15.3 GB | +46.7 GB | ✓ fits |
| 4bit | |
| cerebras/Qwen3-Coder-REAP-25B-A3BOpen weightsCerebras SystemsApache-2.0 | 24.9B | 14.8 GB estimated | +47.2 GB | ✓ fits |
| none recorded | |
| Voxtral Small 24B 2507Open weightsMistral AIApache-2.0 | 24.3B | 14.4 GB estimated | +47.5 GB | ✓ fits |
| none recorded | |
| Devstral-Small-2-24B-Instruct-2512Open weightsMistral AIApache-2.0 | 24B | 14.3 GB estimated | +47.7 GB | ✓ fits |
| 2 recorded | |
| ↳ mlx-community/Devstral-Small-2-24B-Instruct-2512-4bitmlx | 15.1 GB observed | 17.9 GB | +44.1 GB | ✓ fits |
| 4bit | |
| ↳ mlx-community/Devstral-Small-2-24B-Instruct-2512-4bitmlx | 15.1 GB observed | 17.9 GB | +44.1 GB | ✓ fits |
| 4bit | |
| Mistral Small 3.1 24BOpen weightsMistral AIApache-2.0 | 24B | 14.3 GB estimated | +47.7 GB | ✓ fits |
| none recorded | |
| Mistral Small 3.2 24BOpen weightsMistral AIApache-2.0 | 24B | 14.3 GB estimated | +47.7 GB | ✓ fits |
| 1 recorded | |
| ↳ Intel/Mistral-Small-3.2-24B-Instruct-2506-int4-AutoRound | 15.1 GB observed | 17.9 GB | +44.1 GB | ✓ fits |
| 4bit | |
| cerebras/GLM-4.7-Flash-REAP-23B-A3BOpen weightsCerebras SystemsMIT | 23B | 13.7 GB estimated | +48.3 GB | ✓ fits |
| none recorded | |
| ERNIE-4.5-21B-A3B-PTOpen weightsBaiduApache-2.0 | 21.9B | 13.1 GB estimated | +48.9 GB | ✓ fits |
| none recorded | |
| ERNIE-4.5-21B-A3B-Base-PTOpen weightsBaiduApache-2.0 | 21.8B | 13.1 GB estimated | +49.0 GB | ✓ fits |
| none recorded | |
| ERNIE-4.5-21B-A3B-ThinkingOpen weightsBaiduApache-2.0 | 21.8B | 13.1 GB estimated | +49.0 GB | ✓ fits |
| none recorded | |
| gpt-oss-safeguard-20bOpen weightsOpenAIApache-2.0 | 21.5B | 12.9 GB estimated | +49.1 GB | ✓ fits |
| none recorded | |
| gpt-oss-20bOpen weightsOpenAIApache-2.0 | 20.9B | 12.5 GB estimated | +49.5 GB | ✓ fits |
| 2 recorded | |
| ↳ mlx-community/gpt-oss-20b-MXFP4-Q8mlx | 12.1 GB observed | 14.4 GB | +47.6 GB | ✓ fits |
| 4bit | |
| ↳ mlx-community/gpt-oss-20b-MXFP4-Q8mlx | 12.1 GB observed | 14.4 GB | +47.6 GB | ✓ fits |
| 4bit | |
| gpt-neox-20bOpen weightsEleutherAIApache-2.0 | 20.7B | 12.4 GB estimated | +49.6 GB | ✓ fits |
| none recorded | |
| GPT-NeoXT-Chat-Base-20BOpen weightsTogether AIApache-2.0 | 20B | 12.0 GB estimated | +50.0 GB | ✓ fits |
| none recorded | |
| internlm2-base-20bOpen weightsInternLM (Shanghai AI Laboratory)Other | 20B | 12.0 GB estimated | +50.0 GB | ✓ fits |
| none recorded | |
| internlm2-20bOpen weightsInternLM (Shanghai AI Laboratory)Other | 20B | 12.0 GB estimated | +50.0 GB | ✓ fits |
| none recorded | |
| internlm2-chat-20bOpen weightsInternLM (Shanghai AI Laboratory)Other | 19.9B | 11.9 GB estimated | +50.1 GB | ✓ fits |
| none recorded | |
| pplx-qwen-3-8-27b-dflash2-20260819Open weightsPerplexity AIOther | 18.8B | 11.3 GB estimated | +50.7 GB | ✓ fits |
| none recorded | |
| pplx-computer-qwen-3-8-27b-dflash2-20260824Open weightsPerplexity AIOther | 18.8B | 11.3 GB estimated | +50.7 GB | ✓ fits |
| none recorded | |
| Kimi-VL-A3B-ThinkingOpen weightsMoonshot AIMIT | 16.4B | 9.9 GB estimated | +52.1 GB | ✓ fits |
| none recorded | |
| Kimi-VL-A3B-InstructOpen weightsMoonshot AIMIT | 16.4B | 9.9 GB estimated | +52.1 GB | ✓ fits |
| none recorded | |
| Kimi-VL-A3B-Thinking-2506Open weightsMoonshot AIMIT | 16.4B | 9.9 GB estimated | +52.1 GB | ✓ fits |
| none recorded | |
| Moonlight-16B-A3B-InstructOpen weightsMoonshot AIMIT | 16B | 9.7 GB estimated | +52.3 GB | ✓ fits |
| none recorded | |
| Moonlight-16B-A3BOpen weightsMoonshot AIMIT | 16B | 9.7 GB estimated | +52.3 GB | ✓ fits |
| none recorded | |
| DeepSeek-Coder-V2-Lite-InstructOpen weightsDeepSeekOther | 15.7B | 9.5 GB estimated | +52.5 GB | ✓ fits |
| 2 recorded | |
| ↳ bartowski/DeepSeek-Coder-V2-Lite-Instruct-GGUFgguf | — estimated | 9.5 GB | +52.5 GB | ✓ fits |
| 4bit | |
| ↳ bartowski/DeepSeek-Coder-V2-Lite-Instruct-GGUFgguf | — estimated | 9.5 GB | +52.5 GB | ✓ fits |
| 4bit | |
| DeepSeek-V2-LiteOpen weightsDeepSeekOther | 15.7B | 9.5 GB estimated | +52.5 GB | ✓ fits |
| none recorded | |
| BAGEL-7B-MoTOpen weightsByteDanceApache-2.0 | 14.7B | 8.9 GB estimated | +53.0 GB | ✓ fits |
| none recorded | |
| Phi 4Open weightsMicrosoftMIT | 14.7B | 8.9 GB estimated | +53.1 GB | ✓ fits |
| none recorded | |
| Ministral 3 14B 2512Open weightsMistral AIApache-2.0 | 13.9B | 8.5 GB estimated | +53.5 GB | ✓ fits |
| none recorded | |
| Ministral-3-14B-Reasoning-2512Open weightsMistral AIApache-2.0 | 13.9B | 8.5 GB estimated | +53.5 GB | ✓ fits |
| none recorded | |
| FireLLaVA-13brestricted-weightsFireworks AILlama-2-Community | 13.3B | 8.2 GB estimated | +53.8 GB | ✓ fits |
| none recorded | |
| Llama-2-13b-chat-hfrestricted-weightsMeta AILlama-2-Community | 13B | 8.0 GB estimated | +54.0 GB | ✓ fits |
| none recorded | |
| aya-101Open weightsCohereApache-2.0 | 12.9B | 7.9 GB estimated | +54.1 GB | ✓ fits |
| none recorded | |
| Mistral-Nemo-Base-2407Open weightsMistral AIApache-2.0 | 12.3B | 7.5 GB estimated | +54.5 GB | ✓ fits |
| none recorded | |
| Mistral NemoOpen weightsMistral AIApache-2.0 | 12.3B | 7.5 GB estimated | +54.5 GB | ✓ fits |
| none recorded | |
| gemma-4-12B-itOpen weightsGoogleApache-2.0 | 12B | 7.4 GB estimated | +54.6 GB | ✓ fits |
| none recorded | |
| FLUX.1-Fill-devOpen weightsBlack Forest LabsOther | 11.9B | 7.3 GB estimated | +54.7 GB | ✓ fits |
| none recorded | |
| FLUX.1-Kontext-devOpen weightsBlack Forest LabsOther | 11.9B | 7.3 GB estimated | +54.7 GB | ✓ fits |
| none recorded | |
| FLUX.1-Krea-devOpen weightsBlack Forest LabsOther | 11.9B | 7.3 GB estimated | +54.7 GB | ✓ fits |
| none recorded | |
| FLUX.1-devOpen weightsBlack Forest LabsOther | 11.9B | 7.3 GB estimated | +54.7 GB | ✓ fits |
| none recorded | |
| FLUX.1-schnellOpen weightsBlack Forest LabsApache-2.0 | 11.9B | 7.3 GB estimated | +54.7 GB | ✓ fits |
| none recorded | |
| KaLM-Embedding-Gemma3-12B-2511Open weightsTencentOther | 11.8B | 7.3 GB estimated | +54.7 GB | ✓ fits |
| none recorded | |
| Nous-Hermes-2-SOLAR-10.7BOpen weightsNous ResearchApache-2.0 | 10.7B | 6.7 GB estimated | +55.3 GB | ✓ fits |
| none recorded | |
| Llama-3.2-11B-Vision-Instructrestricted-weightsMeta AILlama-3.2-Community | 10.7B | 6.6 GB estimated | +55.4 GB | ✓ fits |
| none recorded | |
| GLM-4.6V-FlashOpen weightsZ.ai (Zhipu AI)MIT | 10.3B | 6.4 GB estimated | +55.6 GB | ✓ fits |
| none recorded | |
| GLM-4.1V-9B-ThinkingOpen weightsZ.ai (Zhipu AI)MIT | 10.3B | 6.4 GB estimated | +55.6 GB | ✓ fits |
| none recorded | |
| Kimi-Audio-7BOpen weightsMoonshot AIMIT | 9.77B | 6.1 GB estimated | +55.9 GB | ✓ fits |
| none recorded | |
| Kimi-Audio-7B-InstructOpen weightsMoonshot AIMIT | 9.77B | 6.1 GB estimated | +55.9 GB | ✓ fits |
| none recorded | |
| Qwen3.5-9BOpen weightsQwenApache-2.0 | 9.65B | 6.0 GB estimated | +56.0 GB | ✓ fits |
| 3 recorded | |
| ↳ Intel/Qwen3.5-9B-int4-AutoRound | 9.0 GB observed | 10.8 GB | +51.2 GB | ✓ fits |
| 4bit | |
| ↳ unsloth/Qwen3.5-9B-GGUFgguf | — estimated | 6.0 GB | +56.0 GB | ✓ fits |
| 4bit | |
| ↳ unsloth/Qwen3.5-9B-GGUFgguf | — estimated | 6.0 GB | +56.0 GB | ✓ fits |
| 4bit | |
| glm-4-9b-chatOpen weightsZ.ai (Zhipu AI)Other | 9.4B | 5.9 GB estimated | +56.1 GB | ✓ fits |
| none recorded | |
| academic-ds-9BOpen weightsByteDanceApache-2.0 | 9.37B | 5.9 GB estimated | +56.1 GB | ✓ fits |
| none recorded | |
| FLUX.2-klein-9BOpen weightsBlack Forest LabsOther | 9.08B | 5.7 GB estimated | +56.3 GB | ✓ fits |
| none recorded | |
| FLUX.2-klein-base-9BOpen weightsBlack Forest LabsOther | 9.08B | 5.7 GB estimated | +56.3 GB | ✓ fits |
| none recorded | |
| FLUX.2-klein-9b-kvOpen weightsBlack Forest LabsOther | 9.08B | 5.7 GB estimated | +56.3 GB | ✓ fits |
| 1 recorded | |
| ↳ black-forest-labs/FLUX.2-klein-9b-kv-fp8 | — estimated | 5.7 GB | +56.3 GB | ✓ fits |
| 4bit | |
| Qianfan-VL-8BOpen weightsBaiduOther | 8.81B | 5.6 GB estimated | +56.4 GB | ✓ fits |
| none recorded | |
| internlm3-8b-instructOpen weightsInternLM (Shanghai AI Laboratory)Apache-2.0 | 8.8B | 5.6 GB estimated | +56.4 GB | ✓ fits |
| 2 recorded | |
| ↳ internlm/internlm3-8b-instruct-ggufgguf | — estimated | 5.6 GB | +56.4 GB | ✓ fits |
| 4bit | |
| ↳ internlm/internlm3-8b-instruct-ggufgguf | — estimated | 5.6 GB | +56.4 GB | ✓ fits |
| 4bit | |
| granite-4.1-8bOpen weightsIBMApache-2.0 | 8.79B | 5.6 GB estimated | +56.4 GB | ✓ fits |
| none recorded | |
| Qwen3 VL 8B InstructOpen weightsQwenApache-2.0 | 8.77B | 5.5 GB estimated | +56.5 GB | ✓ fits |
| 1 recorded | |
| ↳ amd/Qwen3-VL-8B-Instruct-w8a8-llmcompressor | 10.6 GB observed | 12.7 GB | +49.3 GB | ✓ fits |
| 4bit | |
| VibeVoice-ASROpen weightsMicrosoftMIT | 8.67B | 5.5 GB estimated | +56.5 GB | ✓ fits |
| none recorded | |
| Molmo2-8BOpen weightsAllen Institute for AIApache-2.0 | 8.66B | 5.5 GB estimated | +56.5 GB | ✓ fits |
| none recorded | |
| aya-vision-8brestricted-weightsCohereCC-BY-NC-4.0 | 8.63B | 5.5 GB estimated | +56.5 GB | ✓ fits |
| none recorded | |
| Intern-S1-miniOpen weightsInternLM (Shanghai AI Laboratory)Apache-2.0 | 8.54B | 5.4 GB estimated | +56.6 GB | ✓ fits |
| none recorded | |
| GKA-primed-HQwen3-8B-ReasonerOpen weightsAmazon Web ServicesApache-2.0 | 8.5B | 5.4 GB estimated | +56.6 GB | ✓ fits |
| none recorded | |
| GDN-primed-HQwen3-8B-InstructOpen weightsAmazon Web ServicesApache-2.0 | 8.5B | 5.4 GB estimated | +56.6 GB | ✓ fits |
| none recorded | |
| LFM2.5-8B-A1BOpen weightsLiquid AIOther | 8.47B | 5.4 GB estimated | +56.6 GB | ✓ fits |
| 2 recorded | |
| ↳ LiquidAI/LFM2.5-8B-A1B-GGUFgguf | — estimated | 5.4 GB | +56.6 GB | ✓ fits |
| 4bit | |
| ↳ LiquidAI/LFM2.5-8B-A1B-GGUFgguf | — estimated | 5.4 GB | +56.6 GB | ✓ fits |
| 4bit | |
| olmOCR-2-7B-1025Open weightsAllen Institute for AIApache-2.0 | 8.29B | 5.3 GB estimated | +56.7 GB | ✓ fits |
| 1 recorded | |
| ↳ allenai/olmOCR-2-7B-1025-FP8 | 10.1 GB observed | 12.1 GB | +49.9 GB | ✓ fits |
| 4bit | |
| Qwen2.5-VL-7B-InstructOpen weightsQwenApache-2.0 | 8.29B | 5.3 GB estimated | +56.7 GB | ✓ fits |
| none recorded | |
| UI-TARS 7BOpen weightsByteDanceApache-2.0 | 8.29B | 5.3 GB estimated | +56.7 GB | ✓ fits |
| none recorded | |
| UI-TARS-7B-DPOOpen weightsByteDanceApache-2.0 | 8.29B | 5.3 GB estimated | +56.7 GB | ✓ fits |
| none recorded | |
| UI-TARS-7B-SFTOpen weightsByteDanceApache-2.0 | 8.29B | 5.3 GB estimated | +56.7 GB | ✓ fits |
| none recorded | |
| Seed-Coder-8B-InstructOpen weightsByteDanceMIT | 8.25B | 5.3 GB estimated | +56.8 GB | ✓ fits |
| none recorded | |
| Seed-Coder-8B-BaseOpen weightsByteDanceMIT | 8.25B | 5.3 GB estimated | +56.8 GB | ✓ fits |
| none recorded | |
| Seed-Coder-8B-ReasoningOpen weightsByteDanceMIT | 8.25B | 5.3 GB estimated | +56.8 GB | ✓ fits |
| none recorded | |
| Stable-DiffCoder-8B-BaseOpen weightsByteDanceMIT | 8.25B | 5.3 GB estimated | +56.8 GB | ✓ fits |
| none recorded | |
| Stable-DiffCoder-8B-InstructOpen weightsByteDanceMIT | 8.25B | 5.3 GB estimated | +56.8 GB | ✓ fits |
| none recorded | |
| DeepSeek-R1-0528-Qwen3-8BOpen weightsDeepSeekMIT | 8.19B | 5.2 GB estimated | +56.8 GB | ✓ fits |
| none recorded | |
| Qwen3 8BOpen weightsQwenApache-2.0 | 8.19B | 5.2 GB estimated | +56.8 GB | ✓ fits |
| 2 recorded | |
| ↳ mlx-community/Qwen3-8B-4bitmlx | 4.6 GB observed | 5.8 GB | +56.2 GB | ✓ fits |
| 4bit | |
| ↳ mlx-community/Qwen3-8B-4bitmlx | 4.6 GB observed | 5.8 GB | +56.2 GB | ✓ fits |
| 4bit | |
| granite-guardian-3.3-8bOpen weightsIBMApache-2.0 | 8.17B | 5.2 GB estimated | +56.8 GB | ✓ fits |
| none recorded | |
| granite-3.0-8b-instructOpen weightsIBMApache-2.0 | 8.17B | 5.2 GB estimated | +56.8 GB | ✓ fits |
| none recorded | |
| stable-diffusion-3.5-largeOpen weightsStability AIOther | 8.15B | 5.2 GB estimated | +56.8 GB | ✓ fits |
| none recorded | |
| ERNIE-Image-TurboOpen weightsBaiduApache-2.0 | 8.03B | 5.1 GB estimated | +56.9 GB | ✓ fits |
| none recorded | |
| ERNIE-ImageOpen weightsBaiduApache-2.0 | 8.03B | 5.1 GB estimated | +56.9 GB | ✓ fits |
| none recorded | |
| Meta-Llama-3-8B-Instructrestricted-weightsMeta AILlama-3-Community | 8.03B | 5.1 GB estimated | +56.9 GB | ✓ fits |
| none recorded | |
| Llama 3.1 8Brestricted-weightsMeta AILlama-3.1-Community | 8.03B | 5.1 GB estimated | +56.9 GB | ✓ fits |
| 1 recorded | |
| ↳ NousResearch/Hermes-3-Llama-3.1-8B-GGUFgguf | — estimated | 5.1 GB | +56.9 GB | ✓ fits |
| 4bit | |
| NousResearch/Meta-Llama-3.1-8Brestricted-weightsNous ResearchLlama-3.1-Community | 8.03B | 5.1 GB estimated | +56.9 GB | ✓ fits |
| none recorded | |
| NousResearch/Meta-Llama-3.1-8B-Instructrestricted-weightsNous ResearchLlama-3.1-Community | 8.03B | 5.1 GB estimated | +56.9 GB | ✓ fits |
| none recorded | |
| Hermes-2-Theta-Llama-3-8BOpen weightsNous ResearchApache-2.0 | 8.03B | 5.1 GB estimated | +56.9 GB | ✓ fits |
| none recorded | |
| Llama 3.1 8B Instructrestricted-weightsMeta AILlama-3.1-Community | 8.03B | 5.1 GB estimated | +56.9 GB | ✓ fits |
| 2 recorded | |
| ↳ mlx-community/Llama-3.1-8B-Instruct-4bitmlx | 4.5 GB observed | 5.7 GB | +56.3 GB | ✓ fits |
| 4bit | |
| ↳ bartowski/Meta-Llama-3.1-8B-Instruct-GGUFgguf | — estimated | 5.1 GB | +56.9 GB | ✓ fits |
| 4bit | |
| NousResearch/Meta-Llama-3-8BOpen weightsNous ResearchOther | 8.03B | 5.1 GB estimated | +56.9 GB | ✓ fits |
| none recorded | |
| Meta-Llama-3-8Brestricted-weightsMeta AILlama-3-Community | 8.03B | 5.1 GB estimated | +56.9 GB | ✓ fits |
| none recorded | |
| NousResearch/Meta-Llama-3-8B-InstructOpen weightsNous ResearchOther | 8.03B | 5.1 GB estimated | +56.9 GB | ✓ fits |
| none recorded | |
| Hermes-3-Llama-3.1-8Brestricted-weightsNous ResearchLlama-3-Community | 8.03B | 5.1 GB estimated | +56.9 GB | ✓ fits |
| 1 recorded | |
| ↳ NousResearch/Hermes-3-Llama-3.1-8B-GGUFgguf | — estimated | 5.1 GB | +56.9 GB | ✓ fits |
| 4bit | |
| Salesforce/Llama-xLAM-2-8b-fc-rrestricted-weightsSalesforceCC-BY-NC-4.0 | 8.03B | 5.1 GB estimated | +56.9 GB | ✓ fits |
| none recorded | |
| Hy-MT2-7BOpen weightsTencentApache-2.0 | 8.03B | 5.1 GB estimated | +56.9 GB | ✓ fits |
| 2 recorded | |
| ↳ tencent/Hy-MT2-7B-GGUFgguf | — estimated | 5.1 GB | +56.9 GB | ✓ fits |
| 4bit | |
| ↳ tencent/Hy-MT2-7B-GGUFgguf | — estimated | 5.1 GB | +56.9 GB | ✓ fits |
| 4bit | |
| c4ai-command-r7b-12-2024restricted-weightsCohereCC-BY-NC-4.0 | 8.03B | 5.1 GB estimated | +56.9 GB | ✓ fits |
| none recorded | |
| aya-expanse-8brestricted-weightsCohereCC-BY-NC-4.0 | 8.03B | 5.1 GB estimated | +56.9 GB | ✓ fits |
| none recorded | |
| aya-23-8Brestricted-weightsCohereCC-BY-NC-4.0 | 8.03B | 5.1 GB estimated | +56.9 GB | ✓ fits |
| none recorded | |
| Ministral-8B-Instruct-2410Open weightsMistral AIOther | 8.02B | 5.1 GB estimated | +56.9 GB | ✓ fits |
| none recorded | |
| gemma-4-E4B-itOpen weightsGoogleApache-2.0 | 8B | 5.1 GB estimated | +56.9 GB | ✓ fits |
| none recorded | |
| ERNIE-Image-AesOpen weightsBaiduApache-2.0 | 7.94B | 5.1 GB estimated | +56.9 GB | ✓ fits |
| none recorded | |
| internlm2_5-7b-chatOpen weightsInternLM (Shanghai AI Laboratory)Other | 7.74B | 5.0 GB estimated | +57.0 GB | ✓ fits |
| none recorded | |
| internlm2-chat-7bOpen weightsInternLM (Shanghai AI Laboratory)Other | 7.74B | 5.0 GB estimated | +57.0 GB | ✓ fits |
| none recorded | |
| StripedHyena-Hessian-7BOpen weightsTogether AIApache-2.0 | 7.65B | 4.9 GB estimated | +57.1 GB | ✓ fits |
| none recorded | |
| StripedHyena-Nous-7BOpen weightsTogether AIApache-2.0 | 7.65B | 4.9 GB estimated | +57.1 GB | ✓ fits |
| none recorded | |
| SynLogic-7BOpen weightsMiniMaxMIT | 7.62B | 4.9 GB estimated | +57.1 GB | ✓ fits |
| none recorded | |
| Qwen2.5 7B InstructOpen weightsQwenApache-2.0 | 7.62B | 4.9 GB estimated | +57.1 GB | ✓ fits |
| 2 recorded | |
| ↳ bartowski/Qwen2.5-7B-Instruct-GGUFgguf | — estimated | 4.9 GB | +57.1 GB | ✓ fits |
| 4bit | |
| ↳ bartowski/Qwen2.5-7B-Instruct-GGUFgguf | — estimated | 4.9 GB | +57.1 GB | ✓ fits |
| 4bit | |
| BFS-Prover-V2-7BOpen weightsByteDanceApache-2.0 | 7.62B | 4.9 GB estimated | +57.1 GB | ✓ fits |
| none recorded | |
| DeepSeek-R1-Distill-Qwen-7BOpen weightsDeepSeekMIT | 7.62B | 4.9 GB estimated | +57.1 GB | ✓ fits |
| none recorded | |
| Seed-X-PPO-7BOpen weightsByteDanceOther | 7.51B | 4.8 GB estimated | +57.2 GB | ✓ fits |
| none recorded | |
| Hunyuan-7B-InstructOpen weightsTencent | 7.5B | 4.8 GB estimated | +57.2 GB | ✓ fits |
| none recorded | |
| OLMo-2-1124-7B-InstructOpen weightsAllen Institute for AIApache-2.0 | 7.3B | 4.7 GB estimated | +57.3 GB | ✓ fits |
| none recorded | |
| OLMo-2-1124-7BOpen weightsAllen Institute for AIApache-2.0 | 7.3B | 4.7 GB estimated | +57.3 GB | ✓ fits |
| none recorded | |
| Olmo-3-7B-InstructOpen weightsAllen Institute for AIApache-2.0 | 7.3B | 4.7 GB estimated | +57.3 GB | ✓ fits |
| none recorded | |
| Olmo-3-7B-ThinkOpen weightsAllen Institute for AIApache-2.0 | 7.3B | 4.7 GB estimated | +57.3 GB | ✓ fits |
| none recorded | |
| Olmo-3-1025-7BOpen weightsAllen Institute for AIApache-2.0 | 7.3B | 4.7 GB estimated | +57.3 GB | ✓ fits |
| none recorded | |
| wildguardOpen weightsAllen Institute for AIApache-2.0 | 7.25B | 4.7 GB estimated | +57.3 GB | ✓ fits |
| none recorded | |
| Mistral-7B-v0.3Open weightsMistral AIApache-2.0 | 7.25B | 4.7 GB estimated | +57.3 GB | ✓ fits |
| none recorded | |
| Mistral-7B-Instruct-v0.3Open weightsMistral AIApache-2.0 | 7.25B | 4.7 GB estimated | +57.3 GB | ✓ fits |
| none recorded | |
| Mistral-7B-Instruct-v0.1Open weightsMistral AIApache-2.0 | 7.24B | 4.7 GB estimated | +57.3 GB | ✓ fits |
| none recorded | |
| Mistral-7B-v0.1Open weightsMistral AIApache-2.0 | 7.24B | 4.7 GB estimated | +57.3 GB | ✓ fits |
| 1 recorded | |
| ↳ NousResearch/Hermes-2-Pro-Mistral-7B-GGUFgguf | — estimated | 4.7 GB | +57.3 GB | ✓ fits |
| 4bit | |
| Mistral-7B-Instruct-v0.2Open weightsMistral AIApache-2.0 | 7.24B | 4.7 GB estimated | +57.3 GB | ✓ fits |
| none recorded | |
| SFR-Embedding-2_Rrestricted-weightsSalesforceCC-BY-NC-4.0 | 7.11B | 4.6 GB estimated | +57.4 GB | ✓ fits |
| none recorded | |
| SFR-Embedding-Mistralrestricted-weightsSalesforceCC-BY-NC-4.0 | 7.11B | 4.6 GB estimated | +57.4 GB | ✓ fits |
| none recorded | |
| SeedVR-7BOpen weightsByteDanceApache-2.0 | 7B | 4.5 GB estimated | +57.5 GB | ✓ fits |
| none recorded | |
| Qwen2-AudioOpen weightsQwen Team | 7B | 4.5 GB estimated | +57.5 GB | ✓ fits |
| none recorded | |
| internlm-xcomposer2-7bOpen weightsInternLM (Shanghai AI Laboratory)Other | 7B | 4.5 GB estimated | +57.5 GB | ✓ fits |
| none recorded | |
| fireworks-ai/mistral-7b-eagle-head-experimentalOpen weightsFireworks AIApache-2.0 | 7B | 4.5 GB estimated | +57.5 GB | ✓ fits |
| none recorded | |
| internlm2-base-7bOpen weightsInternLM (Shanghai AI Laboratory)Other | 7B | 4.5 GB estimated | +57.5 GB | ✓ fits |
| none recorded | |
| SeedVR2-7BOpen weightsByteDanceApache-2.0 | 7B | 4.5 GB estimated | +57.5 GB | ✓ fits |
| none recorded | |
| Qwen2.5-OmniOpen weightsQwen Team | 7B | 4.5 GB estimated | +57.5 GB | ✓ fits |
| none recorded | |
| togethercomputer/Llama-2-7B-32K-Instructrestricted-weightsTogether AILlama-2-Community | 7B | 4.5 GB estimated | +57.5 GB | ✓ fits |
| none recorded | |
| internlm-xcomposer-7bOpen weightsInternLM (Shanghai AI Laboratory)Apache-2.0 | 7B | 4.5 GB estimated | +57.5 GB | ✓ fits |
| none recorded | |
| togethercomputer/LLaMA-2-7B-32Krestricted-weightsTogether AILlama-2-Community | 7B | 4.5 GB estimated | +57.5 GB | ✓ fits |
| none recorded | |
| internlm2-7bOpen weightsInternLM (Shanghai AI Laboratory)Other | 7B | 4.5 GB estimated | +57.5 GB | ✓ fits |
| none recorded | |
| internlm-chat-7bOpen weightsInternLM (Shanghai AI Laboratory) | 7B | 4.5 GB estimated | +57.5 GB | ✓ fits |
| none recorded | |
| Pythia-Chat-Base-7BOpen weightsTogether AIApache-2.0 | 7B | 4.5 GB estimated | +57.5 GB | ✓ fits |
| none recorded | |
| RedPajama-INCITE-7B-BaseOpen weightsTogether AIApache-2.0 | 7B | 4.5 GB estimated | +57.5 GB | ✓ fits |
| none recorded | |
| RedPajama-INCITE-7B-ChatOpen weightsTogether AIApache-2.0 | 7B | 4.5 GB estimated | +57.5 GB | ✓ fits |
| none recorded | |
| RedPajama-INCITE-7B-InstructOpen weightsTogether AIApache-2.0 | 7B | 4.5 GB estimated | +57.5 GB | ✓ fits |
| none recorded | |
| pythia-6.9bOpen weightsEleutherAIApache-2.0 | 6.99B | 4.5 GB estimated | +57.5 GB | ✓ fits |
| none recorded | |
| OLMoE-1B-7B-0924Open weightsAllen Institute for AIApache-2.0 | 6.92B | 4.5 GB estimated | +57.5 GB | ✓ fits |
| none recorded | |
| OLMoE-1B-7B-0125-InstructOpen weightsAllen Institute for AIApache-2.0 | 6.92B | 4.5 GB estimated | +57.5 GB | ✓ fits |
| none recorded | |
| deepseek-coder-7b-instruct-v1.5Open weightsDeepSeekOther | 6.91B | 4.5 GB estimated | +57.5 GB | ✓ fits |
| none recorded | |
| deepseek-coder-6.7b-instructOpen weightsDeepSeekOther | 6.74B | 4.4 GB estimated | +57.6 GB | ✓ fits |
| none recorded | |
| NousResearch/Llama-2-7b-hfOpen weightsNous Research | 6.74B | 4.4 GB estimated | +57.6 GB | ✓ fits |
| none recorded | |
| NousResearch/Llama-2-7b-chat-hfOpen weightsNous Research | 6.74B | 4.4 GB estimated | +57.6 GB | ✓ fits |
| none recorded | |
| Llama-2-7b-hfrestricted-weightsMeta AILlama-2-Community | 6.74B | 4.4 GB estimated | +57.6 GB | ✓ fits |
| none recorded | |
| Llama-2-7b-chat-hfrestricted-weightsMeta AILlama-2-Community | 6.74B | 4.4 GB estimated | +57.6 GB | ✓ fits |
| none recorded | |
| Nous-Hermes-llama-2-7bOpen weightsNous ResearchMIT | 6.74B | 4.4 GB estimated | +57.6 GB | ✓ fits |
| none recorded | |
| evo-1-8k-baseOpen weightsTogether AIApache-2.0 | 6.45B | 4.2 GB estimated | +57.8 GB | ✓ fits |
| none recorded | |
| evo-1-131k-baseOpen weightsTogether AIApache-2.0 | 6.45B | 4.2 GB estimated | +57.8 GB | ✓ fits |
| none recorded | |
| chatglm3-6bOpen weightsZ.ai (Zhipu AI) | 6.24B | 4.1 GB estimated | +57.9 GB | ✓ fits |
| none recorded | |
| GPT-JT-6B-v1Open weightsTogether AIApache-2.0 | 6B | 4.0 GB estimated | +58.0 GB | ✓ fits |
| none recorded | |
| gpt-j-6bOpen weightsEleutherAIApache-2.0 | 6B | 4.0 GB estimated | +58.0 GB | ✓ fits |
| none recorded | |
| chatglm2-6bOpen weightsZ.ai (Zhipu AI) | 6B | 4.0 GB estimated | +58.0 GB | ✓ fits |
| none recorded | |
| gemma-4-E2B-itOpen weightsGoogleApache-2.0 | 5.12B | 3.4 GB estimated | +58.6 GB | ✓ fits |
| none recorded | |
| Molmo2-4BOpen weightsAllen Institute for AIApache-2.0 | 4.85B | 3.3 GB estimated | +58.7 GB | ✓ fits |
| none recorded | |
| Qianfan-OCROpen weightsBaiduApache-2.0 | 4.74B | 3.2 GB estimated | +58.8 GB | ✓ fits |
| none recorded | |
| Voxtral-Mini-3B-2507Open weightsMistral AIApache-2.0 | 4.68B | 3.2 GB estimated | +58.8 GB | ✓ fits |
| none recorded | |
| Qwen3.5-4BOpen weightsQwenApache-2.0 | 4.66B | 3.2 GB estimated | +58.8 GB | ✓ fits |
| 5 recorded | |
| ↳ Intel/Qwen3.5-4B-int4-AutoRound | 4.6 GB observed | 5.8 GB | +56.2 GB | ✓ fits |
| 4bit | |
| ↳ unsloth/Qwen3.5-4B-GGUFgguf | — estimated | 3.2 GB | +58.8 GB | ✓ fits |
| 4bit | |
| ↳ bartowski/Qwen_Qwen3.5-4B-GGUFgguf | — estimated | 3.2 GB | +58.8 GB | ✓ fits |
| 4bit | |
| ↳ bartowski/Qwen_Qwen3.5-4B-GGUFgguf | — estimated | 3.2 GB | +58.8 GB | ✓ fits |
| 4bit | |
| +1 more artifacts on the model page. | |||||||
| Qwen3-VL-4B-InstructOpen weightsQwenApache-2.0 | 4.44B | 3.0 GB estimated | +59.0 GB | ✓ fits |
| none recorded | |
| Voxtral-Mini-4B-Realtime-2602Open weightsMistral AIApache-2.0 | 4.43B | 3.0 GB estimated | +59.0 GB | ✓ fits |
| none recorded | |
| Gemma 3 4Brestricted-weightsGoogleGemma-Terms | 4.3B | 3.0 GB estimated | +59.0 GB | ✓ fits |
| 1 recorded | |
| ↳ mlx-community/gemma-3-4b-it-qat-4bitmlx | 3.0 GB observed | 4.0 GB | +58.0 GB | ✓ fits |
| 4bit | |
| Ministral-3-3B-Reasoning-2512Open weightsMistral AIApache-2.0 | 4.25B | 3.0 GB estimated | +59.0 GB | ✓ fits |
| none recorded | |
| Phi-3.5-vision-instructOpen weightsMicrosoftMIT | 4.15B | 2.9 GB estimated | +59.1 GB | ✓ fits |
| none recorded | |
| instructblip-flan-t5-xlOpen weightsSalesforceMIT | 4.02B | 2.8 GB estimated | +59.2 GB | ✓ fits |
| none recorded | |
| pplx-embed-context-v1-4bOpen weightsPerplexity AIMIT | 4.02B | 2.8 GB estimated | +59.2 GB | ✓ fits |
| none recorded | |
| pplx-embed-v1-4bOpen weightsPerplexity AIMIT | 4.02B | 2.8 GB estimated | +59.2 GB | ✓ fits |
| none recorded | |
| Qwen3-4BOpen weightsQwenApache-2.0 | 4.02B | 2.8 GB estimated | +59.2 GB | ✓ fits |
| none recorded | |
| TRELLIS.2-4BOpen weightsMicrosoftMIT | 4B | 2.8 GB estimated | +59.2 GB | ✓ fits |
| none recorded | |
| granite-vision-4.1-4bOpen weightsIBMApache-2.0 | 4B | 2.8 GB estimated | +59.2 GB | ✓ fits |
| none recorded | |
| blip2-flan-t5-xlOpen weightsSalesforceMIT | 3.94B | 2.8 GB estimated | +59.2 GB | ✓ fits |
| none recorded | |
| FLUX.2-klein-4BOpen weightsBlack Forest LabsApache-2.0 | 3.88B | 2.7 GB estimated | +59.3 GB | ✓ fits |
| none recorded | |
| FLUX.2-klein-base-4BOpen weightsBlack Forest LabsApache-2.0 | 3.88B | 2.7 GB estimated | +59.3 GB | ✓ fits |
| none recorded | |
| Cosmos3-EdgeOpen weightsNVIDIAOther | 3.86B | 2.7 GB estimated | +59.3 GB | ✓ fits |
| none recorded | |
| Ministral 3 3B 2512Open weightsMistral AIApache-2.0 | 3.85B | 2.7 GB estimated | +59.3 GB | ✓ fits |
| none recorded | |
| 3b-zh-pretrain-research_releaserestricted-weightsCanopy LabsLlama-3.2-Community | 3.78B | 2.7 GB estimated | +59.3 GB | ✓ fits |
| none recorded | |
| orpheus-3b-0.1-ftOpen weightsCanopy LabsApache-2.0 | 3.78B | 2.7 GB estimated | +59.3 GB | ✓ fits |
| none recorded | |
| 3b-ko-pretrain-research_releaseOpen weightsCanopy LabsApache-2.0 | 3.78B | 2.7 GB estimated | +59.3 GB | ✓ fits |
| none recorded | |
| 3b-fr-pretrain-research_releaserestricted-weightsCanopy LabsLlama-3.2-Community | 3.78B | 2.7 GB estimated | +59.3 GB | ✓ fits |
| none recorded | |
| 3b-hi-pretrain-research_releaserestricted-weightsCanopy LabsLlama-3.2-Community | 3.78B | 2.7 GB estimated | +59.3 GB | ✓ fits |
| none recorded | |
| 3b-de-pretrain-research_releaserestricted-weightsCanopy LabsLlama-3.2-Community | 3.78B | 2.7 GB estimated | +59.3 GB | ✓ fits |
| none recorded | |
| 3b-es_it-pretrain-research_releaserestricted-weightsCanopy LabsLlama-3.2-Community | 3.78B | 2.7 GB estimated | +59.3 GB | ✓ fits |
| none recorded | |
| orpheus-3b-0.1-pretrainedOpen weightsCanopy LabsApache-2.0 | 3.78B | 2.7 GB estimated | +59.3 GB | ✓ fits |
| none recorded | |
| blip2-opt-2.7bOpen weightsSalesforceMIT | 3.74B | 2.6 GB estimated | +59.4 GB | ✓ fits |
| none recorded | |
| Qianfan-VL-3BOpen weightsBaiduOther | 3.71B | 2.6 GB estimated | +59.4 GB | ✓ fits |
| none recorded | |
| M1-3BOpen weightsTogether AIMIT | 3.45B | 2.5 GB estimated | +59.5 GB | ✓ fits |
| none recorded | |
| granite-4.1-3bOpen weightsIBMApache-2.0 | 3.4B | 2.5 GB estimated | +59.5 GB | ✓ fits |
| none recorded | |
| DeepSeek-OCR-2Open weightsDeepSeekApache-2.0 | 3.39B | 2.4 GB estimated | +59.6 GB | ✓ fits |
| none recorded | |
| tiny-aya-globalrestricted-weightsCohereCC-BY-NC-4.0 | 3.35B | 2.4 GB estimated | +59.6 GB | ✓ fits |
| none recorded | |
| tiny-aya-baserestricted-weightsCohereCC-BY-NC-4.0 | 3.35B | 2.4 GB estimated | +59.6 GB | ✓ fits |
| none recorded | |
| DeepSeek-OCROpen weightsDeepSeekMIT | 3.34B | 2.4 GB estimated | +59.6 GB | ✓ fits |
| none recorded | |
| Unlimited-OCROpen weightsBaiduMIT | 3.34B | 2.4 GB estimated | +59.6 GB | ✓ fits |
| none recorded | |
| 3b-fr-ft-research_releaseOpen weightsCanopy LabsApache-2.0 | 3.3B | 2.4 GB estimated | +59.6 GB | ✓ fits |
| none recorded | |
| 3b-hi-ft-research_releaseOpen weightsCanopy LabsApache-2.0 | 3.3B | 2.4 GB estimated | +59.6 GB | ✓ fits |
| none recorded | |
| 3b-ko-ft-research_releaseOpen weightsCanopy LabsApache-2.0 | 3.3B | 2.4 GB estimated | +59.6 GB | ✓ fits |
| none recorded | |
| 3b-es_it-ft-research_releaseOpen weightsCanopy LabsApache-2.0 | 3.3B | 2.4 GB estimated | +59.6 GB | ✓ fits |
| none recorded | |
| 3b-de-ft-research_releaseOpen weightsCanopy LabsApache-2.0 | 3.3B | 2.4 GB estimated | +59.6 GB | ✓ fits |
| none recorded | |
| 3b-zh-ft-research_releaseOpen weightsCanopy LabsApache-2.0 | 3.3B | 2.4 GB estimated | +59.6 GB | ✓ fits |
| none recorded | |
| Llama 3.2 3B Instructrestricted-weightsMeta AILlama-3.2-Community | 3.21B | 2.4 GB estimated | +59.6 GB | ✓ fits |
| none recorded | |
| Llama-3.2-3Brestricted-weightsMeta AILlama-3.2-Community | 3.21B | 2.4 GB estimated | +59.6 GB | ✓ fits |
| none recorded | |
| AI21-Jamba-Reasoning-3BOpen weightsAI21 LabsApache-2.0 | 3.2B | 2.3 GB estimated | +59.7 GB | ✓ fits |
| 1 recorded | |
| ↳ ai21labs/AI21-Jamba-Reasoning-3B-GGUFgguf | — estimated | 2.3 GB | +59.7 GB | ✓ fits |
| 4bit | |
| LFM2.5-VL-3BOpen weightsLiquid AIOther | 3.12B | 2.3 GB estimated | +59.7 GB | ✓ fits |
| 2 recorded | |
| ↳ LiquidAI/LFM2.5-VL-3B-GGUFgguf | — estimated | 2.3 GB | +59.7 GB | ✓ fits |
| 4bit | |
| ↳ LiquidAI/LFM2.5-VL-3B-GGUFgguf | — estimated | 2.3 GB | +59.7 GB | ✓ fits |
| 4bit | |
| Qwen2.5-3B-InstructOpen weightsQwenOther | 3.09B | 2.3 GB estimated | +59.7 GB | ✓ fits |
| none recorded | |
| SmolLM3-3BOpen weightsHugging FaceApache-2.0 | 3.08B | 2.3 GB estimated | +59.7 GB | ✓ fits |
| none recorded | |
| SmolLM3-3B-BaseOpen weightsHugging FaceApache-2.0 | 3.08B | 2.3 GB estimated | +59.7 GB | ✓ fits |
| none recorded | |
| AI21-Jamba2-3BOpen weightsAI21 LabsApache-2.0 | 3.03B | 2.2 GB estimated | +59.8 GB | ✓ fits |
| none recorded | |
| RedPajama-INCITE-Chat-3B-v1Open weightsTogether AIApache-2.0 | 3B | 2.2 GB estimated | +59.8 GB | ✓ fits |
| none recorded | |
| SeedVR2-3BOpen weightsByteDanceApache-2.0 | 3B | 2.2 GB estimated | +59.8 GB | ✓ fits |
| none recorded | |
| SeedVR-3BOpen weightsByteDanceApache-2.0 | 3B | 2.2 GB estimated | +59.8 GB | ✓ fits |
| none recorded | |
| RedPajama-INCITE-Instruct-3B-v1Open weightsTogether AIApache-2.0 | 3B | 2.2 GB estimated | +59.8 GB | ✓ fits |
| none recorded | |
| RedPajama-INCITE-Base-3B-v1Open weightsTogether AIApache-2.0 | 3B | 2.2 GB estimated | +59.8 GB | ✓ fits |
| none recorded | |
| granite-vision-3.3-2bOpen weightsIBMApache-2.0 | 2.98B | 2.2 GB estimated | +59.8 GB | ✓ fits |
| none recorded | |
| pythia-2.8bOpen weightsEleutherAIApache-2.0 | 2.91B | 2.2 GB estimated | +59.8 GB | ✓ fits |
| none recorded | |
| stablelm-3b-4e1tOpen weightsStability AICC-BY-SA-4.0 | 2.8B | 2.1 GB estimated | +59.9 GB | ✓ fits |
| none recorded | |
| phi-2Open weightsMicrosoftMIT | 2.78B | 2.1 GB estimated | +59.9 GB | ✓ fits |
| none recorded | |
| WeMM-Embedding-2BOpen weightsTencentOther | 2.72B | 2.1 GB estimated | +59.9 GB | ✓ fits |
| none recorded | |
| LFM2.5-2.6B (free)Open weightsLiquid AIOther | 2.7B | 2.0 GB estimated | +60.0 GB | ✓ fits |
| 1 recorded | |
| ↳ LiquidAI/LFM2.5-2.6B-GGUFgguf | — estimated | 2.0 GB | +60.0 GB | ✓ fits |
| 4bit | |
| stable-diffusion-xl-base-1.0restricted-weightsStability AIOpenRAIL++-M | 2.57B | 2.0 GB estimated | +60.0 GB | ✓ fits |
| none recorded | |
| sdxl-turboOpen weightsStability AIOther | 2.57B | 2.0 GB estimated | +60.0 GB | ✓ fits |
| none recorded | |
| North-Micro-Vision-InstructOpen weightsCohereApache-2.0 | 2.48B | 1.9 GB estimated | +60.1 GB | ✓ fits |
| none recorded | |
| stable-diffusion-3.5-mediumOpen weightsStability AIOther | 2.47B | 1.9 GB estimated | +60.1 GB | ✓ fits |
| none recorded | |
| UI-TARS-2B-SFTOpen weightsByteDanceApache-2.0 | 2.44B | 1.9 GB estimated | +60.1 GB | ✓ fits |
| none recorded | |
| Cosmos-Reason2-2BOpen weightsNVIDIAOther | 2.44B | 1.9 GB estimated | +60.1 GB | ✓ fits |
| none recorded | |
| MiniMax-Music3Open weightsMiniMax | 2.43B | 1.9 GB estimated | +60.1 GB | ✓ fits |
| none recorded | |
| granite-speech-4.1-2bOpen weightsIBMApache-2.0 | 2.31B | 1.8 GB estimated | +60.2 GB | ✓ fits |
| none recorded | |
| stable-audio-3-mediumOpen weightsStability AIOther | 2.31B | 1.8 GB estimated | +60.2 GB | ✓ fits |
| none recorded | |
| stable-diffusion-xl-refiner-1.0restricted-weightsStability AIOpenRAIL++-M | 2.26B | 1.8 GB estimated | +60.2 GB | ✓ fits |
| none recorded | |
| GLM-ASR-Nano-2512Open weightsZ.ai (Zhipu AI)MIT | 2.26B | 1.8 GB estimated | +60.2 GB | ✓ fits |
| none recorded | |
| granite-speech-4.1-2b-narOpen weightsIBMApache-2.0 | 2.25B | 1.8 GB estimated | +60.2 GB | ✓ fits |
| none recorded | |
| SmolVLM2-2.2B-InstructOpen weightsHugging FaceApache-2.0 | 2.25B | 1.8 GB estimated | +60.2 GB | ✓ fits |
| none recorded | |
| SmolVLM-InstructOpen weightsHugging FaceApache-2.0 | 2.25B | 1.8 GB estimated | +60.2 GB | ✓ fits |
| none recorded | |
| granite-speech-4.1-2b-plusOpen weightsIBMApache-2.0 | 2.11B | 1.7 GB estimated | +60.3 GB | ✓ fits |
| none recorded | |
| stable-diffusion-3-medium-diffusersOpen weightsStability AIOther | 2.08B | 1.7 GB estimated | +60.3 GB | ✓ fits |
| none recorded | |
| cohere-transcribe-03-2026Open sourceCohereApache-2.0 | 2.07B | 1.7 GB estimated | +60.3 GB | ✓ fits |
| none recorded | |
| cohere-transcribe-arabic-07-2026Open weightsCohereApache-2.0 | 2.07B | 1.7 GB estimated | +60.3 GB | ✓ fits |
| none recorded | |
| Hy-MT2-1.8BOpen weightsTencentApache-2.0 | 2.04B | 1.7 GB estimated | +60.3 GB | ✓ fits |
| 2 recorded | |
| ↳ tencent/Hy-MT2-1.8B-GGUFgguf | — estimated | 1.7 GB | +60.3 GB | ✓ fits |
| 4bit | |
| ↳ tencent/Hy-MT2-1.8B-GGUFgguf | — estimated | 1.7 GB | +60.3 GB | ✓ fits |
| 4bit | |
| Youtu-LLM-2BOpen weightsTencentOther | 1.96B | 1.6 GB estimated | +60.4 GB | ✓ fits |
| none recorded | |
| FastVLM-1.5Brestricted-weightsAppleApple-AMLR | 1.91B | 1.6 GB estimated | +60.4 GB | ✓ fits |
| none recorded | |
| internlm2_5-1_8b-chatOpen weightsInternLM (Shanghai AI Laboratory)Other | 1.89B | 1.6 GB estimated | +60.4 GB | ✓ fits |
| none recorded | |
| siglip2-giant-opt-patch16-384Open weightsGoogleApache-2.0 | 1.87B | 1.6 GB estimated | +60.4 GB | ✓ fits |
| none recorded | |
| amazon/GPT-OSS-20B-P-EAGLEOpen weightsAmazon Web ServicesApache-2.0 | 1.8B | 1.5 GB estimated | +60.5 GB | ✓ fits |
| none recorded | |
Headroom = total device memory − reserve − estimate. 34 of 300 models have quantized artifacts recorded; artifact rows use the observed file size when a source publishes it, otherwise the estimate. Sorted by the API (fitting models first). Compare shortlisted models with the + buttons.
MethodEvery figure is an ESTIMATE: weights = params × bytes/param (× 1.15 overhead) unless an artifact's observed file size is available; KV cache uses architecture metadata when known, else 0.5 GB per 8K tokens × batch. /methodology