Skip to content
AI Atlas
Hardware HardwareSystemActive

MacBook Air (Apple M5)

Applespec sheet data quality79

Memory
16–32 GB
Bandwidth
153 GB/s
TDP

Specifications

As published

Every value opens its evidence (source, tier, observed time).

Kind

Source:Apple — Mac tech specsT1observed 11 h agohigh

Kind raw
computer
Spec sheet

Source:Apple — Mac tech specsT1observed 11 h agohigh

Memory

Source:Apple — Mac tech specsT1observed 11 h agohigh

Memory type

Source:Apple — Mac tech specsT1observed 11 h agohigh

Product line

Source:Apple — Mac tech specsT1observed 11 h agohigh

Memory bandwidth

Source:Apple — Mac tech specsT1observed 11 h agohigh

What can this run?

Models estimated to fit Estimated

Estimate from each model's parameter count, the chosen quantization and context — never a measurement.

Open in Run locally
Quantization4bit8bitfp16
Configuration16 GB24 GB32 GB

EstimatedAll fit figures are estimates, not measurements.355 of 536 evaluated models fit

Assumptions (7)
  • Estimated, not measured: weights = parameters × bytes/param × 1.15 runtime overhead (or the observed artifact file size when one is recorded).
  • bytes/param: 4bit = 0.5, 8bit = 1.0, fp16 = 2.0 (uniform quantization, no per-layer exceptions).
  • KV cache: 2 × layers × kv_heads × head_dim × 2 bytes × context × batch when the architecture is known; otherwise 0.5 GB per 8 192 tokens (× batch), independent of architecture (GQA/MLA models need less).
  • A model 'fits' when the estimate is at most the device memory minus 2 GB reserved for the OS and framework.
  • Mixture-of-experts models are estimated on total parameters (all experts must be resident); active parameters are ignored.
  • Device memory uses the largest configuration when several are listed (e.g. Apple silicon tiers).
  • Multi-GPU: device memories are summed; interconnect bandwidth, tensor-parallel replication and pipeline bubbles are not modelled.

This device ships in 3 memory configurations (16 / 24 / 32 GB). Estimates use 16 GB — pick another configuration above. Multiple units: use Run locally (memories are summed; interconnect is not modelled).

ModelParamsEst. memoryHeadroomFitsBreakdown
GLM-4.6V-FlashOpen weightsZ.ai (Zhipu AI)10.3B13.8 GB estimated+0.2 GB✓ fits
  • weights 10.3 GB estimated
  • KV cache 2.00 GB (heuristic)
  • overhead 1.5 GB
  • reserved 2 GB
  • context 32,768 × batch 1
GLM-4.1V-9B-ThinkingOpen weightsZ.ai (Zhipu AI)10.3B13.8 GB estimated+0.2 GB✓ fits
  • weights 10.3 GB estimated
  • KV cache 2.00 GB (heuristic)
  • overhead 1.5 GB
  • reserved 2 GB
  • context 32,768 × batch 1
Kimi-Audio-7BOpen weightsMoonshot AI9.77B13.2 GB estimated+0.8 GB✓ fits
  • weights 9.8 GB estimated
  • KV cache 2.00 GB (heuristic)
  • overhead 1.5 GB
  • reserved 2 GB
  • context 32,768 × batch 1
Kimi-Audio-7B-InstructOpen weightsMoonshot AI9.77B13.2 GB estimated+0.8 GB✓ fits
  • weights 9.8 GB estimated
  • KV cache 2.00 GB (heuristic)
  • overhead 1.5 GB
  • reserved 2 GB
  • context 32,768 × batch 1
Qwen3.5-9BOpen weightsQwen9.65B13.1 GB estimated+0.9 GB✓ fits
  • weights 9.7 GB estimated
  • KV cache 2.00 GB (heuristic)
  • overhead 1.4 GB
  • reserved 2 GB
  • context 32,768 × batch 1
glm-4-9b-chatOpen weightsZ.ai (Zhipu AI)9.4B12.8 GB estimated+1.2 GB✓ fits
  • weights 9.4 GB estimated
  • KV cache 2.00 GB (heuristic)
  • overhead 1.4 GB
  • reserved 2 GB
  • context 32,768 × batch 1
academic-ds-9BOpen weightsByteDance9.37B12.8 GB estimated+1.2 GB✓ fits
  • weights 9.4 GB estimated
  • KV cache 2.00 GB (heuristic)
  • overhead 1.4 GB
  • reserved 2 GB
  • context 32,768 × batch 1
FLUX.2-klein-base-9BOpen weightsBlack Forest Labs9.08B12.4 GB estimated+1.6 GB✓ fits
  • weights 9.1 GB estimated
  • KV cache 2.00 GB (heuristic)
  • overhead 1.4 GB
  • reserved 2 GB
  • context 32,768 × batch 1
FLUX.2-klein-9BOpen weightsBlack Forest Labs9.08B12.4 GB estimated+1.6 GB✓ fits
  • weights 9.1 GB estimated
  • KV cache 2.00 GB (heuristic)
  • overhead 1.4 GB
  • reserved 2 GB
  • context 32,768 × batch 1
FLUX.2-klein-9b-kvOpen weightsBlack Forest Labs9.08B12.4 GB estimated+1.6 GB✓ fits
  • weights 9.1 GB estimated
  • KV cache 2.00 GB (heuristic)
  • overhead 1.4 GB
  • reserved 2 GB
  • context 32,768 × batch 1
Qianfan-VL-8BOpen weightsBaidu8.81B12.1 GB estimated+1.9 GB✓ fits
  • weights 8.8 GB estimated
  • KV cache 2.00 GB (heuristic)
  • overhead 1.3 GB
  • reserved 2 GB
  • context 32,768 × batch 1
internlm3-8b-instructOpen weightsInternLM (Shanghai AI Laboratory)8.8B12.1 GB estimated+1.9 GB✓ fits
  • weights 8.8 GB estimated
  • KV cache 2.00 GB (heuristic)
  • overhead 1.3 GB
  • reserved 2 GB
  • context 32,768 × batch 1
granite-4.1-8bOpen weightsIBM8.79B12.1 GB estimated+1.9 GB✓ fits
  • weights 8.8 GB estimated
  • KV cache 2.00 GB (heuristic)
  • overhead 1.3 GB
  • reserved 2 GB
  • context 32,768 × batch 1
Qwen3 VL 8B InstructOpen weightsQwen8.77B12.1 GB estimated+1.9 GB✓ fits
  • weights 8.8 GB estimated
  • KV cache 2.00 GB (heuristic)
  • overhead 1.3 GB
  • reserved 2 GB
  • context 32,768 × batch 1
VibeVoice-ASROpen weightsMicrosoft8.67B12.0 GB estimated+2.0 GB✓ fits
  • weights 8.7 GB estimated
  • KV cache 2.00 GB (heuristic)
  • overhead 1.3 GB
  • reserved 2 GB
  • context 32,768 × batch 1
Molmo2-8BOpen weightsAllen Institute for AI8.66B12.0 GB estimated+2.0 GB✓ fits
  • weights 8.7 GB estimated
  • KV cache 2.00 GB (heuristic)
  • overhead 1.3 GB
  • reserved 2 GB
  • context 32,768 × batch 1
aya-vision-8brestricted-weightsCohere8.63B11.9 GB estimated+2.1 GB✓ fits
  • weights 8.6 GB estimated
  • KV cache 2.00 GB (heuristic)
  • overhead 1.3 GB
  • reserved 2 GB
  • context 32,768 × batch 1
Intern-S1-miniOpen weightsInternLM (Shanghai AI Laboratory)8.54B11.8 GB estimated+2.2 GB✓ fits
  • weights 8.5 GB estimated
  • KV cache 2.00 GB (heuristic)
  • overhead 1.3 GB
  • reserved 2 GB
  • context 32,768 × batch 1
GKA-primed-HQwen3-8B-ReasonerOpen weightsAmazon Web Services8.5B11.8 GB estimated+2.2 GB✓ fits
  • weights 8.5 GB estimated
  • KV cache 2.00 GB (heuristic)
  • overhead 1.3 GB
  • reserved 2 GB
  • context 32,768 × batch 1
GDN-primed-HQwen3-8B-InstructOpen weightsAmazon Web Services8.5B11.8 GB estimated+2.2 GB✓ fits
  • weights 8.5 GB estimated
  • KV cache 2.00 GB (heuristic)
  • overhead 1.3 GB
  • reserved 2 GB
  • context 32,768 × batch 1
LFM2.5-8B-A1BOpen weightsLiquid AI8.47B11.7 GB estimated+2.3 GB✓ fits
  • weights 8.5 GB estimated
  • KV cache 2.00 GB (heuristic)
  • overhead 1.3 GB
  • reserved 2 GB
  • context 32,768 × batch 1
Qwen2.5-VL-7B-InstructOpen weightsQwen8.29B11.5 GB estimated+2.5 GB✓ fits
  • weights 8.3 GB estimated
  • KV cache 2.00 GB (heuristic)
  • overhead 1.2 GB
  • reserved 2 GB
  • context 32,768 × batch 1
UI-TARS 7BOpen weightsByteDance8.29B11.5 GB estimated+2.5 GB✓ fits
  • weights 8.3 GB estimated
  • KV cache 2.00 GB (heuristic)
  • overhead 1.2 GB
  • reserved 2 GB
  • context 32,768 × batch 1
olmOCR-2-7B-1025Open weightsAllen Institute for AI8.29B11.5 GB estimated+2.5 GB✓ fits
  • weights 8.3 GB estimated
  • KV cache 2.00 GB (heuristic)
  • overhead 1.2 GB
  • reserved 2 GB
  • context 32,768 × batch 1
UI-TARS-7B-SFTOpen weightsByteDance8.29B11.5 GB estimated+2.5 GB✓ fits
  • weights 8.3 GB estimated
  • KV cache 2.00 GB (heuristic)
  • overhead 1.2 GB
  • reserved 2 GB
  • context 32,768 × batch 1
UI-TARS-7B-DPOOpen weightsByteDance8.29B11.5 GB estimated+2.5 GB✓ fits
  • weights 8.3 GB estimated
  • KV cache 2.00 GB (heuristic)
  • overhead 1.2 GB
  • reserved 2 GB
  • context 32,768 × batch 1
Seed-Coder-8B-ReasoningOpen weightsByteDance8.25B11.5 GB estimated+2.5 GB✓ fits
  • weights 8.3 GB estimated
  • KV cache 2.00 GB (heuristic)
  • overhead 1.2 GB
  • reserved 2 GB
  • context 32,768 × batch 1
Seed-Coder-8B-InstructOpen weightsByteDance8.25B11.5 GB estimated+2.5 GB✓ fits
  • weights 8.3 GB estimated
  • KV cache 2.00 GB (heuristic)
  • overhead 1.2 GB
  • reserved 2 GB
  • context 32,768 × batch 1
Stable-DiffCoder-8B-InstructOpen weightsByteDance8.25B11.5 GB estimated+2.5 GB✓ fits
  • weights 8.3 GB estimated
  • KV cache 2.00 GB (heuristic)
  • overhead 1.2 GB
  • reserved 2 GB
  • context 32,768 × batch 1
Stable-DiffCoder-8B-BaseOpen weightsByteDance8.25B11.5 GB estimated+2.5 GB✓ fits
  • weights 8.3 GB estimated
  • KV cache 2.00 GB (heuristic)
  • overhead 1.2 GB
  • reserved 2 GB
  • context 32,768 × batch 1
Seed-Coder-8B-BaseOpen weightsByteDance8.25B11.5 GB estimated+2.5 GB✓ fits
  • weights 8.3 GB estimated
  • KV cache 2.00 GB (heuristic)
  • overhead 1.2 GB
  • reserved 2 GB
  • context 32,768 × batch 1
DeepSeek-R1-0528-Qwen3-8BOpen weightsDeepSeek8.19B11.4 GB estimated+2.6 GB✓ fits
  • weights 8.2 GB estimated
  • KV cache 2.00 GB (heuristic)
  • overhead 1.2 GB
  • reserved 2 GB
  • context 32,768 × batch 1
Qwen3 8BOpen weightsQwen8.19B11.4 GB estimated+2.6 GB✓ fits
  • weights 8.2 GB estimated
  • KV cache 2.00 GB (heuristic)
  • overhead 1.2 GB
  • reserved 2 GB
  • context 32,768 × batch 1
granite-guardian-3.3-8bOpen weightsIBM8.17B11.4 GB estimated+2.6 GB✓ fits
  • weights 8.2 GB estimated
  • KV cache 2.00 GB (heuristic)
  • overhead 1.2 GB
  • reserved 2 GB
  • context 32,768 × batch 1
granite-3.0-8b-instructOpen weightsIBM8.17B11.4 GB estimated+2.6 GB✓ fits
  • weights 8.2 GB estimated
  • KV cache 2.00 GB (heuristic)
  • overhead 1.2 GB
  • reserved 2 GB
  • context 32,768 × batch 1
stable-diffusion-3.5-largeOpen weightsStability AI8.15B11.4 GB estimated+2.6 GB✓ fits
  • weights 8.2 GB estimated
  • KV cache 2.00 GB (heuristic)
  • overhead 1.2 GB
  • reserved 2 GB
  • context 32,768 × batch 1
ERNIE-Image-TurboOpen weightsBaidu8.03B11.2 GB estimated+2.8 GB✓ fits
  • weights 8.0 GB estimated
  • KV cache 2.00 GB (heuristic)
  • overhead 1.2 GB
  • reserved 2 GB
  • context 32,768 × batch 1
ERNIE-ImageOpen weightsBaidu8.03B11.2 GB estimated+2.8 GB✓ fits
  • weights 8.0 GB estimated
  • KV cache 2.00 GB (heuristic)
  • overhead 1.2 GB
  • reserved 2 GB
  • context 32,768 × batch 1
Meta-Llama-3-8Brestricted-weightsMeta AI8.03B11.2 GB estimated+2.8 GB✓ fits
  • weights 8.0 GB estimated
  • KV cache 2.00 GB (heuristic)
  • overhead 1.2 GB
  • reserved 2 GB
  • context 32,768 × batch 1
NousResearch/Meta-Llama-3-8B-InstructOpen weightsNous Research8.03B11.2 GB estimated+2.8 GB✓ fits
  • weights 8.0 GB estimated
  • KV cache 2.00 GB (heuristic)
  • overhead 1.2 GB
  • reserved 2 GB
  • context 32,768 × batch 1
Hermes-2-Theta-Llama-3-8BOpen weightsNous Research8.03B11.2 GB estimated+2.8 GB✓ fits
  • weights 8.0 GB estimated
  • KV cache 2.00 GB (heuristic)
  • overhead 1.2 GB
  • reserved 2 GB
  • context 32,768 × batch 1
NousResearch/Meta-Llama-3.1-8Brestricted-weightsNous Research8.03B11.2 GB estimated+2.8 GB✓ fits
  • weights 8.0 GB estimated
  • KV cache 2.00 GB (heuristic)
  • overhead 1.2 GB
  • reserved 2 GB
  • context 32,768 × batch 1
Llama 3.1 8Brestricted-weightsMeta AI8.03B11.2 GB estimated+2.8 GB✓ fits
  • weights 8.0 GB estimated
  • KV cache 2.00 GB (heuristic)
  • overhead 1.2 GB
  • reserved 2 GB
  • context 32,768 × batch 1
NousResearch/Meta-Llama-3-8BOpen weightsNous Research8.03B11.2 GB estimated+2.8 GB✓ fits
  • weights 8.0 GB estimated
  • KV cache 2.00 GB (heuristic)
  • overhead 1.2 GB
  • reserved 2 GB
  • context 32,768 × batch 1
Hermes-3-Llama-3.1-8Brestricted-weightsNous Research8.03B11.2 GB estimated+2.8 GB✓ fits
  • weights 8.0 GB estimated
  • KV cache 2.00 GB (heuristic)
  • overhead 1.2 GB
  • reserved 2 GB
  • context 32,768 × batch 1
NousResearch/Meta-Llama-3.1-8B-Instructrestricted-weightsNous Research8.03B11.2 GB estimated+2.8 GB✓ fits
  • weights 8.0 GB estimated
  • KV cache 2.00 GB (heuristic)
  • overhead 1.2 GB
  • reserved 2 GB
  • context 32,768 × batch 1
Llama 3.1 8B Instructrestricted-weightsMeta AI8.03B11.2 GB estimated+2.8 GB✓ fits
  • weights 8.0 GB estimated
  • KV cache 2.00 GB (heuristic)
  • overhead 1.2 GB
  • reserved 2 GB
  • context 32,768 × batch 1
Salesforce/Llama-xLAM-2-8b-fc-rrestricted-weightsSalesforce8.03B11.2 GB estimated+2.8 GB✓ fits
  • weights 8.0 GB estimated
  • KV cache 2.00 GB (heuristic)
  • overhead 1.2 GB
  • reserved 2 GB
  • context 32,768 × batch 1
Meta-Llama-3-8B-Instructrestricted-weightsMeta AI8.03B11.2 GB estimated+2.8 GB✓ fits
  • weights 8.0 GB estimated
  • KV cache 2.00 GB (heuristic)
  • overhead 1.2 GB
  • reserved 2 GB
  • context 32,768 × batch 1
Hy-MT2-7BOpen weightsTencent8.03B11.2 GB estimated+2.8 GB✓ fits
  • weights 8.0 GB estimated
  • KV cache 2.00 GB (heuristic)
  • overhead 1.2 GB
  • reserved 2 GB
  • context 32,768 × batch 1
aya-23-8Brestricted-weightsCohere8.03B11.2 GB estimated+2.8 GB✓ fits
  • weights 8.0 GB estimated
  • KV cache 2.00 GB (heuristic)
  • overhead 1.2 GB
  • reserved 2 GB
  • context 32,768 × batch 1
aya-expanse-8brestricted-weightsCohere8.03B11.2 GB estimated+2.8 GB✓ fits
  • weights 8.0 GB estimated
  • KV cache 2.00 GB (heuristic)
  • overhead 1.2 GB
  • reserved 2 GB
  • context 32,768 × batch 1
c4ai-command-r7b-12-2024restricted-weightsCohere8.03B11.2 GB estimated+2.8 GB✓ fits
  • weights 8.0 GB estimated
  • KV cache 2.00 GB (heuristic)
  • overhead 1.2 GB
  • reserved 2 GB
  • context 32,768 × batch 1
Ministral-8B-Instruct-2410Open weightsMistral AI8.02B11.2 GB estimated+2.8 GB✓ fits
  • weights 8.0 GB estimated
  • KV cache 2.00 GB (heuristic)
  • overhead 1.2 GB
  • reserved 2 GB
  • context 32,768 × batch 1
Robostral NavigateProprietaryMistral AI8B11.2 GB estimated+2.8 GB✓ fits
  • weights 8.0 GB estimated
  • KV cache 2.00 GB (heuristic)
  • overhead 1.2 GB
  • reserved 2 GB
  • context 32,768 × batch 1
gemma-4-E4B-itOpen weightsGoogle8B11.2 GB estimated+2.8 GB✓ fits
  • weights 8.0 GB estimated
  • KV cache 2.00 GB (heuristic)
  • overhead 1.2 GB
  • reserved 2 GB
  • context 32,768 × batch 1
ERNIE-Image-AesOpen weightsBaidu7.94B11.1 GB estimated+2.9 GB✓ fits
  • weights 7.9 GB estimated
  • KV cache 2.00 GB (heuristic)
  • overhead 1.2 GB
  • reserved 2 GB
  • context 32,768 × batch 1
internlm2_5-7b-chatOpen weightsInternLM (Shanghai AI Laboratory)7.74B10.9 GB estimated+3.1 GB✓ fits
  • weights 7.7 GB estimated
  • KV cache 2.00 GB (heuristic)
  • overhead 1.2 GB
  • reserved 2 GB
  • context 32,768 × batch 1
internlm2-chat-7bOpen weightsInternLM (Shanghai AI Laboratory)7.74B10.9 GB estimated+3.1 GB✓ fits
  • weights 7.7 GB estimated
  • KV cache 2.00 GB (heuristic)
  • overhead 1.2 GB
  • reserved 2 GB
  • context 32,768 × batch 1
StripedHyena-Hessian-7BOpen weightsTogether AI7.65B10.8 GB estimated+3.2 GB✓ fits
  • weights 7.7 GB estimated
  • KV cache 2.00 GB (heuristic)
  • overhead 1.1 GB
  • reserved 2 GB
  • context 32,768 × batch 1
StripedHyena-Nous-7BOpen weightsTogether AI7.65B10.8 GB estimated+3.2 GB✓ fits
  • weights 7.7 GB estimated
  • KV cache 2.00 GB (heuristic)
  • overhead 1.1 GB
  • reserved 2 GB
  • context 32,768 × batch 1
SynLogic-7BOpen weightsMiniMax7.62B10.8 GB estimated+3.2 GB✓ fits
  • weights 7.6 GB estimated
  • KV cache 2.00 GB (heuristic)
  • overhead 1.1 GB
  • reserved 2 GB
  • context 32,768 × batch 1
Qwen2.5 7B InstructOpen weightsQwen7.62B10.8 GB estimated+3.2 GB✓ fits
  • weights 7.6 GB estimated
  • KV cache 2.00 GB (heuristic)
  • overhead 1.1 GB
  • reserved 2 GB
  • context 32,768 × batch 1
BFS-Prover-V2-7BOpen weightsByteDance7.62B10.8 GB estimated+3.2 GB✓ fits
  • weights 7.6 GB estimated
  • KV cache 2.00 GB (heuristic)
  • overhead 1.1 GB
  • reserved 2 GB
  • context 32,768 × batch 1
DeepSeek-R1-Distill-Qwen-7BOpen weightsDeepSeek7.62B10.8 GB estimated+3.2 GB✓ fits
  • weights 7.6 GB estimated
  • KV cache 2.00 GB (heuristic)
  • overhead 1.1 GB
  • reserved 2 GB
  • context 32,768 × batch 1
Seed-X-PPO-7BOpen weightsByteDance7.51B10.6 GB estimated+3.4 GB✓ fits
  • weights 7.5 GB estimated
  • KV cache 2.00 GB (heuristic)
  • overhead 1.1 GB
  • reserved 2 GB
  • context 32,768 × batch 1
Hunyuan-7B-InstructOpen weightsTencent7.5B10.6 GB estimated+3.4 GB✓ fits
  • weights 7.5 GB estimated
  • KV cache 2.00 GB (heuristic)
  • overhead 1.1 GB
  • reserved 2 GB
  • context 32,768 × batch 1
OLMo-2-1124-7B-InstructOpen weightsAllen Institute for AI7.3B10.4 GB estimated+3.6 GB✓ fits
  • weights 7.3 GB estimated
  • KV cache 2.00 GB (heuristic)
  • overhead 1.1 GB
  • reserved 2 GB
  • context 32,768 × batch 1
OLMo-2-1124-7BOpen weightsAllen Institute for AI7.3B10.4 GB estimated+3.6 GB✓ fits
  • weights 7.3 GB estimated
  • KV cache 2.00 GB (heuristic)
  • overhead 1.1 GB
  • reserved 2 GB
  • context 32,768 × batch 1
Olmo-3-7B-InstructOpen weightsAllen Institute for AI7.3B10.4 GB estimated+3.6 GB✓ fits
  • weights 7.3 GB estimated
  • KV cache 2.00 GB (heuristic)
  • overhead 1.1 GB
  • reserved 2 GB
  • context 32,768 × batch 1
Olmo-3-7B-ThinkOpen weightsAllen Institute for AI7.3B10.4 GB estimated+3.6 GB✓ fits
  • weights 7.3 GB estimated
  • KV cache 2.00 GB (heuristic)
  • overhead 1.1 GB
  • reserved 2 GB
  • context 32,768 × batch 1
Olmo-3-1025-7BOpen weightsAllen Institute for AI7.3B10.4 GB estimated+3.6 GB✓ fits
  • weights 7.3 GB estimated
  • KV cache 2.00 GB (heuristic)
  • overhead 1.1 GB
  • reserved 2 GB
  • context 32,768 × batch 1
wildguardOpen weightsAllen Institute for AI7.25B10.3 GB estimated+3.7 GB✓ fits
  • weights 7.3 GB estimated
  • KV cache 2.00 GB (heuristic)
  • overhead 1.1 GB
  • reserved 2 GB
  • context 32,768 × batch 1
Mistral-7B-v0.3Open weightsMistral AI7.25B10.3 GB estimated+3.7 GB✓ fits
  • weights 7.3 GB estimated
  • KV cache 2.00 GB (heuristic)
  • overhead 1.1 GB
  • reserved 2 GB
  • context 32,768 × batch 1
Mistral-7B-Instruct-v0.3Open weightsMistral AI7.25B10.3 GB estimated+3.7 GB✓ fits
  • weights 7.3 GB estimated
  • KV cache 2.00 GB (heuristic)
  • overhead 1.1 GB
  • reserved 2 GB
  • context 32,768 × batch 1
Mistral-7B-v0.1Open weightsMistral AI7.24B10.3 GB estimated+3.7 GB✓ fits
  • weights 7.2 GB estimated
  • KV cache 2.00 GB (heuristic)
  • overhead 1.1 GB
  • reserved 2 GB
  • context 32,768 × batch 1
Mistral-7B-Instruct-v0.2Open weightsMistral AI7.24B10.3 GB estimated+3.7 GB✓ fits
  • weights 7.2 GB estimated
  • KV cache 2.00 GB (heuristic)
  • overhead 1.1 GB
  • reserved 2 GB
  • context 32,768 × batch 1
Mistral-7B-Instruct-v0.1Open weightsMistral AI7.24B10.3 GB estimated+3.7 GB✓ fits
  • weights 7.2 GB estimated
  • KV cache 2.00 GB (heuristic)
  • overhead 1.1 GB
  • reserved 2 GB
  • context 32,768 × batch 1
SFR-Embedding-2_Rrestricted-weightsSalesforce7.11B10.2 GB estimated+3.8 GB✓ fits
  • weights 7.1 GB estimated
  • KV cache 2.00 GB (heuristic)
  • overhead 1.1 GB
  • reserved 2 GB
  • context 32,768 × batch 1
SFR-Embedding-Mistralrestricted-weightsSalesforce7.11B10.2 GB estimated+3.8 GB✓ fits
  • weights 7.1 GB estimated
  • KV cache 2.00 GB (heuristic)
  • overhead 1.1 GB
  • reserved 2 GB
  • context 32,768 × batch 1
SeedVR-7BOpen weightsByteDance7B10.1 GB estimated+4.0 GB✓ fits
  • weights 7.0 GB estimated
  • KV cache 2.00 GB (heuristic)
  • overhead 1.1 GB
  • reserved 2 GB
  • context 32,768 × batch 1
internlm-xcomposer-7bOpen weightsInternLM (Shanghai AI Laboratory)7B10.1 GB estimated+4.0 GB✓ fits
  • weights 7.0 GB estimated
  • KV cache 2.00 GB (heuristic)
  • overhead 1.1 GB
  • reserved 2 GB
  • context 32,768 × batch 1
internlm2-base-7bOpen weightsInternLM (Shanghai AI Laboratory)7B10.1 GB estimated+4.0 GB✓ fits
  • weights 7.0 GB estimated
  • KV cache 2.00 GB (heuristic)
  • overhead 1.1 GB
  • reserved 2 GB
  • context 32,768 × batch 1
internlm-xcomposer2-7bOpen weightsInternLM (Shanghai AI Laboratory)7B10.1 GB estimated+4.0 GB✓ fits
  • weights 7.0 GB estimated
  • KV cache 2.00 GB (heuristic)
  • overhead 1.1 GB
  • reserved 2 GB
  • context 32,768 × batch 1
Qwen2-AudioOpen weightsQwen Team7B10.1 GB estimated+4.0 GB✓ fits
  • weights 7.0 GB estimated
  • KV cache 2.00 GB (heuristic)
  • overhead 1.1 GB
  • reserved 2 GB
  • context 32,768 × batch 1
SeedVR2-7BOpen weightsByteDance7B10.1 GB estimated+4.0 GB✓ fits
  • weights 7.0 GB estimated
  • KV cache 2.00 GB (heuristic)
  • overhead 1.1 GB
  • reserved 2 GB
  • context 32,768 × batch 1
fireworks-ai/mistral-7b-eagle-head-experimentalOpen weightsFireworks AI7B10.1 GB estimated+4.0 GB✓ fits
  • weights 7.0 GB estimated
  • KV cache 2.00 GB (heuristic)
  • overhead 1.1 GB
  • reserved 2 GB
  • context 32,768 × batch 1
internlm-chat-7bOpen weightsInternLM (Shanghai AI Laboratory)7B10.1 GB estimated+4.0 GB✓ fits
  • weights 7.0 GB estimated
  • KV cache 2.00 GB (heuristic)
  • overhead 1.1 GB
  • reserved 2 GB
  • context 32,768 × batch 1
internlm2-7bOpen weightsInternLM (Shanghai AI Laboratory)7B10.1 GB estimated+4.0 GB✓ fits
  • weights 7.0 GB estimated
  • KV cache 2.00 GB (heuristic)
  • overhead 1.1 GB
  • reserved 2 GB
  • context 32,768 × batch 1
togethercomputer/LLaMA-2-7B-32Krestricted-weightsTogether AI7B10.1 GB estimated+4.0 GB✓ fits
  • weights 7.0 GB estimated
  • KV cache 2.00 GB (heuristic)
  • overhead 1.1 GB
  • reserved 2 GB
  • context 32,768 × batch 1
RedPajama-INCITE-7B-BaseOpen weightsTogether AI7B10.1 GB estimated+4.0 GB✓ fits
  • weights 7.0 GB estimated
  • KV cache 2.00 GB (heuristic)
  • overhead 1.1 GB
  • reserved 2 GB
  • context 32,768 × batch 1
togethercomputer/Llama-2-7B-32K-Instructrestricted-weightsTogether AI7B10.1 GB estimated+4.0 GB✓ fits
  • weights 7.0 GB estimated
  • KV cache 2.00 GB (heuristic)
  • overhead 1.1 GB
  • reserved 2 GB
  • context 32,768 × batch 1
RedPajama-INCITE-7B-ChatOpen weightsTogether AI7B10.1 GB estimated+4.0 GB✓ fits
  • weights 7.0 GB estimated
  • KV cache 2.00 GB (heuristic)
  • overhead 1.1 GB
  • reserved 2 GB
  • context 32,768 × batch 1
RedPajama-INCITE-7B-InstructOpen weightsTogether AI7B10.1 GB estimated+4.0 GB✓ fits
  • weights 7.0 GB estimated
  • KV cache 2.00 GB (heuristic)
  • overhead 1.1 GB
  • reserved 2 GB
  • context 32,768 × batch 1
Pythia-Chat-Base-7BOpen weightsTogether AI7B10.1 GB estimated+4.0 GB✓ fits
  • weights 7.0 GB estimated
  • KV cache 2.00 GB (heuristic)
  • overhead 1.1 GB
  • reserved 2 GB
  • context 32,768 × batch 1
Qwen2.5-OmniOpen weightsQwen Team7B10.1 GB estimated+4.0 GB✓ fits
  • weights 7.0 GB estimated
  • KV cache 2.00 GB (heuristic)
  • overhead 1.1 GB
  • reserved 2 GB
  • context 32,768 × batch 1
pythia-6.9bOpen weightsEleutherAI6.99B10.0 GB estimated+4.0 GB✓ fits
  • weights 7.0 GB estimated
  • KV cache 2.00 GB (heuristic)
  • overhead 1.1 GB
  • reserved 2 GB
  • context 32,768 × batch 1
OLMoE-1B-7B-0125-InstructOpen weightsAllen Institute for AI6.92B10.0 GB estimated+4.0 GB✓ fits
  • weights 6.9 GB estimated
  • KV cache 2.00 GB (heuristic)
  • overhead 1.0 GB
  • reserved 2 GB
  • context 32,768 × batch 1
OLMoE-1B-7B-0924Open weightsAllen Institute for AI6.92B10.0 GB estimated+4.0 GB✓ fits
  • weights 6.9 GB estimated
  • KV cache 2.00 GB (heuristic)
  • overhead 1.0 GB
  • reserved 2 GB
  • context 32,768 × batch 1
deepseek-coder-7b-instruct-v1.5Open weightsDeepSeek6.91B9.9 GB estimated+4.0 GB✓ fits
  • weights 6.9 GB estimated
  • KV cache 2.00 GB (heuristic)
  • overhead 1.0 GB
  • reserved 2 GB
  • context 32,768 × batch 1
deepseek-coder-6.7b-instructOpen weightsDeepSeek6.74B9.8 GB estimated+4.3 GB✓ fits
  • weights 6.7 GB estimated
  • KV cache 2.00 GB (heuristic)
  • overhead 1.0 GB
  • reserved 2 GB
  • context 32,768 × batch 1
NousResearch/Llama-2-7b-chat-hfOpen weightsNous Research6.74B9.8 GB estimated+4.3 GB✓ fits
  • weights 6.7 GB estimated
  • KV cache 2.00 GB (heuristic)
  • overhead 1.0 GB
  • reserved 2 GB
  • context 32,768 × batch 1
Llama-2-7b-chat-hfrestricted-weightsMeta AI6.74B9.8 GB estimated+4.3 GB✓ fits
  • weights 6.7 GB estimated
  • KV cache 2.00 GB (heuristic)
  • overhead 1.0 GB
  • reserved 2 GB
  • context 32,768 × batch 1
Llama-2-7b-hfrestricted-weightsMeta AI6.74B9.8 GB estimated+4.3 GB✓ fits
  • weights 6.7 GB estimated
  • KV cache 2.00 GB (heuristic)
  • overhead 1.0 GB
  • reserved 2 GB
  • context 32,768 × batch 1
NousResearch/Llama-2-7b-hfOpen weightsNous Research6.74B9.8 GB estimated+4.3 GB✓ fits
  • weights 6.7 GB estimated
  • KV cache 2.00 GB (heuristic)
  • overhead 1.0 GB
  • reserved 2 GB
  • context 32,768 × batch 1
Nous-Hermes-llama-2-7bOpen weightsNous Research6.74B9.8 GB estimated+4.3 GB✓ fits
  • weights 6.7 GB estimated
  • KV cache 2.00 GB (heuristic)
  • overhead 1.0 GB
  • reserved 2 GB
  • context 32,768 × batch 1
evo-1-131k-baseOpen weightsTogether AI6.45B9.4 GB estimated+4.6 GB✓ fits
  • weights 6.5 GB estimated
  • KV cache 2.00 GB (heuristic)
  • overhead 1.0 GB
  • reserved 2 GB
  • context 32,768 × batch 1
evo-1-8k-baseOpen weightsTogether AI6.45B9.4 GB estimated+4.6 GB✓ fits
  • weights 6.5 GB estimated
  • KV cache 2.00 GB (heuristic)
  • overhead 1.0 GB
  • reserved 2 GB
  • context 32,768 × batch 1
chatglm3-6bOpen weightsZ.ai (Zhipu AI)6.24B9.2 GB estimated+4.8 GB✓ fits
  • weights 6.2 GB estimated
  • KV cache 2.00 GB (heuristic)
  • overhead 0.9 GB
  • reserved 2 GB
  • context 32,768 × batch 1
chatglm2-6bOpen weightsZ.ai (Zhipu AI)6B8.9 GB estimated+5.1 GB✓ fits
  • weights 6.0 GB estimated
  • KV cache 2.00 GB (heuristic)
  • overhead 0.9 GB
  • reserved 2 GB
  • context 32,768 × batch 1
GPT-JT-6B-v1Open weightsTogether AI6B8.9 GB estimated+5.1 GB✓ fits
  • weights 6.0 GB estimated
  • KV cache 2.00 GB (heuristic)
  • overhead 0.9 GB
  • reserved 2 GB
  • context 32,768 × batch 1
gpt-j-6bOpen weightsEleutherAI6B8.9 GB estimated+5.1 GB✓ fits
  • weights 6.0 GB estimated
  • KV cache 2.00 GB (heuristic)
  • overhead 0.9 GB
  • reserved 2 GB
  • context 32,768 × batch 1
gemma-4-E2B-itOpen weightsGoogle5.12B7.9 GB estimated+6.1 GB✓ fits
  • weights 5.1 GB estimated
  • KV cache 2.00 GB (heuristic)
  • overhead 0.8 GB
  • reserved 2 GB
  • context 32,768 × batch 1
Molmo2-4BOpen weightsAllen Institute for AI4.85B7.6 GB estimated+6.4 GB✓ fits
  • weights 4.8 GB estimated
  • KV cache 2.00 GB (heuristic)
  • overhead 0.7 GB
  • reserved 2 GB
  • context 32,768 × batch 1
Qianfan-OCROpen weightsBaidu4.74B7.5 GB estimated+6.5 GB✓ fits
  • weights 4.7 GB estimated
  • KV cache 2.00 GB (heuristic)
  • overhead 0.7 GB
  • reserved 2 GB
  • context 32,768 × batch 1
Voxtral-Mini-3B-2507Open weightsMistral AI4.68B7.4 GB estimated+6.6 GB✓ fits
  • weights 4.7 GB estimated
  • KV cache 2.00 GB (heuristic)
  • overhead 0.7 GB
  • reserved 2 GB
  • context 32,768 × batch 1
Qwen3.5-4BOpen weightsQwen4.66B7.4 GB estimated+6.6 GB✓ fits
  • weights 4.7 GB estimated
  • KV cache 2.00 GB (heuristic)
  • overhead 0.7 GB
  • reserved 2 GB
  • context 32,768 × batch 1
Qwen3-VL-4B-InstructOpen weightsQwen4.44B7.1 GB estimated+6.9 GB✓ fits
  • weights 4.4 GB estimated
  • KV cache 2.00 GB (heuristic)
  • overhead 0.7 GB
  • reserved 2 GB
  • context 32,768 × batch 1
Voxtral-Mini-4B-Realtime-2602Open weightsMistral AI4.43B7.1 GB estimated+6.9 GB✓ fits
  • weights 4.4 GB estimated
  • KV cache 2.00 GB (heuristic)
  • overhead 0.7 GB
  • reserved 2 GB
  • context 32,768 × batch 1
Gemma 3 4Brestricted-weightsGoogle4.3B7.0 GB estimated+7.0 GB✓ fits
  • weights 4.3 GB estimated
  • KV cache 2.00 GB (heuristic)
  • overhead 0.7 GB
  • reserved 2 GB
  • context 32,768 × batch 1

First 120 models (fitting first). The full list is in Run locally.

MethodEstimated, not measured: weights = parameters × bytes/param × 1.15 runtime overhead (or the observed artifact file size when one is recorded). /methodology

Relations

In the graph

Explore graph
Manufactured by
Apple
Uses
Apple M5

Timeline

Events

Full timeline

No events yet

Events are generated by the change engine when a material property, price or result changes.

Sources

Where these facts come from

Source documents
SourceDocumentTypeTierLast observedSnapshots
Apple — Mac tech specsapple.com/macbook-air/specs spec_pageT1· Official11 h ago1

Tier 1 = official/primary, 2 = quality secondary, 3 = community, 4 = unverified. Every snapshot is archived; see all sources and the methodology.