Samprix

LLM VRAM Calculator

Pick a local model, a quantization and a context length to see how much GPU memory it needs, and which graphics cards and Macs it fits on.

Specs last checked: how
Model
Quantization
Context length
Memory needed 5.9 GB Qwen3 8B · Q4_K_M · 4,096 tokens
  • Model file 4.8 GB listed as 5.2 GB
  • Context 0.6 GB
  • Allowance 0.5 GB
File size from Ollama library: qwen3; context memory from the publisher’s config, checked Oct 2, 2026.

Memory by context length

+141 MB per 1K tokens

Where it fits at 4K context

Memory available

Graphics cards

Macs

A Mac gives a model about 75% of its unified memory (our estimate). A card that’s short can still run it with part in system RAM, much more slowly. Check your machine
ModelQuantizationFileAt 4K contextAt max contextSmallest card
Qwen3 0.6BQ4_K_M0.523 GB1.4 GB5.4 GB8 GB VRAM
Qwen3 1.7BQ4_K_M1.4 GB2.2 GB6.2 GB8 GB VRAM
Qwen3 4BQ4_K_M2.6 GB3.5 GB8.5 GB8 GB VRAM
Qwen3 8BQ4_K_M5.2 GB5.9 GB11 GB8 GB VRAM
Q8_08.9 GB9.4 GB14 GB12 GB VRAM
Qwen3 14BQ4_K_M9.3 GB9.8 GB15 GB12 GB VRAM
Q8_016 GB16 GB22 GB24 GB VRAM
Qwen3 32BQ4_K_M20 GB20 GB29 GB24 GB VRAM
Q8_035 GB34 GB43 GBLarger than we track
Qwen2.5 0.5BQ4_K_M0.398 GB0.9 GB1.2 GB8 GB VRAM
Qwen2.5 1.5BQ4_K_M0.986 GB1.5 GB2.3 GB8 GB VRAM
Qwen2.5 3BQ4_K_M1.9 GB2.4 GB3.4 GB8 GB VRAM
Qwen2.5 7BQ4_K_M4.7 GB5.1 GB6.6 GB8 GB VRAM
Qwen2.5 14BQ4_K_M9 GB9.6 GB15 GB12 GB VRAM
Qwen2.5 32BQ4_K_M20 GB20 GB27 GB24 GB VRAM
Qwen2.5 72BQ4_K_M47 GB46 GB54 GBLarger than we track
Qwen2.5 Coder 1.5BQ4_K_M0.986 GB1.5 GB2.3 GB8 GB VRAM
Qwen2.5 Coder 3BQ4_K_M1.9 GB2.4 GB3.4 GB8 GB VRAM
Qwen2.5 Coder 7BQ4_K_M4.7 GB5.1 GB6.6 GB8 GB VRAM
Qwen2.5 Coder 14BQ4_K_M9 GB9.6 GB15 GB12 GB VRAM
Qwen2.5 Coder 32BQ4_K_M20 GB20 GB27 GB24 GB VRAM
Phi-4 14BQ4_K_M9.1 GB9.8 GB12 GB12 GB VRAM
Phi-4 mini 3.8BQ4_K_M2.5 GB3.3 GB19 GB8 GB VRAM
Phi-3 mini 3.8BQ4_K_M2.4 GB4.2 GB4.2 GB8 GB VRAM
Phi-3 medium 14BQ4_K_M8.6 GB9.3 GB34 GB12 GB VRAM
Mistral 7BQ4_K_M4.4 GB5.1 GB8.6 GB8 GB VRAM
Mistral Nemo 12BQ4_K_M7.5 GB8.1 GB27 GB12 GB VRAM
Mistral Small 22BQ4_K_M13 GB13 GB20 GB16 GB VRAM
Mistral Small 24BQ4_K_M14 GB14 GB19 GB16 GB VRAM
DeepSeek-R1 Distill 1.5BQ4_K_M1.1 GB1.6 GB5 GB8 GB VRAM
DeepSeek-R1 Distill 7BQ4_K_M4.7 GB5.1 GB12 GB8 GB VRAM
DeepSeek-R1 Distill 8BQ4_K_M4.9 GB5.6 GB21 GB8 GB VRAM
DeepSeek-R1 Distill 14BQ4_K_M9 GB9.6 GB33 GB12 GB VRAM
DeepSeek-R1 Distill 32BQ4_K_M20 GB20 GB51 GB24 GB VRAM
DeepSeek-R1 Distill 70BQ4_K_M43 GB42 GB81 GBLarger than we track
DeepSeek-R1 0528 8BQ4_K_M5.2 GB5.9 GB23 GB8 GB VRAM
SmolLM2 360MQ4_K_M0.271 GB0.9 GB1.1 GB8 GB VRAM
SmolLM2 1.7BQ4_K_M1.1 GB2.3 GB3 GB8 GB VRAM
OLMo 2 7BQ4_K_M4.5 GB6.7 GB6.7 GB8 GB VRAM
OLMo 2 13BQ4_K_M8.4 GB11 GB11 GB12 GB VRAM
TinyLlama 1.1BQ4_K_M0.669 GB1.2 GB1.2 GB8 GB VRAM
Dolphin 3.0 8BQ4_K_M4.9 GB5.6 GB21 GB8 GB VRAM

Did this tool do what you needed?

© 2026 Samprix. Not affiliated with OpenAI, Anthropic or Google.