DeepSeek-R1 Distill Qwen 32B · Q4_K_M
GGUF · Q4_K_M
- Model-size signal
- 20 GB
- Minimum load
- 23 GB
- Planning memory
- 27 GB
- GPU VRAM
- 24 GB
24 GB dedicated VRAM meets the 23 GB minimum-load planning level, but not the 27 GB recommended headroom.
Expect less room for long context or concurrent GPU work.