Can RTX 4090 run DeepSeek-R1 671B?
❌ Not feasible
DeepSeek-R1 671B @ Q4_K_M · 8K context
Needed
413.1 GB
Usable
24.0 GB
VRAM breakdown
| Weights (Q4_K_M) | 411.0 GB |
| KV cache (8K context) | 0.6 GB |
| Runtime overhead | 1.5 GB |
| Total | 413.1 GB |
Computed at 8K context with fp16 KV cache. Longer contexts need more VRAM — use the VRAM calculator for other settings.
Recommended quantization
Does not fit at any quantization
Too big for local? Rent a cloud GPU · Vast.ai →Smaller models this GPU runs well
FAQ
- How much VRAM does DeepSeek-R1 671B need?
- At Q4_K_M with 8K context: 411.0GB weights + 0.6GB KV cache + 1.5GB runtime overhead = 413.1GB total. RTX 4090 offers 24.0GB usable VRAM, so the verdict is: Not feasible.
- What is the best quantization for DeepSeek-R1 671B on RTX 4090?
- None — even Q2_K (268.0GB) exceeds this GPU's 24.0GB usable VRAM. Use a smaller model, multiple GPUs, or a cloud GPU.
Check another combination in the GPU Checker →
Data verified 2026-08-04