CanRunAI

Open-weight local model

DeepSeek R1 Distill Qwen 14B

Reasoning-focused dense model. Long reasoning output can make generation feel slower than chat models.

Parameters
14.77B
Recommended quantization
Q4_K_M
Weight estimate
9.3GB
Maximum context
128K

Official model card

Quantization memory at the default context

QuantizationWeightsKV cacheTotal VRAM
FP1629.80GB1.61GB34.39GB
Q8_016.40GB1.61GB19.65GB
Q6_K12.60GB1.61GB15.47GB
Q5_K_M10.66GB1.61GB13.34GB
Q4_K_M9.30GB1.61GB11.84GB
Q3_K_M7.70GB1.61GB10.11GB
Q2_K4.85GB1.61GB7.26GB

GPUs to check for this model