CanRunAI

オープンウェイトモデル

DeepSeek R1 Distill Qwen 14B

Reasoning-focused dense model. Long reasoning output can make generation feel slower than chat models.

パラメータ
14.77B
推奨量子化
Q4_K_M
重み推定
9.3GB
最大コンテキスト
128K

公式モデルカード

既定コンテキストでの量子化メモリ

量子化重みKV cache合計VRAM
FP1629.80GB1.61GB34.39GB
Q8_016.40GB1.61GB19.65GB
Q6_K12.60GB1.61GB15.47GB
Q5_K_M10.66GB1.61GB13.34GB
Q4_K_M9.30GB1.61GB11.84GB
Q3_K_M7.70GB1.61GB10.11GB
Q2_K4.85GB1.61GB7.26GB

このモデルを確認するGPU