CanRunAI

オープンウェイトモデル

DeepSeek R1 Distill Qwen 32B

Demanding dense reasoning model. On 24GB cards, lower quantization or short context may be necessary.

パラメータ
32.764B
推奨量子化
Q4_K_M
重み推定
20.4GB
最大コンテキスト
128K

公式モデルカード

既定コンテキストでの量子化メモリ

量子化重みKV cache合計VRAM
FP1666.00GB2.15GB74.75GB
Q8_036.00GB2.15GB41.75GB
Q6_K27.95GB2.15GB32.89GB
Q5_K_M23.65GB2.15GB28.16GB
Q4_K_M20.40GB2.15GB24.59GB
Q3_K_M16.90GB2.15GB20.74GB
Q2_K10.75GB2.15GB13.97GB

このモデルを確認するGPU