CanRunAI

オープンウェイトモデル

DeepSeek R1 Distill Qwen 1.5B

Small reasoning checkpoint; long reasoning responses can reduce perceived chat speed.

パラメータ
1.777B
推奨量子化
Q4_K_M
重み推定
1.3GB
最大コンテキスト
128K

公式モデルカード

既定コンテキストでの量子化メモリ

量子化重みKV cache合計VRAM
FP163.70GB0.23GB4.73GB
Q8_02.10GB0.23GB3.13GB
Q6_K1.52GB0.23GB2.55GB
Q5_K_M1.28GB0.23GB2.32GB
Q4_K_M1.30GB0.23GB2.33GB
Q3_K_M1.10GB0.23GB2.13GB

このモデルを確認するGPU