CanRunAI

开源权重本地模型

Qwen 2.5 Coder 7B Instruct

Popular code-specialized model that is practical on an 8GB GPU at 4-bit with modest context.

参数量
7.616B
推荐量化
Q4_K_M
权重估算
5.1GB
最大上下文
32K

官方模型卡

默认上下文下的量化显存

量化权重KV cache总显存
FP1615.60GB0.47GB17.63GB
Q8_08.70GB0.47GB10.04GB
Q6_K6.50GB0.47GB7.77GB
Q5_K_M5.50GB0.47GB6.77GB
Q4_K_M5.10GB0.47GB6.37GB
Q3_K_M4.30GB0.47GB5.57GB

检查这些 GPU 能否运行该模型