Open-weight local model
OpenAI gpt-oss-20b
Ships in native MXFP4. The estimate follows that checkpoint rather than treating it as generic GGUF Q4.
- Parameters
- 21.512B
- Recommended quantization
- MXFP4
- Weight estimate
- 12.8GB
- Maximum context
- 128K
Quantization memory at the default context
| Quantization | Weights | KV cache | Total VRAM |
|---|---|---|---|
| MXFP4 | 12.80GB | 0.20GB | 14.28GB |
GPUs to check for this model
- Can RTX 4070 12GB run it?
- Can RTX 4070 Ti SUPER 16GB run it?
- Can Apple MacBook Pro M4 Max (40-core GPU) 48GB run it?
- Can RTX 5070 SUPER 18GB run it?
- Can Apple MacBook Pro M4 Pro (16-core GPU) 24GB run it?
- Can Apple MacBook Pro M4 Pro (20-core GPU) 24GB run it?
- Can Apple iMac M4 (10-core GPU) 24GB run it?
- Can Apple Mac mini M4 (10-core GPU) 24GB run it?
- Can Apple MacBook Pro M4 (10-core GPU) 24GB run it?
- Can Apple iMac M3 (10-core GPU) 24GB run it?