Alternatives to DeepSeek-V3.1
DeepSeek-V3.1 alternatives
DeepSeek-V3.1's smallest published quantization is 137.32 GiB. The models below do the same job with a different trade — less memory, comparable size, or a more permissive licence.
From the file· sizes from summed file bytes
Meaningfully smaller
under 70% of its smallest quantization
| Model | Params | Smallest | Quants | Licence |
|---|---|---|---|---|
| Qwen3-Coder-30B-A3B-InstructMoE | 30.5B | 7.46 GiB | 46 | apache-2.0 |
| Qwen3.6-27B | 27.8B | 8.74 GiB | 40 | apache-2.0 |
| Qwen3.8-27B | 27.8B | 8.39 GiB | 22 | apache-2.0 |
| DeepSeek-V4-FlashMoE | 291B | 76.87 GiB | 12 | mit |
| gemma-4-12B-it-qat-q4_0-unquantized | 12.0B | 6.50 GiB | 2 | apache-2.0 |
| gemma-4-E4B-it | 8.0B | 3.30 GiB | 36 | apache-2.0 |
| Qwen3-30B-A3B-Thinking-2507MoE | 30.5B | 7.05 GiB | 51 | apache-2.0 |
| Qwen3-4B | 4.0B | 1.01 GiB | 29 | — |
Comparable in size
within ±40%, so a like-for-like swap
| Model | Params | Smallest | Quants | Licence |
|---|---|---|---|---|
| GLM-5.2MoE | 753B | 169.33 GiB | 20 | mit |
More permissively licensed
licences that allow commercial use
| Model | Params | Smallest | Quants | Licence |
|---|---|---|---|---|
| Qwen3-Coder-30B-A3B-InstructMoE | 30.5B | 7.46 GiB | 46 | apache-2.0 |
| Qwen3.6-27B | 27.8B | 8.74 GiB | 40 | apache-2.0 |
| Qwen3.8-27B | 27.8B | 8.39 GiB | 22 | apache-2.0 |
| DeepSeek-V4-FlashMoE | 291B | 76.87 GiB | 12 | mit |
| gemma-4-12B-it-qat-q4_0-unquantized | 12.0B | 6.50 GiB | 2 | apache-2.0 |
| gemma-4-E4B-it | 8.0B | 3.30 GiB | 36 | apache-2.0 |
| Qwen3-30B-A3B-Thinking-2507MoE | 30.5B | 7.05 GiB | 51 | apache-2.0 |
| Qwen3-8B | 8.2B | 2.12 GiB | 51 | apache-2.0 |