Alternatives to GLM-4.6V
GLM-4.6V alternatives
GLM-4.6V's smallest published quantization is 33.46 GiB. The models below do the same job with a different trade — less memory, comparable size, or a more permissive licence.
From the file· sizes from summed file bytes
Meaningfully smaller
under 70% of its smallest quantization
| Model | Params | Smallest | Quants | Licence |
|---|---|---|---|---|
| Qwen3.6-35B-A3BMoE | 36.0B | 8.77 GiB | 61 | apache-2.0 |
| Qwen3.6-27B-Fable-Fusion-711-Uncensored-Heretic-NM-DAU-MTP | 27.8B | 10.12 GiB | 34 | apache-2.0 |
| Qwen3.5-9B | 9.7B | 2.97 GiB | 34 | apache-2.0 |
| gemma-4-26B-A4B-itMoE | 26.5B | 8.99 GiB | 46 | apache-2.0 |
| gemma-4-12B-it | 12.0B | 3.92 GiB | 59 | apache-2.0 |
| Qwen3.5-4B | 4.7B | 1.42 GiB | 57 | apache-2.0 |
| Qwythos-9B-Claude-Mythos-5-1M | 9.4B | 5.38 GiB | 16 | apache-2.0 |
| gemma-4-31B-it | 31.3B | 7.95 GiB | 55 | apache-2.0 |
Comparable in size
within ±40%, so a like-for-like swap
| Model | Params | Smallest | Quants | Licence |
|---|---|---|---|---|
| Qwen3.5-122B-A10BMoE | 125B | 26.92 GiB | 58 | apache-2.0 |
| Qwen3.6-35B-A3B-Uncensored-HauhauCS-Aggressive | 34.7B | 34.37 GiB | 1 | apache-2.0 |
| Step-3.7-Flash | 201B | 40.03 GiB | 49 | apache-2.0 |
| Qwopus3.5-122B-A10B-Kimi-K2.6-destill-healed-abliteratedMoE | 123B | 42.66 GiB | 13 | other |
| Llama-4-Scout-17B-16E-InstructMoE | 109B | 24.51 GiB | 48 | other |
More permissively licensed
licences that allow commercial use
| Model | Params | Smallest | Quants | Licence |
|---|---|---|---|---|
| Qwen3.6-35B-A3BMoE | 36.0B | 8.77 GiB | 61 | apache-2.0 |
| Qwen3.6-27B-Fable-Fusion-711-Uncensored-Heretic-NM-DAU-MTP | 27.8B | 10.12 GiB | 34 | apache-2.0 |
| Qwen3.5-9B | 9.7B | 2.97 GiB | 34 | apache-2.0 |
| gemma-4-26B-A4B-itMoE | 26.5B | 8.99 GiB | 46 | apache-2.0 |
| gemma-4-12B-it | 12.0B | 3.92 GiB | 59 | apache-2.0 |
| Qwen3.5-4B | 4.7B | 1.42 GiB | 57 | apache-2.0 |
| Qwythos-9B-Claude-Mythos-5-1M | 9.4B | 5.38 GiB | 16 | apache-2.0 |
| gemma-4-31B-it | 31.3B | 7.95 GiB | 55 | apache-2.0 |