Qwen3.6-9B-Heretic-Uncensored-Thinking-Sweet-Madness alternatives

Qwen3.6-9B-Heretic-Uncensored-Thinking-Sweet-Madness's smallest published quantization is 16.09 GiB. The models below do the same job with a different trade — less memory, comparable size, or a more permissive licence.

From the file· sizes from summed file bytes

Meaningfully smaller

under 70% of its smallest quantization
ModelParamsSmallestQuantsLicence
Qwen3.6-35B-A3BMoE36.0B8.77 GiB61apache-2.0
Qwen3.6-27B-Fable-Fusion-711-Uncensored-Heretic-NM-DAU-MTP27.8B10.12 GiB34apache-2.0
Qwen3.5-9B9.7B2.97 GiB34apache-2.0
gemma-4-26B-A4B-itMoE26.5B8.99 GiB46apache-2.0
gemma-4-12B-it12.0B3.92 GiB59apache-2.0
Qwen3.5-4B4.7B1.42 GiB57apache-2.0
Qwythos-9B-Claude-Mythos-5-1M9.4B5.38 GiB16apache-2.0
gemma-4-31B-it31.3B7.95 GiB55apache-2.0

Comparable in size

within ±40%, so a like-for-like swap
ModelParamsSmallestQuantsLicence
gemma-4-26B-A4B-it-qat-q4_0-unquantizedMoE26.5B13.45 GiB1apache-2.0
gemma-4-31B-it-qat-q4_0-unquantized32.7B16.44 GiB1apache-2.0
Qwen3.6-40B-Claude-4.6-Opus-Deckard-Heretic-Uncensored-Thinking39.5B13.76 GiB30apache-2.0
Qwen3.6-35B-A3B-Claude-4.7-Opus-Reasoning-DistilledMoE36.0B12.34 GiB12apache-2.0
Qwen3.6-35B-A3B-uncensored-heretic-Native-MTP-PreservedMoE35.1B16.08 GiB10apache-2.0
Qwen3.6-35B-A3B-uncensored-hereticMoE35.1B15.71 GiB10apache-2.0
Qwopus3.6-27B-v227.8B12.39 GiB10apache-2.0
Qwen3.6-27B-uncensored-heretic-v227.4B12.39 GiB11apache-2.0

More permissively licensed

licences that allow commercial use
ModelParamsSmallestQuantsLicence
Qwen3.6-35B-A3BMoE36.0B8.77 GiB61apache-2.0
Qwen3.6-27B-Fable-Fusion-711-Uncensored-Heretic-NM-DAU-MTP27.8B10.12 GiB34apache-2.0
Qwen3.5-9B9.7B2.97 GiB34apache-2.0
gemma-4-26B-A4B-itMoE26.5B8.99 GiB46apache-2.0
gemma-4-12B-it12.0B3.92 GiB59apache-2.0
Qwen3.5-4B4.7B1.42 GiB57apache-2.0
Qwythos-9B-Claude-Mythos-5-1M9.4B5.38 GiB16apache-2.0
gemma-4-31B-it31.3B7.95 GiB55apache-2.0