Quantization format

IQ3_M

IQ3_M has a nominal rate of bits per weight. We have no measured files carrying this label yet.

From the file· 906 files measured
Nominal bpw
from the block layout
Measured average
3.871
906 files
Range
2.453–13.190
varies by architecture
File sizes
0.03 GiB+
up to 454.54 GiB

Real files

one per model, most downloaded first
ModelSizeEffective bpwvs nominal
Qwen3.6-35B-A3B16.57 GiB3.959
gemma-4-26B-A4B-it12.37 GiB4.003
gemma-4-12B-it5.34 GiB3.836
Qwen3.5-4B2.31 GiB4.263
Hy3133.80 GiB3.847
gemma-4-E4B-it4.39 GiB4.717
gemma-4-31B-it14.09 GiB3.870
Llama-3.2-1B-Instruct0.61 GiB4.255
gemma-4-E2B-it2.92 GiB4.895
Qwen3-8B3.63 GiB3.806
Qwen3.6-27B-Fable-Fusion-711-Uncensored-Heretic-NM-DAU-MTP26.65 GiB8.239
Qwen3.5-0.8B0.47 GiB4.649
Llama-3.1-8B-Instruct3.52 GiB3.771
Ornith-1.0-35B15.74 GiB3.901
Qwopus3.6-27B-Coder12.14 GiB3.753
Llama-3.2-3B-Instruct1.49 GiB3.983
ThinkingCap-Qwen3.6-27B12.95 GiB4.066
Qwythos-9B-v24.53 GiB4.027
Qwen3.5-122B-A10B57.41 GiB3.943
Qwen3-Coder-Next34.13 GiB3.679
Qwen2.5-32B-Instruct13.79 GiB3.616
Qwen3.5-35B-A3B16.57 GiB3.959
Qwen2.5-Coder-7B-Instruct3.33 GiB3.754
Qwen2.5-7B-Instruct3.33 GiB3.754
Qwen3-0.6B0.38 GiB4.288
From the filewhat these mean

Other quantizations