Quantization format

UD_IQ1_M

UD_IQ1_M has a nominal rate of bits per weight. We have no measured files carrying this label yet.

From the file· 178 files measured
Nominal bpw
from the block layout
Measured average
2.311
178 files
Range
1.467–4.479
varies by architecture
File sizes
0.21 GiB+
up to 288.02 GiB

Real files

one per model, most downloaded first
ModelSizeEffective bpwvs nominal
Qwen3.6-35B-A3B9.36 GiB2.236
DeepSeek-V4-Flash80.93 GiB2.389
Qwen3-Coder-30B-A3B-Instruct8.97 GiB2.523
Qwen3-VL-30B-A3B-Instruct9.00 GiB2.489
Llama-3.2-1B-Instruct0.41 GiB2.843
Qwen3-8B2.23 GiB2.341
Qwen3-4B1.06 GiB2.274
GLM-5.2212.80 GiB2.426
Llama-3.1-8B-Instruct2.13 GiB2.284
Ornith-1.0-35B10.29 GiB2.549
Llama-3.2-3B-Instruct0.89 GiB2.392
Kimi-K2.7-Code283.04 GiB2.297
Qwen3.5-122B-A10B31.87 GiB2.189
Qwen3-Coder-Next20.21 GiB2.179
gemma-3-1b-it0.52 GiB4.479
Qwen3-14B3.79 GiB2.202
Qwen3-VL-4B-Instruct1.06 GiB2.061
Qwen3-0.6B0.21 GiB2.350
Qwen3-VL-2B-Instruct0.52 GiB2.113
Qwen2.5-VL-7B-Instruct2.05 GiB2.123
Qwen3-30B-A3B9.00 GiB2.533
Qwen3-1.7B0.52 GiB2.213
Laguna-S-2.133.19 GiB2.425
Qwen3-4B-Instruct-25071.06 GiB2.274
gemma-3-12b-it3.03 GiB2.133
From the filewhat these mean

Other quantizations