Quantization format

UD_IQ1_S

UD_IQ1_S has a nominal rate of bits per weight. We have no measured files carrying this label yet.

From the file· 157 files measured
Nominal bpw
from the block layout
Measured average
2.160
157 files
Range
1.387–4.457
varies by architecture
File sizes
0.20 GiB+
up to 265.74 GiB

Real files

one per model, most downloaded first
ModelSizeEffective bpwvs nominal
DeepSeek-V4-Flash76.87 GiB2.270
Qwen3-Coder-30B-A3B-Instruct8.30 GiB2.336
Qwen3-VL-30B-A3B-Instruct8.46 GiB2.340
Llama-3.2-1B-Instruct0.39 GiB2.729
Qwen3-8B2.12 GiB2.222
Qwen3-4B1.01 GiB2.154
GLM-5.2201.15 GiB2.294
Llama-3.1-8B-Instruct2.02 GiB2.156
Ornith-1.0-35B9.80 GiB2.429
Llama-3.2-3B-Instruct0.85 GiB2.272
Qwen3-Coder-Next20.03 GiB2.160
gemma-3-1b-it0.52 GiB4.457
Qwen3-14B3.56 GiB2.073
Qwen3-VL-4B-Instruct1.01 GiB1.953
Qwen3-0.6B0.20 GiB2.285
Qwen3-VL-2B-Instruct0.50 GiB2.022
Qwen2.5-VL-7B-Instruct1.93 GiB2.001
Qwen3-30B-A3B8.42 GiB2.369
Qwen3-1.7B0.50 GiB2.118
Laguna-S-2.131.45 GiB2.298
Qwen3-4B-Instruct-25071.01 GiB2.154
gemma-3-12b-it2.85 GiB2.007
GLM-4.7-Flash8.61 GiB2.370
Qwen3-30B-A3B-Instruct-25078.42 GiB2.370
DeepSeek-R1-0528-Qwen3-8B2.11 GiB2.216
From the filewhat these mean

Other quantizations