Quantization format

IQ1_S

IQ1_S has a nominal rate of bits per weight. We have no measured files carrying this label yet.

From the file· 192 files measured
Nominal bpw
from the block layout
Measured average
2.075
192 files
Range
1.561–7.254
varies by architecture
File sizes
0.02 GiB+
up to 528.03 GiB

Real files

one per model, most downloaded first
ModelSizeEffective bpwvs nominal
Hy359.65 GiB1.715
gemma-4-E2B-it2.16 GiB3.616
Qwen3.5-122B-A10B26.92 GiB1.849
Qwen3-Coder-Next15.44 GiB1.665
jina-embeddings-v5-text-small0.19 GiB2.792
Laguna-S-2.123.15 GiB1.691
Phi-3.5-mini-instruct0.78 GiB1.762
gemma-2-2b-it0.78 GiB2.546
Step-3.7-Flash40.03 GiB1.708
Mistral-7B-Instruct-v0.31.50 GiB1.783
Meta-Llama-3-8B-Instruct1.88 GiB2.013
Llama-3.1-70B-Instruct14.29 GiB1.740
jina-embeddings-v5-text-nano0.09 GiB3.756
Ornith-1.0-397B76.16 GiB1.649
Qwen3.5-397B-A17B82.64 GiB1.760
Qwen2-7B-Instruct1.77 GiB2.000
Mixtral-8x22B-v0.127.61 GiB1.687
DeepSeek-V3-0324124.38 GiB1.561
Mistral-Small-Instruct-24094.50 GiB1.737
Llama-3-8B-Instruct-32k-v0.11.88 GiB2.012
Yi-Coder-1.5B-Chat0.46 GiB2.661
Yi-Coder-9B-Chat1.88 GiB1.825
Mathstral-7B-v0.11.50 GiB1.783
solar-pro-preview-instruct4.46 GiB1.730
llama-3-firefunction-v214.29 GiB1.740
From the filewhat these mean

Other quantizations