Quantization format

UD_Q6_K

UD_Q6_K has a nominal rate of bits per weight. We have no measured files carrying this label yet.

From the file· 28 files measured
Nominal bpw
from the block layout
Measured average
6.740
28 files
Range
6.366–8.138
varies by architecture
File sizes
3.09 GiB+
up to 788.42 GiB

Real files

one per model, most downloaded first
ModelSizeEffective bpwvs nominal
Qwen3.6-35B-A3B27.30 GiB6.522
gemma-4-26B-A4B-it21.58 GiB6.984
Qwen-AgentWorld-35B-A3B27.30 GiB6.765
GLM-5.2582.88 GiB6.646
Ornith-1.0-35B27.30 GiB6.765
Qwen3.5-122B-A10B96.97 GiB6.659
Qwen3-Coder-Next61.27 GiB6.606
Laguna-S-2.191.19 GiB6.663
Nemotron-3-Nano-Omni-30B-A3B-Reasoning-BF1631.28 GiB8.138
Step-3.7-Flash152.12 GiB6.489
LFM2.5-8B-A1B6.60 GiB6.697
Ornith-1.0-397B306.89 GiB6.644
Qwen3.5-397B-A17B304.17 GiB6.477
KAT-Coder-V2.5-Dev27.95 GiB6.927
North-Mini-Code-1.023.76 GiB6.696
MiniMax-M2.7175.15 GiB6.578
Nanbeige4.2-3B3.09 GiB6.366
MiniMax-M3330.08 GiB6.639
NVIDIA-Nemotron-3-Super-120B-A12B-BF16106.87 GiB7.427
Mistral-Small-4-119B-260392.60 GiB6.662
Huihui-gemma-4-26B-A4B-it-abliterated21.33 GiB6.903
MiMo-V2.5239.34 GiB6.615
GLM-5.1578.66 GiB6.594
NVIDIA-Nemotron-3-Ultra-550B-A55B-BF16429.09 GiB6.576
MiMo-V2.5-Pro788.42 GiB6.619
From the filewhat these mean

Other quantizations