Quantization publisher

city96

city96 publishes 241 quantizations across 19 models in our index, averaging 6.735 effective bits per weight. Their files differ in size from other publishers' builds of the same nominal quantization on 25 of the pairs we can compare — the same label does not mean the same file.

From the file· summed file bytes
Repositories
19
Quantizations
241
Models covered
19
Avg effective bpw
6.735
across their files

Same model, same quant label, different bytes

largest disagreements first
ModelQuantcity96vsTheirsDifference
FLUX.2-devQ8_032.60 GiBgguf-org51.67 GiB-36.9%
FLUX.2-devQ4_017.97 GiBgguf-org29.28 GiB-38.6%
HunyuanVideoQ8_013.01 GiBcalcuis20.97 GiB-37.9%
FLUX.2-devQ2_K11.98 GiBgguf-org18.79 GiB-36.3%
stable-diffusion-3.5-largeQ5_15.84 GiBsecond-state9.57 GiB-39.0%
stable-diffusion-3.5-largeQ5_05.38 GiBsecond-state8.82 GiB-39.1%
stable-diffusion-3.5-largeQ4_14.91 GiBsecond-state8.07 GiB-39.2%
stable-diffusion-3.5-largeQ4_04.44 GiBsecond-state7.32 GiB-39.3%
t5-v1_1-xxlF168.87 GiBchatpig10.40 GiB-14.7%
FLUX.2-devQ5_K_S21.63 GiBunsloth22.12 GiB-2.2%
FLUX.2-devQ6_K25.51 GiBgguf-org25.98 GiB-1.8%
FLUX.2-devQ5_123.46 GiBgguf-org23.92 GiB-1.9%
FLUX.2-devQ4_119.80 GiBgguf-org20.27 GiB-2.3%
FLUX.2-devQ5_K_S21.63 GiBgguf-org22.10 GiB-2.1%
FLUX.2-devQ4_K_S17.97 GiBgguf-org18.44 GiB-2.5%
FLUX.2-devQ5_021.63 GiBgguf-org22.10 GiB-2.1%
FLUX.2-devQ4_K_S17.97 GiBunsloth18.43 GiB-2.5%
Qwen-ImageQ5_K_S13.15 GiBunsloth13.32 GiB-1.3%
Qwen-ImageQ2_K6.58 GiBunsloth6.72 GiB-2.2%
FLUX.2-devQ3_K_S14.70 GiBgguf-org14.55 GiB+1.0%
FLUX.2-devQ5_K_M22.41 GiBunsloth22.28 GiB+0.5%
FLUX.2-devQ3_K_M14.86 GiBunsloth14.74 GiB+0.8%
FLUX.2-devQ3_K_S14.70 GiBunsloth14.57 GiB+0.8%
Qwen-ImageQ4_K_S11.31 GiBunsloth11.43 GiB-1.0%
FLUX.2-devQ4_K_M18.70 GiBunsloth18.59 GiB+0.6%

A quantization label describes a target, not a recipe. Publishers make different choices about which tensors to keep at higher precision, and some apply an importance matrix while others don't — so two files both honestly labelled the same thing can differ measurably in size and in quality.

Models they publish

ModelQuantizationsSmallest
FLUX.1-dev113.76 GiB
FLUX.1-schnell113.73 GiB
FLUX.2-dev1411.98 GiB
umt5-xxl102.66 GiB
Qwen-Image146.58 GiB
t5-v1_1-xxl111.96 GiB
Wan2.1-I2V-14B-480P147.38 GiB
stable-diffusion-3.5-large64.44 GiB
LTX-Video150.92 GiB
Wan2.1-T2V-14B146.51 GiB
stable-diffusion-3.5-medium131.35 GiB
stable-diffusion-3-medium111.20 GiB
Wan2.1-I2V-14B-720P147.38 GiB
HiDream-I1-Full146.11 GiB
Wan2.1-FLF2V-14B-720P147.38 GiB
HunyuanVideo135.67 GiB
llava-llama-3-8b-v1_1-transformers153.41 GiB
HunyuanVideo-I2V135.67 GiB
AuraFlow-v0.3142.23 GiB