DavidAU · text

Qwen3-The-Xiaolong-Omega-Directive-22B-uncensored-abliterated

DavidAU/Qwen3-The-Xiaolong-Omega-Directive-22B-uncensored-abliterated

Qwen3-The-Xiaolong-Omega-Directive-22B-uncensored-abliterated at Q4_K_M is exactly 13,333,473,760 bytes (12.42 GiB / 13.33 GB) — an effective 4.841 bits per weight, not the nominal 4.

From the file· summed from 1 file(s)
Parameters
22.0B
Architecture
qwen3
Context
native (config.json)
License
apache-2.0

Shipped quantizations

exact bytes, summed from published files
QuantSizeExact bytesEffective bpwTensorsPublisher
I1-IQ1_S4.75 GiB5,102,210,7201.852mradermacher
I1-IQ1_M5.14 GiB5,521,845,9202.005mradermacher
I1-IQ2_XXS5.79 GiB6,221,237,9202.259mradermacher
I1-IQ2_XS6.36 GiB6,831,541,9202.480mradermacher
I1-IQ2_S6.71 GiB7,202,470,5602.615mradermacher
I1-IQ2_M7.23 GiB7,761,984,1602.818mradermacher
I1-Q2_K_S7.31 GiB7,843,965,6002.848mradermacher
Q2_K7.85 GiB8,424,038,8803.058YOLOLOBOOM
Q2_K7.85 GiB8,424,038,8803.058DavidAU
I1-Q2_K7.85 GiB8,424,041,1203.058mradermacher
I1-IQ3_XXS8.13 GiB8,729,868,9603.169mradermacher
I1-IQ3_XS8.70 GiB9,339,451,0403.391mradermacher
Q3_K_S9.11 GiB9,780,424,1603.551YOLOLOBOOM
Q3_K_S9.11 GiB9,780,424,1603.551DavidAU
I1-Q3_K_S9.11 GiB9,780,426,4003.551mradermacher
I1-IQ3_S9.15 GiB9,823,598,2403.567mradermacher
I1-IQ3_M9.43 GiB10,122,319,5203.675mradermacher
Q3_K_M10.07 GiB10,808,110,5603.924YOLOLOBOOM
Q3_K_M10.07 GiB10,808,110,5603.924DavidAU
I1-Q3_K_M10.07 GiB10,808,112,8003.924mradermacher
Q3_K_L10.90 GiB11,707,919,8404.251YOLOLOBOOM
Q3_K_L10.90 GiB11,707,919,8404.251DavidAU
I1-Q3_K_L10.90 GiB11,707,922,0804.251mradermacher
I1-IQ4_XS11.17 GiB11,990,090,4004.353mradermacher
IQ4_XS11.26 GiB12,087,572,9604.388YOLOLOBOOM
IQ4_XS11.26 GiB12,087,572,9604.388DavidAU
I1-Q4_011.77 GiB12,642,562,7204.590mradermacher
Q4_K_S11.81 GiB12,684,175,8404.605DavidAU
Q4_K_S11.81 GiB12,684,175,8404.605YOLOLOBOOM
I1-Q4_K_S11.81 GiB12,684,178,0804.605mradermacher
Q4_K_M12.42 GiB13,333,473,7604.841DavidAU
Q4_K_M12.42 GiB13,333,473,7604.841YOLOLOBOOM
I1-Q4_K_M12.42 GiB13,333,476,0004.841mradermacher
I1-Q4_112.98 GiB13,932,106,4005.058mradermacher
Q5_K_S14.21 GiB15,260,641,7605.540DavidAU
Q5_K_S14.21 GiB15,260,641,7605.540YOLOLOBOOM
I1-Q5_K_S14.21 GiB15,260,644,0005.540mradermacher
Q5_K_M14.56 GiB15,636,654,5605.677DavidAU
Q5_K_M14.56 GiB15,636,654,5605.677YOLOLOBOOM
I1-Q5_K_M14.56 GiB15,636,656,8005.677mradermacher

Compare with

same modality, comparable size

Will it run on your card?

full quant x context sweep

Why other calculators give a different number

A parameters × bits ÷ 8 estimate puts Q4_K_M at roughly 11.54 GiB. The real file is 12.42 GiB, because a quantization is a mixture and some tensors are always kept at higher precision.

Architecture

Architecture unavailable — this repository is gated and no ungated mirror was found. Exact file sizes above are still authoritative; only the KV math needs the config.

Questions people ask

How much VRAM does Qwen3-The-Xiaolong-Omega-Directive-22B-uncensored-abliterated need?
Q4_K_M is exactly 13,333,473,760 bytes (12.42 GiB) in weights. Add the KV cache, which depends on your context length, plus roughly half a gigabyte of runtime overhead.
Which quantization of Qwen3-The-Xiaolong-Omega-Directive-22B-uncensored-abliterated should I use?
Q4_K_M is the usual default. Pick the largest quantization that fits your card at the context you actually need — the table above gives exact sizes for every one published.