DavidAU · text

Qwen2.5-MOE-2X1.5B-DeepSeek-Uncensored-Censored-4B

DavidAU/Qwen2.5-MOE-2X1.5B-DeepSeek-Uncensored-Censored-4B

Qwen2.5-MOE-2X1.5B-DeepSeek-Uncensored-Censored-4B at Q4_K_M is exactly 2,517,671,648 bytes (2.34 GiB / 2.52 GB) — an effective 4.925 bits per weight, not the nominal 4.

From the file· summed from 1 file(s)
Parameters
4.1B
Architecture
qwen2moe
Context
native (config.json)
License
apache-2.0

Shipped quantizations

exact bytes, summed from published files
QuantSizeExact bytesEffective bpwTensorsPublisher
Q2_K1.48 GiB1,590,256,3523.111DavidAU
Q3_K_S1.73 GiB1,855,147,2323.629DavidAU
Q3_K_M1.89 GiB2,024,180,9603.960DavidAU
Q3_K_L2.02 GiB2,173,062,3684.251DavidAU
IQ4_XS2.11 GiB2,267,813,6004.437DavidAU
Q4_K_S2.22 GiB2,382,909,1524.662DavidAU
Q4_K_M2.34 GiB2,517,671,6484.925DavidAU
Q5_K_S2.65 GiB2,849,189,6005.574DavidAU
Q5_K_M2.73 GiB2,926,690,0165.726DavidAU
Q6_K3.13 GiB3,361,272,0326.576DavidAU
Q8_04.05 GiB4,351,589,6008.513DavidAU
F167.62 GiB8,185,077,18416.013DavidAU

Compare with

same modality, comparable size

Will it run on your card?

full quant x context sweep

Why other calculators give a different number

A parameters × bits ÷ 8 estimate puts Q4_K_M at roughly 2.14 GiB. The real file is 2.34 GiB, because a quantization is a mixture and some tensors are always kept at higher precision.

Architecture

Architecture unavailable — this repository is gated and no ungated mirror was found. Exact file sizes above are still authoritative; only the KV math needs the config.

Questions people ask

How much VRAM does Qwen2.5-MOE-2X1.5B-DeepSeek-Uncensored-Censored-4B need?
Q4_K_M is exactly 2,517,671,648 bytes (2.34 GiB) in weights. Add the KV cache, which depends on your context length, plus roughly half a gigabyte of runtime overhead.
Which quantization of Qwen2.5-MOE-2X1.5B-DeepSeek-Uncensored-Censored-4B should I use?
Q4_K_M is the usual default. Pick the largest quantization that fits your card at the context you actually need — the table above gives exact sizes for every one published.
Qwen2.5-MOE-2X1.5B-DeepSeek-Uncensored-Censored-4B — VRAM requirements, exact quant sizes — ossmodeldb