DavidAU · text

MN-26B-Oblivion-Uncensored

DavidAU/MN-26B-Oblivion-Uncensored

MN-26B-Oblivion-Uncensored at Q4_K_M is exactly 16,292,150,432 bytes (15.17 GiB / 16.29 GB) — an effective 5.090 bits per weight, not the nominal 4.

From the file· summed from 1 file(s)
Parameters
25.6B
Architecture
llama
Context
native (config.json)
License
apache-2.0

Shipped quantizations

exact bytes, summed from published files
QuantSizeExact bytesEffective bpwTensorsPublisher
IQ2_M9.15 GiB9,819,651,2323.068DavidAU
IQ3_M11.63 GiB12,485,164,1923.901DavidAU
IQ4_XS13.67 GiB14,673,452,1924.584DavidAU
IQ4_NL14.38 GiB15,438,093,4724.823DavidAU
Q4_K_S14.42 GiB15,483,313,3124.837DavidAU
Q4_K_M15.17 GiB16,292,150,4325.090DavidAU
Q5_K_S17.23 GiB18,496,658,5925.779DavidAU
Q5_K_M17.66 GiB18,966,674,5925.925DavidAU
Q6_K20.31 GiB21,808,356,5126.813DavidAU
Q8_025.93 GiB27,847,335,0728.700DavidAU

Compare with

same modality, comparable size

Will it run on your card?

full quant x context sweep

Why other calculators give a different number

A parameters × bits ÷ 8 estimate puts Q4_K_M at roughly 13.41 GiB. The real file is 15.17 GiB, because a quantization is a mixture and some tensors are always kept at higher precision.

Architecture

Architecture unavailable — this repository is gated and no ungated mirror was found. Exact file sizes above are still authoritative; only the KV math needs the config.

Questions people ask

How much VRAM does MN-26B-Oblivion-Uncensored need?
Q4_K_M is exactly 16,292,150,432 bytes (15.17 GiB) in weights. Add the KV cache, which depends on your context length, plus roughly half a gigabyte of runtime overhead.
Which quantization of MN-26B-Oblivion-Uncensored should I use?
Q4_K_M is the usual default. Pick the largest quantization that fits your card at the context you actually need — the table above gives exact sizes for every one published.