Theory-of-mind · text

gemma-4-26B-A4B-it-abliterix-v6

Theory-of-mind/gemma-4-26B-A4B-it-abliterix-v6

gemma-4-26B-A4B-it-abliterix-v6 at I1-IQ1_S is exactly 8,290,271,488 bytes (7.72 GiB / 8.29 GB) — an effective 2.628 bits per weight, not the nominal 1.

From the file· summed from 1 file(s)
Parameters
25.2B
Architecture
gemma4
Context
native (config.json)
License
gemma

Shipped quantizations

exact bytes, summed from published files
QuantSizeExact bytesEffective bpwTensorsPublisher
I1-IQ1_S7.72 GiB8,290,271,4882.628mradermacher
I1-IQ1_M8.07 GiB8,668,657,4082.748mradermacher
I1-IQ2_XXS8.66 GiB9,299,300,6082.948mradermacher
I1-IQ2_XS9.14 GiB9,816,430,8483.112mradermacher
I1-IQ2_S9.20 GiB9,873,201,4083.130mradermacher
I1-IQ2_M9.67 GiB10,377,715,9683.290mradermacher
I1-Q2_K9.86 GiB10,582,737,6643.355mradermacher
I1-Q2_K_S9.89 GiB10,624,482,0483.368mradermacher
I1-IQ3_XXS10.55 GiB11,325,694,2083.591mradermacher
I1-IQ3_XS10.84 GiB11,636,068,0963.689mradermacher
I1-Q3_K_S11.38 GiB12,222,409,9843.875mradermacher
I1-IQ3_S11.38 GiB12,222,409,9843.875mradermacher
I1-IQ3_M11.54 GiB12,392,563,9683.929mradermacher
I1-Q3_K_M12.37 GiB13,286,734,0804.213mradermacher
I1-Q3_K_L12.88 GiB13,824,488,7044.383mradermacher
I1-IQ4_XS12.96 GiB13,917,726,4644.412mradermacher
I1-Q4_013.49 GiB14,488,056,5764.593mradermacher
I1-Q4_K_S14.40 GiB15,464,825,6004.903mradermacher
I1-Q4_114.87 GiB15,969,576,7045.063mradermacher
I1-Q4_K_M15.64 GiB16,796,016,3845.325mradermacher
I1-Q5_K_S16.75 GiB17,986,733,8245.703mradermacher
I1-Q5_K_M17.82 GiB19,132,890,8806.066mradermacher
I1-Q6_K21.08 GiB22,638,399,7447.177mradermacher

Compare with

same modality, comparable size

Will it run on your card?

full quant x context sweep

Why other calculators give a different number

A parameters × bits ÷ 8 estimate puts I1-IQ1_S at roughly 13.22 GiB. The real file is 7.72 GiB, because a quantization is a mixture and some tensors are always kept at higher precision.

Architecture

Architecture unavailable — this repository is gated and no ungated mirror was found. Exact file sizes above are still authoritative; only the KV math needs the config.

Questions people ask

How much VRAM does gemma-4-26B-A4B-it-abliterix-v6 need?
I1-IQ1_S is exactly 8,290,271,488 bytes (7.72 GiB) in weights. Add the KV cache, which depends on your context length, plus roughly half a gigabyte of runtime overhead.
Which quantization of gemma-4-26B-A4B-it-abliterix-v6 should I use?
Q4_K_M is the usual default. Pick the largest quantization that fits your card at the context you actually need — the table above gives exact sizes for every one published.