huihui-ai · text

Huihui-Kimi-Linear-48B-A3B-Instruct-abliterated

huihui-ai/Huihui-Kimi-Linear-48B-A3B-Instruct-abliterated

Huihui-Kimi-Linear-48B-A3B-Instruct-abliterated at I1-IQ1_S is exactly 10,081,754,784 bytes (9.39 GiB / 10.08 GB) — an effective 1.642 bits per weight, not the nominal 1.

From the file· summed from 1 file(s)
Parameters
49.1B
Architecture
kimi-linear
Context
native (config.json)
License
mit

Shipped quantizations

exact bytes, summed from published files
QuantSizeExact bytesEffective bpwTensorsPublisher
I1-IQ1_S9.39 GiB10,081,754,7841.642mradermacher
I1-IQ1_M10.42 GiB11,189,045,2801.822mradermacher
I1-IQ2_XXS12.14 GiB13,034,529,4402.123mradermacher
I1-IQ2_XS13.52 GiB14,518,879,3922.365mradermacher
I1-IQ2_S13.65 GiB14,656,787,6162.387mradermacher
I1-IQ2_M15.03 GiB16,133,174,9442.627mradermacher
I1-Q2_K_S15.52 GiB16,669,413,1522.715mradermacher
I1-Q2_K16.79 GiB18,028,533,5362.936mradermacher
I1-IQ3_XXS17.69 GiB18,992,831,1363.093mradermacher
I1-IQ3_XS18.78 GiB20,166,372,7683.284mradermacher
I1-Q3_K_S19.86 GiB21,325,598,1123.473mradermacher
I1-IQ3_S19.86 GiB21,325,598,1123.473mradermacher
I1-IQ3_M20.07 GiB21,548,385,6963.509mradermacher
I1-Q3_K_M21.87 GiB23,486,104,9923.825mradermacher
I1-Q3_K_L23.76 GiB25,509,791,1364.154mradermacher
I1-IQ4_XS24.47 GiB26,270,981,1524.278mradermacher
I1-Q4_025.96 GiB27,869,756,9604.539mradermacher
I1-Q4_K_S26.04 GiB27,956,051,4884.553mradermacher
I1-Q4_K_M27.66 GiB29,702,759,9684.837mradermacher
I1-Q4_128.72 GiB30,838,178,3365.022mradermacher
I1-Q5_K_S31.56 GiB33,885,947,4245.519mradermacher
I1-Q5_K_M32.47 GiB34,867,654,1765.678mradermacher
I1-Q6_K37.59 GiB40,364,127,9046.574mradermacher

Compare with

same modality, comparable size

Will it run on your card?

full quant x context sweep

Why other calculators give a different number

A parameters × bits ÷ 8 estimate puts I1-IQ1_S at roughly 25.73 GiB. The real file is 9.39 GiB, because a quantization is a mixture and some tensors are always kept at higher precision.

Architecture

Architecture unavailable — this repository is gated and no ungated mirror was found. Exact file sizes above are still authoritative; only the KV math needs the config.

Questions people ask

How much VRAM does Huihui-Kimi-Linear-48B-A3B-Instruct-abliterated need?
I1-IQ1_S is exactly 10,081,754,784 bytes (9.39 GiB) in weights. Add the KV cache, which depends on your context length, plus roughly half a gigabyte of runtime overhead.
Which quantization of Huihui-Kimi-Linear-48B-A3B-Instruct-abliterated should I use?
Q4_K_M is the usual default. Pick the largest quantization that fits your card at the context you actually need — the table above gives exact sizes for every one published.