Kotokin · text

Merged-RP-Stew-V2-51B

Kotokin/Merged-RP-Stew-V2-51B

Merged-RP-Stew-V2-51B at I1-IQ1_S is exactly 11,024,504,672 bytes (10.27 GiB / 11.02 GB) — an effective 1.725 bits per weight, not the nominal 1.

From the file· summed from 1 file(s)
Parameters
51.1B
Architecture
llama
Context
native (config.json)
License
other

Shipped quantizations

exact bytes, summed from published files
QuantSizeExact bytesEffective bpwTensorsPublisher
I1-IQ1_S10.27 GiB11,024,504,6721.725mradermacher
I1-IQ1_M11.21 GiB12,039,493,4721.884mradermacher
I1-IQ2_XXS12.79 GiB13,731,141,4722.149mradermacher
I1-IQ2_XS14.18 GiB15,228,966,7522.383mradermacher
I1-IQ2_S14.98 GiB16,089,843,5522.518mradermacher
I1-IQ2_M16.25 GiB17,443,161,9522.729mradermacher
I1-Q2_K17.67 GiB18,973,673,3122.969mradermacher
I1-IQ3_XXS18.39 GiB19,743,803,2323.090mradermacher
I1-IQ3_XS19.62 GiB21,070,886,7523.297mradermacher
I1-Q3_K_S20.63 GiB22,152,968,0323.466mradermacher
I1-IQ3_S20.71 GiB22,240,704,3523.480mradermacher
I1-IQ3_M21.48 GiB23,069,325,1523.610mradermacher
I1-Q3_K_M23.01 GiB24,703,170,4003.866mradermacher
I1-Q3_K_L25.07 GiB26,921,695,0724.213mradermacher
I1-IQ4_XS25.52 GiB27,401,807,7124.288mradermacher
I1-Q4_026.99 GiB28,982,781,7924.535mradermacher
I1-Q4_K_S27.09 GiB29,087,377,2484.552mradermacher
I1-Q4_K_M28.56 GiB30,670,128,9924.799mradermacher
I1-Q5_K_S32.80 GiB35,214,927,7125.510mradermacher
I1-Q5_K_M33.65 GiB36,136,159,0725.655mradermacher
I1-Q6_K39.06 GiB41,943,816,0326.563mradermacher

Compare with

same modality, comparable size

Will it run on your card?

full quant x context sweep

Why other calculators give a different number

A parameters × bits ÷ 8 estimate puts I1-IQ1_S at roughly 26.78 GiB. The real file is 10.27 GiB, because a quantization is a mixture and some tensors are always kept at higher precision.

Architecture

Architecture unavailable — this repository is gated and no ungated mirror was found. Exact file sizes above are still authoritative; only the KV math needs the config.

Questions people ask

How much VRAM does Merged-RP-Stew-V2-51B need?
I1-IQ1_S is exactly 11,024,504,672 bytes (10.27 GiB) in weights. Add the KV cache, which depends on your context length, plus roughly half a gigabyte of runtime overhead.
Which quantization of Merged-RP-Stew-V2-51B should I use?
Q4_K_M is the usual default. Pick the largest quantization that fits your card at the context you actually need — the table above gives exact sizes for every one published.