NightPrince · text

Muslim-6B-V1.0

NightPrince/Muslim-6B-V1.0

Muslim-6B-V1.0 at Q4_K_M is exactly 3,664,678,976 bytes (3.41 GiB / 3.66 GB) — an effective 4.933 bits per weight, not the nominal 4.

From the file· summed from 1 file(s)
Parameters
5.9B
Architecture
qwen3
Context
native (config.json)
License
apache-2.0

Shipped quantizations

exact bytes, summed from published files
QuantSizeExact bytesEffective bpwTensorsPublisher
I1-IQ1_S1.42 GiB1,520,697,9522.047mradermacher
I1-IQ1_M1.52 GiB1,628,340,8322.192mradermacher
I1-IQ2_XXS1.68 GiB1,807,745,6322.433mradermacher
I1-IQ2_XS1.83 GiB1,968,964,1922.650mradermacher
I1-IQ2_S1.92 GiB2,063,766,1122.778mradermacher
I1-IQ2_M2.06 GiB2,207,289,9522.971mradermacher
I1-Q2_K_S2.12 GiB2,271,035,7123.057mradermacher
Q2_K2.26 GiB2,430,103,6163.271mradermacher
I1-Q2_K2.26 GiB2,430,103,8723.271mradermacher
I1-IQ3_XXS2.28 GiB2,443,096,6723.288mradermacher
I1-IQ3_XS2.46 GiB2,646,249,7923.562mradermacher
Q3_K_S2.57 GiB2,756,350,0163.710mradermacher
I1-Q3_K_S2.57 GiB2,756,350,2723.710mradermacher
I1-IQ3_S2.58 GiB2,775,150,9123.735mradermacher
I1-IQ3_M2.67 GiB2,870,198,5923.863mradermacher
Q3_K_M2.83 GiB3,038,953,5364.090mradermacher
I1-Q3_K_M2.83 GiB3,038,953,7924.090mradermacher
Q3_K_L3.06 GiB3,285,532,7364.422mradermacher
I1-Q3_K_L3.06 GiB3,285,532,9924.422mradermacher
I1-IQ4_XS3.10 GiB3,331,981,6324.485mradermacher
IQ4_XS3.12 GiB3,355,328,5764.516mradermacher
I1-Q4_03.25 GiB3,489,513,7924.697mradermacher
I1-IQ4_NL3.26 GiB3,497,869,6324.708mradermacher
Q4_K_S3.26 GiB3,500,163,1364.711mradermacher
I1-Q4_K_S3.26 GiB3,500,163,3924.711mradermacher
Q4_K_M3.41 GiB3,664,678,9764.933mradermacher
I1-Q4_K_M3.41 GiB3,664,679,2324.933mradermacher
I1-Q4_13.56 GiB3,820,798,2725.143mradermacher
Q5_K_S3.88 GiB4,161,421,3765.601mradermacher
I1-Q5_K_S3.88 GiB4,161,421,6325.601mradermacher
Q5_K_M3.96 GiB4,256,469,0565.729mradermacher
I1-Q5_K_M3.96 GiB4,256,469,3125.729mradermacher
Q6_K4.55 GiB4,885,246,0166.575mradermacher
I1-Q6_K4.55 GiB4,885,246,2726.575mradermacher
Q8_05.89 GiB6,324,652,8968.513mradermacher
F1611.08 GiB11,896,550,49616.012mradermacher

Compare with

same modality, comparable size

Will it run on your card?

full quant x context sweep

Why other calculators give a different number

A parameters × bits ÷ 8 estimate puts Q4_K_M at roughly 3.11 GiB. The real file is 3.41 GiB, because a quantization is a mixture and some tensors are always kept at higher precision.

Architecture

Architecture unavailable — this repository is gated and no ungated mirror was found. Exact file sizes above are still authoritative; only the KV math needs the config.

Questions people ask

How much VRAM does Muslim-6B-V1.0 need?
Q4_K_M is exactly 3,664,678,976 bytes (3.41 GiB) in weights. Add the KV cache, which depends on your context length, plus roughly half a gigabyte of runtime overhead.
Which quantization of Muslim-6B-V1.0 should I use?
Q4_K_M is the usual default. Pick the largest quantization that fits your card at the context you actually need — the table above gives exact sizes for every one published.