King3Djbl · text

FableForge-1.5B-v2

King3Djbl/FableForge-1.5B-v2

FableForge-1.5B-v2 at Q4_K_M is exactly 986,047,904 bytes (0.92 GiB / 0.99 GB) — an effective 5.110 bits per weight, not the nominal 4.

From the file· summed from 1 file(s)
Parameters
1.5B
Architecture
qwen2
Context
native (config.json)
License
apache-2.0

Shipped quantizations

exact bytes, summed from published files
QuantSizeExact bytesEffective bpwTensorsPublisher
IQ1_S0.41 GiB436,527,2642.262fableforge-ai
IQ2_XXS0.48 GiB511,017,1522.648fableforge-ai
IQ2_XS0.51 GiB550,326,4322.852fableforge-ai
IQ2_S0.53 GiB563,809,4402.922fableforge-ai
IQ3_XXS0.55 GiB586,478,7843.039fableforge-ai
Q2_K0.63 GiB676,304,2883.505fableforge-ai
IQ3_XS0.68 GiB731,698,8483.792fableforge-ai
Q3_K_S0.71 GiB760,944,0323.943fableforge-ai
IQ3_S0.71 GiB762,406,5603.951fableforge-ai
IQ3_M0.72 GiB776,663,7124.025fableforge-ai
Q3_K_M0.77 GiB824,178,0804.271fableforge-ai
Q3_K_L0.82 GiB880,162,2084.561fableforge-ai
IQ4_XS0.83 GiB895,731,3924.642fableforge-ai
Q4_00.87 GiB934,954,4004.845fableforge-ai
IQ4_NL0.87 GiB936,330,9124.852fableforge-ai
Q4_K_S0.88 GiB940,311,9684.873fableforge-ai
Q4_K_M0.92 GiB986,047,9045.110fableforge-ai
Q4_10.95 GiB1,016,841,6325.270fableforge-ai
Q5_K_S1.02 GiB1,098,728,8645.694fableforge-ai
Q5_K_M1.05 GiB1,125,049,7605.830fableforge-ai
Q6_K1.19 GiB1,272,739,2326.596fableforge-ai
Q8_01.53 GiB1,646,572,4488.533fableforge-ai
F162.88 GiB3,093,668,76816.032fableforge-ai

Compare with

same modality, comparable size

Will it run on your card?

full quant x context sweep

Why other calculators give a different number

A parameters × bits ÷ 8 estimate puts Q4_K_M at roughly 0.81 GiB. The real file is 0.92 GiB, because a quantization is a mixture and some tensors are always kept at higher precision.

Architecture

Architecture unavailable — this repository is gated and no ungated mirror was found. Exact file sizes above are still authoritative; only the KV math needs the config.

Questions people ask

How much VRAM does FableForge-1.5B-v2 need?
Q4_K_M is exactly 986,047,904 bytes (0.92 GiB) in weights. Add the KV cache, which depends on your context length, plus roughly half a gigabyte of runtime overhead.
Which quantization of FableForge-1.5B-v2 should I use?
Q4_K_M is the usual default. Pick the largest quantization that fits your card at the context you actually need — the table above gives exact sizes for every one published.