King3Djbl · text

mythos-v2-8b

King3Djbl/mythos-v2-8b

mythos-v2-8b at Q4_K_M is exactly 5,027,783,616 bytes (4.68 GiB / 5.03 GB) — an effective 4.911 bits per weight, not the nominal 4.

From the file· summed from 1 file(s)
Parameters
8.2B
Architecture
qwen3
Context
native (config.json)
License
apache-2.0

Shipped quantizations

exact bytes, summed from published files
QuantSizeExact bytesEffective bpwTensorsPublisher
IQ4_XS1.68 GiB1,798,732,9921.757King3Djbl
IQ2_XXS2.32 GiB2,490,111,1682.432King3Djbl
Q2_K3.06 GiB3,281,732,5443.205King3Djbl
IQ3_XXS3.14 GiB3,369,632,9603.291King3Djbl
Q3_K_M3.84 GiB4,124,160,9604.028King3Djbl
Q4_04.45 GiB4,774,749,1204.664King3Djbl
Q4_K_M4.68 GiB5,027,783,6164.911King3Djbl
Q5_K_M5.45 GiB5,851,112,3845.715King3Djbl
Q6_K6.26 GiB6,725,899,2006.569King3Djbl
Q8_08.11 GiB8,709,518,2728.507King3Djbl
F1615.26 GiB16,388,043,71216.006King3Djbl

Compare with

same modality, comparable size

Will it run on your card?

full quant x context sweep

Why other calculators give a different number

A parameters × bits ÷ 8 estimate puts Q4_K_M at roughly 4.29 GiB. The real file is 4.68 GiB, because a quantization is a mixture and some tensors are always kept at higher precision.

Architecture

Architecture unavailable — this repository is gated and no ungated mirror was found. Exact file sizes above are still authoritative; only the KV math needs the config.

Questions people ask

How much VRAM does mythos-v2-8b need?
Q4_K_M is exactly 5,027,783,616 bytes (4.68 GiB) in weights. Add the KV cache, which depends on your context length, plus roughly half a gigabyte of runtime overhead.
Which quantization of mythos-v2-8b should I use?
Q4_K_M is the usual default. Pick the largest quantization that fits your card at the context you actually need — the table above gives exact sizes for every one published.