ZERO-POINT-INTELLIGENCE-LTD · vision language

MARTHA-2B-Qwen3.5-Omni

ZERO-POINT-INTELLIGENCE-LTD/MARTHA-2B-Qwen3.5-Omni

MARTHA-2B-Qwen3.5-Omni at I1-IQ1_S is exactly 722,962,240 bytes (0.67 GiB / 0.72 GB) — an effective 3.074 bits per weight, not the nominal 1.

From the file· summed from 1 file(s)
Parameters
1.9B
Architecture
qwen35
Context
native (config.json)
License
apache-2.0

Shipped quantizations

exact bytes, summed from published files
QuantSizeExact bytesEffective bpwTensorsPublisher
I1-IQ1_S0.67 GiB722,962,2403.074mradermacher
I1-IQ1_M0.70 GiB748,204,8643.181mradermacher
I1-IQ2_XXS0.74 GiB790,275,9043.360mradermacher
I1-IQ2_XS0.77 GiB824,719,1683.506mradermacher
I1-IQ2_S0.77 GiB832,091,9683.537mradermacher
I1-IQ2_M0.81 GiB865,748,8003.680mradermacher
I1-IQ3_XXS0.86 GiB927,754,0483.944mradermacher
I1-Q2_K_S0.88 GiB944,163,6484.014mradermacher
I1-Q2_K0.90 GiB968,543,0404.117mradermacher
I1-Q3_K_S0.95 GiB1,020,174,1444.337mradermacher
I1-IQ3_XS0.96 GiB1,027,202,8804.367mradermacher
I1-IQ3_S0.98 GiB1,051,090,7524.468mradermacher
I1-IQ3_M0.99 GiB1,059,446,5924.504mradermacher
I1-Q3_K_M1.02 GiB1,099,259,7124.673mradermacher
I1-Q3_K_L1.08 GiB1,164,533,5684.951mradermacher
I1-IQ4_XS1.11 GiB1,195,963,2005.084mradermacher
I1-Q4_01.12 GiB1,204,847,4245.122mradermacher
I1-Q4_K_S1.13 GiB1,212,056,3845.153mradermacher
I1-IQ4_NL1.15 GiB1,231,586,1125.236mradermacher
I1-Q4_K_M1.19 GiB1,274,397,5045.418mradermacher
I1-Q4_11.20 GiB1,288,282,9445.477mradermacher
I1-Q5_K_S1.28 GiB1,374,077,7605.841mradermacher
I1-Q5_K_M1.31 GiB1,411,121,9845.999mradermacher
I1-Q6_K1.45 GiB1,556,391,7446.617mradermacher

Compare with

same modality, comparable size

Will it run on your card?

full quant x context sweep

Why other calculators give a different number

A parameters × bits ÷ 8 estimate puts I1-IQ1_S at roughly 0.99 GiB. The real file is 0.67 GiB, because a quantization is a mixture and some tensors are always kept at higher precision.

Architecture

Architecture unavailable — this repository is gated and no ungated mirror was found. Exact file sizes above are still authoritative; only the KV math needs the config.

Questions people ask

How much VRAM does MARTHA-2B-Qwen3.5-Omni need?
I1-IQ1_S is exactly 722,962,240 bytes (0.67 GiB) in weights. Add the KV cache, which depends on your context length, plus roughly half a gigabyte of runtime overhead.
Which quantization of MARTHA-2B-Qwen3.5-Omni should I use?
Q4_K_M is the usual default. Pick the largest quantization that fits your card at the context you actually need — the table above gives exact sizes for every one published.