FireRedTeam · text

FireRed-Image-Edit-1.1

FireRedTeam/FireRed-Image-Edit-1.1

FireRed-Image-Edit-1.1 at Q4_K_M is exactly 13,065,746,976 bytes (12.17 GiB / 13.07 GB) — an effective 5.116 bits per weight, not the nominal 4.

From the file· summed from 1 file(s)
Parameters
20.4B
Architecture
qwen_image
Context
native (config.json)
License
apache-2.0

Shipped quantizations

exact bytes, summed from published files
QuantSizeExact bytesEffective bpwTensorsPublisher
Q8_07.54 GiB8,098,523,7123.171FireRedTeam
Q8_07.54 GiB8,098,523,7123.171cusiman
Q3_K_S8.52 GiB9,144,597,0563.581vantagewithai
Q3_K_M9.19 GiB9,868,900,9283.864vantagewithai
Q4_011.15 GiB11,975,457,3444.689vantagewithai
Q4_K_S11.41 GiB12,249,135,6804.796vantagewithai
Q4_111.96 GiB12,843,678,2405.029FireRedTeam
Q4_111.96 GiB12,843,678,2405.029cusiman
Q4_112.03 GiB12,912,097,8565.056vantagewithai
Q4_K_M12.17 GiB13,065,746,9765.116cusiman
Q4_K_M12.17 GiB13,065,746,9765.116FireRedTeam
Q4_K_M12.21 GiB13,107,919,4245.133vantagewithai
Q5_K_S13.15 GiB14,117,698,1125.528vantagewithai
Q5_013.40 GiB14,386,657,8565.633vantagewithai
Q5_K_M13.85 GiB14,869,723,7125.823vantagewithai
Q5_114.33 GiB15,391,717,9526.027vantagewithai
Q6_K15.67 GiB16,824,990,2726.588vantagewithai
Q8_020.27 GiB21,761,817,1528.521vantagewithai
BF1638.07 GiB40,872,114,72016.004FireRedTeam

Compare with

same modality, comparable size

Will it run on your card?

full quant x context sweep

Why other calculators give a different number

A parameters × bits ÷ 8 estimate puts Q4_K_M at roughly 10.70 GiB. The real file is 12.17 GiB, because a quantization is a mixture and some tensors are always kept at higher precision.

Architecture

Architecture unavailable — this repository is gated and no ungated mirror was found. Exact file sizes above are still authoritative; only the KV math needs the config.

Questions people ask

How much VRAM does FireRed-Image-Edit-1.1 need?
Q4_K_M is exactly 13,065,746,976 bytes (12.17 GiB) in weights. Add the KV cache, which depends on your context length, plus roughly half a gigabyte of runtime overhead.
Which quantization of FireRed-Image-Edit-1.1 should I use?
Q4_K_M is the usual default. Pick the largest quantization that fits your card at the context you actually need — the table above gives exact sizes for every one published.