Qwen · text

Qwen-Image-Layered

Qwen/Qwen-Image-Layered

Qwen-Image-Layered at Q4_K_M is exactly 13,244,770,944 bytes (12.34 GiB / 13.24 GB) — an effective 5.186 bits per weight, not the nominal 4.

From the file· summed from 1 file(s)
Parameters
20.4B
Architecture
qwen_image
Context
native (config.json)
License
apache-2.0

Shipped quantizations

exact bytes, summed from published files
QuantSizeExact bytesEffective bpwTensorsPublisher
Q2_K6.89 GiB7,394,011,7762.895unsloth
Q3_K_S8.59 GiB9,218,927,2323.610unsloth
Q3_K_M9.24 GiB9,920,817,7923.885unsloth
Q3_K_L9.85 GiB10,581,420,6724.143unsloth
Q4_011.04 GiB11,852,786,3044.641unsloth
Q4_K_S11.56 GiB12,410,759,8084.860unsloth
Q4_111.96 GiB12,843,690,6245.029unsloth
Q4_K_M12.34 GiB13,244,770,9445.186unsloth
Q5_K_S13.34 GiB14,325,623,4245.609unsloth
Q5_013.41 GiB14,400,825,9845.639unsloth
Q5_K_M14.00 GiB15,027,513,9845.884unsloth
Q5_114.33 GiB15,391,730,3046.027unsloth
Q6_K15.70 GiB16,852,429,4406.599unsloth
Q8_020.27 GiB21,761,829,5048.521unsloth
BF1638.07 GiB40,872,127,10416.004unsloth
F1638.07 GiB40,872,127,10416.004unsloth

Compare with

same modality, comparable size

Will it run on your card?

full quant x context sweep

Why other calculators give a different number

A parameters × bits ÷ 8 estimate puts Q4_K_M at roughly 10.70 GiB. The real file is 12.34 GiB, because a quantization is a mixture and some tensors are always kept at higher precision.

Architecture

Architecture unavailable — this repository is gated and no ungated mirror was found. Exact file sizes above are still authoritative; only the KV math needs the config.

Questions people ask

How much VRAM does Qwen-Image-Layered need?
Q4_K_M is exactly 13,244,770,944 bytes (12.34 GiB) in weights. Add the KV cache, which depends on your context length, plus roughly half a gigabyte of runtime overhead.
Which quantization of Qwen-Image-Layered should I use?
Q4_K_M is the usual default. Pick the largest quantization that fits your card at the context you actually need — the table above gives exact sizes for every one published.