sch0tten · text

Qwen3.5-27B-Uncensored

sch0tten/Qwen3.5-27B-Uncensored

Qwen3.5-27B-Uncensored at Q4_K_M is exactly 16,547,399,136 bytes (15.41 GiB / 16.55 GB) — an effective 4.922 bits per weight, not the nominal 4.

From the file· summed from 1 file(s)
Parameters
26.9B
Architecture
qwen35
Context
native (config.json)
License

Shipped quantizations

exact bytes, summed from published files
QuantSizeExact bytesEffective bpwTensorsPublisher
I1-IQ1_S6.66 GiB7,149,823,7122.127mradermacher
I1-IQ1_M7.11 GiB7,631,083,2322.270mradermacher
I1-IQ2_XXS7.85 GiB8,433,182,4322.508mradermacher
I1-IQ2_XS8.47 GiB9,090,590,4322.704mradermacher
I1-IQ2_S8.72 GiB9,362,912,9922.785mradermacher
I1-IQ2_M9.32 GiB10,004,592,3522.976mradermacher
I1-Q2_K_S9.54 GiB10,248,324,8323.048mradermacher
Q2_K9.98 GiB10,711,664,0963.186mradermacher
I1-Q2_K9.98 GiB10,711,664,3523.186mradermacher
I1-IQ3_XXS10.42 GiB11,186,370,2723.327mradermacher
I1-IQ3_XS11.15 GiB11,967,129,3123.559mradermacher
Q3_K_S11.24 GiB12,073,952,7363.591mradermacher
I1-Q3_K_S11.24 GiB12,073,952,9923.591mradermacher
I1-IQ3_S11.57 GiB12,419,327,7123.694mradermacher
I1-IQ3_M11.72 GiB12,580,873,9523.742mradermacher
Q3_K_M12.39 GiB13,301,442,0163.956mradermacher
I1-Q3_K_M12.39 GiB13,301,442,2723.956mradermacher
Q3_K_L13.36 GiB14,344,775,1364.267mradermacher
I1-Q3_K_L13.36 GiB14,344,775,3924.267mradermacher
I1-IQ4_XS14.05 GiB15,082,505,9524.486mradermacher
IQ4_XS14.15 GiB15,193,916,8964.519mradermacher
I1-Q4_014.46 GiB15,521,433,3124.617mradermacher
Q4_K_S14.52 GiB15,586,313,6964.636mradermacher
I1-Q4_K_S14.52 GiB15,586,313,9524.636mradermacher
Q4_K_M15.41 GiB16,547,399,1364.922mradermacher
I1-Q4_K_M15.41 GiB16,547,399,3924.922mradermacher
I1-Q4_115.91 GiB17,078,240,9925.080mradermacher
Q5_K_S17.40 GiB18,679,612,8965.556mradermacher
I1-Q5_K_S17.40 GiB18,679,613,1525.556mradermacher
Q5_K_M17.91 GiB19,231,098,3365.720mradermacher
I1-Q5_K_M17.91 GiB19,231,098,5925.720mradermacher
Q6_K20.57 GiB22,082,528,7366.568mradermacher
I1-Q6_K20.57 GiB22,082,528,9926.568mradermacher
Q8_026.63 GiB28,595,762,6568.506mradermacher

Compare with

same modality, comparable size

Will it run on your card?

full quant x context sweep

Why other calculators give a different number

A parameters × bits ÷ 8 estimate puts Q4_K_M at roughly 14.09 GiB. The real file is 15.41 GiB, because a quantization is a mixture and some tensors are always kept at higher precision.

Architecture

Architecture unavailable — this repository is gated and no ungated mirror was found. Exact file sizes above are still authoritative; only the KV math needs the config.

Questions people ask

How much VRAM does Qwen3.5-27B-Uncensored need?
Q4_K_M is exactly 16,547,399,136 bytes (15.41 GiB) in weights. Add the KV cache, which depends on your context length, plus roughly half a gigabyte of runtime overhead.
Which quantization of Qwen3.5-27B-Uncensored should I use?
Q4_K_M is the usual default. Pick the largest quantization that fits your card at the context you actually need — the table above gives exact sizes for every one published.