huihui-ai · text

Qwen3-14B-abliterated

huihui-ai/Qwen3-14B-abliterated

Qwen3-14B-abliterated at Q4_K_M is exactly 9,001,749,568 bytes (8.38 GiB / 9.00 GB) — an effective 4.876 bits per weight, not the nominal 4.

From the file· summed from 1 file(s)
Parameters
14.8B
Architecture
qwen3
Context
native (config.json)
License
apache-2.0

Shipped quantizations

exact bytes, summed from published files
QuantSizeExact bytesEffective bpwTensorsPublisher
IQ2_XS4.37 GiB4,691,585,0882.541bartowski
IQ2_S4.62 GiB4,963,308,6082.689bartowski
IQ2_M4.96 GiB5,322,937,4082.883bartowski
Q2_K5.36 GiB5,753,979,9683.117bartowski
IQ3_XXS5.53 GiB5,942,662,2083.219bartowski
IQ3_XS5.94 GiB6,375,297,0883.454bartowski
Q2_K_L6.07 GiB6,513,659,9683.529bartowski
Q3_K_S6.20 GiB6,657,101,8883.606bartowski
IQ3_M6.41 GiB6,883,405,8883.729bartowski
Q3_K_M6.82 GiB7,321,309,2483.966bartowski
Q3_K_L7.36 GiB7,900,647,4884.280bartowski
IQ4_XS7.55 GiB8,110,726,2084.394bartowski
IQ4_NL7.95 GiB8,541,359,1684.627bartowski
Q4_07.96 GiB8,542,997,5684.628bartowski
Q4_K_S7.98 GiB8,573,471,8084.644bartowski
Q4_K_M8.38 GiB9,001,749,5684.876bartowski
Q4_18.74 GiB9,389,517,8885.086bartowski
Q4_K_L8.92 GiB9,579,106,3685.189bartowski
Q5_K_S9.56 GiB10,263,891,0085.560bartowski
Q5_K_M9.79 GiB10,514,566,2085.696bartowski
Q5_K_L10.24 GiB10,994,683,9685.956bartowski
Q6_K11.29 GiB12,121,933,8886.566bartowski
Q6_K_L11.64 GiB12,498,735,1686.771bartowski
Q8_014.62 GiB15,698,530,3688.504bartowski
BF1627.51 GiB29,543,419,68016.004bartowski

Compare with

same modality, comparable size

Will it run on your card?

full quant x context sweep

Why other calculators give a different number

A parameters × bits ÷ 8 estimate puts Q4_K_M at roughly 7.74 GiB. The real file is 8.38 GiB, because a quantization is a mixture and some tensors are always kept at higher precision.

Architecture

Architecture unavailable — this repository is gated and no ungated mirror was found. Exact file sizes above are still authoritative; only the KV math needs the config.

Questions people ask

How much VRAM does Qwen3-14B-abliterated need?
Q4_K_M is exactly 9,001,749,568 bytes (8.38 GiB) in weights. Add the KV cache, which depends on your context length, plus roughly half a gigabyte of runtime overhead.
Which quantization of Qwen3-14B-abliterated should I use?
Q4_K_M is the usual default. Pick the largest quantization that fits your card at the context you actually need — the table above gives exact sizes for every one published.