Quantization format

UD_IQ2_M

UD_IQ2_M has a nominal rate of bits per weight. We have no measured files carrying this label yet.

From the file· 206 files measured
Nominal bpw
from the block layout
Measured average
2.937
206 files
Range
1.874–5.853
varies by architecture
File sizes
0.10 GiB+
up to 329.27 GiB

Real files

one per model, most downloaded first
ModelSizeEffective bpwvs nominal
Qwen3.6-27B10.10 GiB3.123
Qwen3.6-35B-A3B10.73 GiB2.564
Qwen3.5-9B3.40 GiB3.024
gemma-4-26B-A4B-it9.33 GiB3.018
gemma-4-12B-it3.92 GiB2.818
DeepSeek-V4-Flash84.68 GiB2.500
Qwen3.5-4B1.64 GiB3.022
gemma-4-E4B-it3.30 GiB3.547
Qwen3-Coder-30B-A3B-Instruct10.09 GiB2.840
gemma-4-31B-it10.01 GiB2.751
Qwen3-VL-30B-A3B-Instruct10.10 GiB2.792
Qwen-AgentWorld-35B-A3B10.77 GiB2.669
Llama-3.2-1B-Instruct0.50 GiB3.472
gemma-4-E2B-it2.13 GiB3.577
Qwen3-8B2.90 GiB3.038
Qwen3.5-0.8B0.35 GiB3.407
Qwen3-4B1.43 GiB3.049
GLM-5.2222.19 GiB2.534
Llama-3.1-8B-Instruct2.80 GiB2.992
Ornith-1.0-35B10.77 GiB2.669
Llama-3.2-3B-Instruct1.17 GiB3.128
Kimi-K2.7-Code296.14 GiB2.403
Qwen3.5-122B-A10B36.46 GiB2.504
Qwen3-Coder-Next23.25 GiB2.506
gemma-3-1b-it0.54 GiB4.625
From the filewhat these mean

Other quantizations