Quantization format

IQ1_M

IQ1_M has a nominal rate of bits per weight. We have no measured files carrying this label yet.

From the file· 255 files measured
Nominal bpw
from the block layout
Measured average
2.170
255 files
Range
1.660–7.918
varies by architecture
File sizes
0.02 GiB+
up to 220.44 GiB

Real files

one per model, most downloaded first
ModelSizeEffective bpwvs nominal
Qwen3.6-35B-A3B8.77 GiB2.096
Hy366.22 GiB1.904
gemma-4-31B-it9.42 GiB2.587
Qwen3.5-122B-A10B29.61 GiB2.033
Qwen3-Coder-Next16.11 GiB1.737
Qwen3.5-35B-A3B8.77 GiB2.096
jina-embeddings-v5-text-small0.20 GiB2.900
Laguna-S-2.125.75 GiB1.881
Phi-3.5-mini-instruct0.85 GiB1.920
gemma-2-2b-it0.81 GiB2.674
Inkling210.69 GiB1.900
Llama-3.3-70B-Instruct15.60 GiB1.899
DeepSeek-R1-Distill-Llama-70B15.60 GiB1.899
Step-3.7-Flash44.41 GiB1.894
Qwen2.5-72B-Instruct22.11 GiB2.612
Mistral-7B-Instruct-v0.31.64 GiB1.940
Meta-Llama-3-8B-Instruct2.01 GiB2.154
Llama-3.1-70B-Instruct15.60 GiB1.899
jina-embeddings-v5-text-nano0.09 GiB3.842
Ornith-1.0-397B85.09 GiB1.842
Qwen3.5-397B-A17B91.53 GiB1.949
Qwen2-7B-Instruct1.90 GiB2.145
Ornith-1.0-35B-AEON-Ultimate-Uncensored-BF167.67 GiB1.877
Mixtral-8x22B-v0.130.48 GiB1.862
DeepSeek-V3-0324138.66 GiB1.740
From the filewhat these mean

Other quantizations