Quantization format

F32

Files labelled F32 average 31.257 effective bits per weight across 158 real quantizations — not the nominal 32. That is -2% more than the label implies, because a quantization is a mixture: some tensors are always kept at higher precision.

From the file· 158 files measured
Nominal bpw
32
from the block layout
Measured average
31.257
158 files
Range
5.061–32.414
varies by architecture
File sizes
0.08 GiB+
up to 262.84 GiB

Note

Full precision. Rarely distributed for inference.

Real files

one per model, most downloaded first
ModelSizeEffective bpwvs nominal
embeddinggemma-300m1.13 GiB32.172+1%
nemotron-3.5-asr-streaming-0.6b2.38 GiB32.004+0%
parakeet-unified-en-0.6b2.30 GiB32.001+0%
Llama-3.1-8B-Instruct29.92 GiB32.008+0%
parakeet-tdt-0.6b-v32.34 GiB32.003+0%
whisper-medium2.85 GiB32.021+0%
canary-180m-flash0.70 GiB32.007+0%
Qwen2.5-3B-Instruct11.50 GiB32.015+0%
Phi-3.5-mini-instruct14.24 GiB32.002+0%
gemma-2-2b-it9.74 GiB32.019+0%
Mistral-Nemo-Instruct-240745.63 GiB32.005+0%
parakeet-tdt-0.6b-v22.30 GiB32.001+0%
DeepSeek-R1-Distill-Qwen-7B28.38 GiB32.006+0%
nomic-embed-text-v1.50.51 GiB32.043+0%
DeepSeek-R1-Distill-Qwen-14B55.03 GiB32.003+0%
Mistral-Small-24B-Instruct-250187.82 GiB32.003+0%
GigaAM-v30.82 GiB31.767+-1%
Mistral-7B-Instruct-v0.327.00 GiB32.001+0%
umt5-xxl21.17 GiB32.009+0%
VibeThinker-3B11.50 GiB32.015+0%
Qwen3-TTS-12Hz-0.6B-Base28.88 GiB
phi-454.61 GiB32.002+0%
canary-1b-v23.65 GiB32.004+0%
DeepSeek-R1-Distill-Qwen-1.5B6.63 GiB32.027+0%
Qwen2-7B-Instruct28.38 GiB32.006+0%
From the filewhat these mean

Other quantizations