Quantization format

Q5_K

Q5_K has a nominal rate of bits per weight. We have no measured files carrying this label yet.

From the file· 108 files measured
Nominal bpw
from the block layout
Measured average
6.374
108 files
Range
4.174–21.264
varies by architecture
File sizes
0.03 GiB+
up to 208.25 GiB

Real files

one per model, most downloaded first
ModelSizeEffective bpwvs nominal
embeddinggemma-300m0.30 GiB8.417
Qwythos-9B-Claude-Mythos-5-1M6.19 GiB5.647
FLUX.2-klein-9B6.35 GiB6.004
whisper-medium0.50 GiB5.647
whisper-large-v31.01 GiB5.604
whisper-large-v3-turbo0.53 GiB5.677
Qwen3-ASR-1.7B1.63 GiB5.942
Qwen3.6-35B-A3B-Claude-4.7-Opus-Reasoning-Distilled23.61 GiB5.640
Wan2.2-Animate-14B12.21 GiB6.074
Mistral-Small-24B-Instruct-250115.61 GiB5.689
Mistral-7B-Instruct-v0.34.78 GiB5.669
Llama-3.1-8B5.34 GiB5.711
WAN2.2-14B-Rapid-AllInOne12.13 GiB6.010
mistral-small-3.1-24b-instruct-2503-hf15.61 GiB5.689
gemma-2-9b-it6.19 GiB5.754
Qwen3.6-27B-DFlash1.14 GiB5.668
Codestral-22B-v0.114.64 GiB5.654
parakeet-ctc-0.6b8.11 GiB
whisper-small0.16 GiB5.798
DeepSeek-Coder-V2-Lite-Instruct11.04 GiB6.036
Gemma-4-31B-StyleTune18.17 GiB4.777
Qwen3.6-35B-A3B-DFlash0.26 GiB5.810
DA3-BASE1.35 GiB
Hermes-3-Llama-3.1-8B5.34 GiB5.711
Qwen2-0.5B-Instruct0.39 GiB6.803
From the filewhat these mean

Other quantizations