Quantization publisher

servantofares

servantofares publishes 12 quantizations across 1 models in our index, averaging 5.242 effective bits per weight. Their files differ in size from other publishers' builds of the same nominal quantization on 2 of the pairs we can compare — the same label does not mean the same file.

From the file· summed file bytes
Repositories
1
Quantizations
12
Models covered
1
Avg effective bpw
5.242
across their files

Same model, same quant label, different bytes

largest disagreements first
ModelQuantservantofaresvsTheirsDifference
GLM-4.7-Flash-hereticQ5_121.31 GiBDavidAU29.75 GiB-28.4%
GLM-4.7-Flash-hereticQ4_K_M17.24 GiBDavidAU24.06 GiB-28.4%

A quantization label describes a target, not a recipe. Publishers make different choices about which tensors to keep at higher precision, and some apply an importance matrix while others don't — so two files both honestly labelled the same thing can differ measurably in size and in quality.

Models they publish

ModelQuantizationsSmallest
GLM-4.7-Flash-heretic129.61 GiB