Quantization publisher

google

google publishes 4 quantizations across 4 models in our index, averaging 5.966 effective bits per weight. Their files differ in size from other publishers' builds of the same nominal quantization on 4 of the pairs we can compare — the same label does not mean the same file.

From the file· summed file bytes
Repositories
4
Quantizations
4
Models covered
4
Avg effective bpw
5.966
across their files

Same model, same quant label, different bytes

largest disagreements first
ModelQuantgooglevsTheirsDifference
gemma-3-12b-itQ4_07.52 GiBunsloth6.43 GiB+16.9%
gemma-3-4b-itQ4_02.94 GiBbartowski2.21 GiB+33.1%
gemma-3-1b-itQ4_00.93 GiBunsloth0.67 GiB+39.0%
gemma-4-12B-it-qat-q4_0-unquantizedQ4_06.50 GiBlmstudio-community6.50 GiB0.0%

A quantization label describes a target, not a recipe. Publishers make different choices about which tensors to keep at higher precision, and some apply an importance matrix while others don't — so two files both honestly labelled the same thing can differ measurably in size and in quality.

Models they publish

ModelQuantizationsSmallest
gemma-4-12B-it-qat-q4_0-unquantized16.50 GiB
gemma-3-1b-it10.93 GiB
gemma-3-4b-it12.94 GiB
gemma-3-12b-it17.52 GiB