Quantization publisher

nvidia

nvidia publishes 1 quantizations across 1 models in our index, averaging 5.712 effective bits per weight.

From the file· summed file bytes
Repositories
1
Quantizations
1
Models covered
1
Avg effective bpw
5.712
across their files

Models they publish

ModelQuantizationsSmallest
NVIDIA-Nemotron-3-Nano-4B-FP812.64 GiB