Nemotron-3-Super-49B-v1 Q4_K_M

This file is exactly 86,051,080,384 bytes — 80.14 GiB / 86.05 GB at an effective 5.705 bits per weight. The nominal rate for Q4_K_M is lower; mixed-precision tensors make the real figure higher, always.

From the file· 1 file(s), summed

Get it

80.14 GiB · 1 file
llama.cpp
llama-cli -hf timteh673/Nemotron-3-Super-120B-A12B-Uncensored-GGUF:Q4_K_M

Downloads and runs in one step, resolving the quantization by name.

Hugging Face CLI
hf download timteh673/Nemotron-3-Super-120B-A12B-Uncensored-GGUF Nemotron-3-Super-120B-A12B-Uncensored-Q4_K_M.gguf
Direct download

Straight from the Hugging Face CDN — we host nothing and earn nothing from this. Verify what you received against the exact byte count above; a size mismatch is the usual cause of a file that will not load.

Size
80.14 GiB
86.05 GB
Effective bpw
5.705
from real bytes ÷ params
Tensors
Header
GGUF metadata