OFFELLIA_Quantis · Brunobkr

OFFELLIA_Quantis Q4_K_M

This file is exactly 76,480,954,528 bytes — 71.23 GiB / 76.48 GB.

From the file· 10 file(s), summed

Get it

71.23 GiB · 10 files
llama.cpp
llama-cli -hf Brunobkr/OFFELLIA_Quantis:Q4_K_M

Downloads and runs in one step, resolving the quantization by name.

Hugging Face CLI
hf download Brunobkr/OFFELLIA_Quantis --include "ΩFFΣLLIα_DeepSeek-Coder-V2-Lite-Instruct-instruct-Q4_K_M.gguf"

This quantization is split across 10 files totalling 71.23 GiB. You need every part — the first is not a smaller version of the model.

Size
71.23 GiB
76.48 GB
Effective bpw
from real bytes ÷ params
Tensors
Header
GGUF metadata
OFFELLIA_Quantis Q4_K_M — exact size and VRAM — ossmodeldb