NightPrince · text

Muslim-6B-v3

NightPrince/Muslim-6B-v3

Muslim-6B-v3 at Q4_K_M is exactly 3,664,678,944 bytes (3.41 GiB / 3.66 GB) — an effective 4.933 bits per weight, not the nominal 4.

From the file· summed from 1 file(s)
Parameters
5.9B
Architecture
qwen3
Context
native (config.json)
License
apache-2.0

Shipped quantizations

exact bytes, summed from published files
QuantSizeExact bytesEffective bpwTensorsPublisher
I1-IQ1_S1.42 GiB1,520,697,9202.047mradermacher
I1-IQ1_M1.52 GiB1,628,340,8002.192mradermacher
I1-IQ2_XXS1.68 GiB1,807,745,6002.433mradermacher
I1-IQ2_XS1.83 GiB1,968,964,1602.650mradermacher
I1-IQ2_S1.92 GiB2,063,766,0802.778mradermacher
I1-IQ2_M2.06 GiB2,207,289,9202.971mradermacher
I1-Q2_K_S2.12 GiB2,271,035,6803.057mradermacher
Q2_K2.26 GiB2,430,103,5843.271mradermacher
I1-Q2_K2.26 GiB2,430,103,8403.271mradermacher
I1-IQ3_XXS2.28 GiB2,443,096,6403.288mradermacher
I1-IQ3_XS2.46 GiB2,646,249,7603.562mradermacher
Q3_K_S2.57 GiB2,756,349,9843.710mradermacher
I1-Q3_K_S2.57 GiB2,756,350,2403.710mradermacher
I1-IQ3_S2.58 GiB2,775,150,8803.735mradermacher
I1-IQ3_M2.67 GiB2,870,198,5603.863mradermacher
Q3_K_M2.83 GiB3,038,953,5044.090mradermacher
I1-Q3_K_M2.83 GiB3,038,953,7604.090mradermacher
Q3_K_L3.06 GiB3,285,532,7044.422mradermacher
I1-Q3_K_L3.06 GiB3,285,532,9604.422mradermacher
I1-IQ4_XS3.10 GiB3,331,981,6004.485mradermacher
IQ4_XS3.12 GiB3,355,328,5444.516mradermacher
I1-Q4_03.25 GiB3,489,513,7604.697mradermacher
I1-IQ4_NL3.26 GiB3,497,869,6004.708mradermacher
Q4_K_S3.26 GiB3,500,163,1044.711mradermacher
I1-Q4_K_S3.26 GiB3,500,163,3604.711mradermacher
Q4_K_M3.41 GiB3,664,678,9444.933mradermacher
I1-Q4_K_M3.41 GiB3,664,679,2004.933mradermacher
I1-Q4_13.56 GiB3,820,798,2405.143mradermacher
Q5_K_S3.88 GiB4,161,421,3445.601mradermacher
I1-Q5_K_S3.88 GiB4,161,421,6005.601mradermacher
Q5_K_M3.96 GiB4,256,469,0245.729mradermacher
I1-Q5_K_M3.96 GiB4,256,469,2805.729mradermacher
Q6_K4.55 GiB4,885,245,9846.575mradermacher
I1-Q6_K4.55 GiB4,885,246,2406.575mradermacher
Q8_05.89 GiB6,324,652,8648.513mradermacher
F1611.08 GiB11,896,550,46416.012mradermacher

Compare with

same modality, comparable size

Will it run on your card?

full quant x context sweep

Why other calculators give a different number

A parameters × bits ÷ 8 estimate puts Q4_K_M at roughly 3.11 GiB. The real file is 3.41 GiB, because a quantization is a mixture and some tensors are always kept at higher precision.

Architecture

Architecture unavailable — this repository is gated and no ungated mirror was found. Exact file sizes above are still authoritative; only the KV math needs the config.

Questions people ask

How much VRAM does Muslim-6B-v3 need?
Q4_K_M is exactly 3,664,678,944 bytes (3.41 GiB) in weights. Add the KV cache, which depends on your context length, plus roughly half a gigabyte of runtime overhead.
Which quantization of Muslim-6B-v3 should I use?
Q4_K_M is the usual default. Pick the largest quantization that fits your card at the context you actually need — the table above gives exact sizes for every one published.