DavidAU · text

Qwen3.5-13B-Strict-Instruct

DavidAU/Qwen3.5-13B-Strict-Instruct

Qwen3.5-13B-Strict-Instruct at I1-IQ1_S is exactly 3,148,430,880 bytes (2.93 GiB / 3.15 GB) — an effective 2.029 bits per weight, not the nominal 1.

From the file· summed from 1 file(s)
Parameters
12.4B
Architecture
qwen35
Context
native (config.json)
License
apache-2.0

Shipped quantizations

exact bytes, summed from published files
QuantSizeExact bytesEffective bpwTensorsPublisher
I1-IQ1_S2.93 GiB3,148,430,8802.029mradermacher
I1-IQ1_M3.15 GiB3,378,683,4242.177mradermacher
I1-IQ2_XXS3.50 GiB3,762,437,6642.425mradermacher
I1-IQ2_XS3.80 GiB4,075,732,5122.627mradermacher
I1-IQ2_S3.95 GiB4,238,016,0322.731mradermacher
I1-IQ2_M4.23 GiB4,545,019,4242.929mradermacher
I1-Q2_K_S4.35 GiB4,673,216,0323.012mradermacher
I1-Q2_K4.53 GiB4,868,251,1683.137mradermacher
I1-IQ3_XXS4.77 GiB5,116,558,8803.297mradermacher
I1-IQ3_XS5.18 GiB5,563,514,4003.585mradermacher
I1-Q3_K_S5.35 GiB5,747,932,7043.704mradermacher
I1-IQ3_S5.36 GiB5,754,617,3763.709mradermacher
I1-IQ3_M5.57 GiB5,981,896,2243.855mradermacher
I1-Q3_K_M5.85 GiB6,282,575,3924.049mradermacher
I1-Q3_K_L6.14 GiB6,596,099,6164.251mradermacher
I1-IQ4_XS6.44 GiB6,913,089,0564.455mradermacher
I1-IQ4_NL6.77 GiB7,267,606,0484.684mradermacher
I1-Q4_06.78 GiB7,280,188,9604.692mradermacher
I1-Q4_K_S6.80 GiB7,301,160,4804.705mradermacher
I1-Q4_K_M7.20 GiB7,732,387,3604.983mradermacher
I1-Q4_17.43 GiB7,973,494,3045.139mradermacher
I1-Q5_K_S8.09 GiB8,685,674,0165.598mradermacher
I1-Q5_K_M8.39 GiB9,010,994,7205.807mradermacher
I1-Q6_K9.50 GiB10,199,055,9046.573mradermacher

Compare with

same modality, comparable size

Will it run on your card?

full quant x context sweep

Why other calculators give a different number

A parameters × bits ÷ 8 estimate puts I1-IQ1_S at roughly 6.50 GiB. The real file is 2.93 GiB, because a quantization is a mixture and some tensors are always kept at higher precision.

Architecture

Architecture unavailable — this repository is gated and no ungated mirror was found. Exact file sizes above are still authoritative; only the KV math needs the config.

Questions people ask

How much VRAM does Qwen3.5-13B-Strict-Instruct need?
I1-IQ1_S is exactly 3,148,430,880 bytes (2.93 GiB) in weights. Add the KV cache, which depends on your context length, plus roughly half a gigabyte of runtime overhead.
Which quantization of Qwen3.5-13B-Strict-Instruct should I use?
Q4_K_M is the usual default. Pick the largest quantization that fits your card at the context you actually need — the table above gives exact sizes for every one published.