wangzhang · text

Qwen3.5-122B-A10B-abliterix

wangzhang/Qwen3.5-122B-A10B-abliterix

Qwen3.5-122B-A10B-abliterix at I1-IQ1_S is exactly 24,837,711,552 bytes (23.13 GiB / 24.84 GB) — an effective 1.627 bits per weight, not the nominal 1.

From the file· summed from 1 file(s)
Parameters
122B
Architecture
qwen35moe
Context
native (config.json)
License
apache-2.0

Shipped quantizations

exact bytes, summed from published files
QuantSizeExact bytesEffective bpwTensorsPublisher
I1-IQ1_S23.13 GiB24,837,711,5521.627mradermacher
I1-IQ1_M25.70 GiB27,598,640,8321.808mradermacher
I1-IQ2_XXS29.99 GiB32,200,189,6322.110mradermacher
I1-IQ2_XS33.43 GiB35,890,865,8562.351mradermacher
I1-IQ2_S33.77 GiB36,257,134,2722.375mradermacher
I1-IQ2_M37.20 GiB39,938,373,3122.616mradermacher
I1-Q2_K_S38.88 GiB41,742,411,4562.735mradermacher
I1-Q2_K41.52 GiB44,577,400,5122.920mradermacher
I1-IQ3_XXS43.91 GiB47,148,234,4323.089mradermacher
I1-IQ3_XS46.72 GiB50,166,302,4003.287mradermacher
I1-Q3_K_S49.29 GiB52,921,517,7603.467mradermacher
I1-IQ3_S49.29 GiB52,924,024,5123.467mradermacher
I1-IQ3_M50.09 GiB53,788,853,9523.524mradermacher
I1-Q3_K_M54.58 GiB58,602,063,5523.839mradermacher
I1-Q3_K_L58.85 GiB63,189,321,4084.140mradermacher
I1-IQ4_XS60.76 GiB65,244,882,6244.274mradermacher
I1-Q4_064.57 GiB69,336,000,1924.543mradermacher
I1-Q4_K_S64.86 GiB69,639,956,1604.562mradermacher
I1-Q4_K_M69.11 GiB74,208,241,3444.862mradermacher
I1-Q4_171.35 GiB76,614,624,9605.019mradermacher
I1-Q5_K_S78.41 GiB84,196,419,2645.516mradermacher
I1-Q5_K_M80.98 GiB86,955,517,6325.697mradermacher
I1-Q6_K93.42 GiB100,307,732,1606.572mradermacher
BF166 shards227.54 GiB244,314,011,32816.006random257
BF1615 shards232.24 GiB249,362,982,72016.337random257

Compare with

same modality, comparable size

Will it run on your card?

full quant x context sweep

Why other calculators give a different number

A parameters × bits ÷ 8 estimate puts I1-IQ1_S at roughly 63.97 GiB. The real file is 23.13 GiB, because a quantization is a mixture and some tensors are always kept at higher precision.

Architecture

Architecture unavailable — this repository is gated and no ungated mirror was found. Exact file sizes above are still authoritative; only the KV math needs the config.

Questions people ask

How much VRAM does Qwen3.5-122B-A10B-abliterix need?
I1-IQ1_S is exactly 24,837,711,552 bytes (23.13 GiB) in weights. Add the KV cache, which depends on your context length, plus roughly half a gigabyte of runtime overhead.
Which quantization of Qwen3.5-122B-A10B-abliterix should I use?
Q4_K_M is the usual default. Pick the largest quantization that fits your card at the context you actually need — the table above gives exact sizes for every one published.