Quantization publisher
lmstudio-community
lmstudio-community publishes 889 quantizations across 225 models in our index, averaging 6.199 effective bits per weight. Their files differ in size from other publishers' builds of the same nominal quantization on 25 of the pairs we can compare — the same label does not mean the same file.
From the file· summed file bytes
Repositories
225
Quantizations
889
Models covered
225
Avg effective bpw
6.199
across their files
Same model, same quant label, different bytes
largest disagreements first
| Model | Quant | lmstudio-community | vs | Theirs | Difference |
|---|---|---|---|---|---|
| Laguna-S-2.1 | Q4_K_M | 66.28 GiB | poolside | 89.44 GiB | -25.9% |
| MiniMax-M2.7 | Q6_K | 174.87 GiB | bartowski | 183.52 GiB | -4.7% |
| Mistral-Small-4-119B-2603 | Q6_K | 90.96 GiB | bartowski | 95.75 GiB | -5.0% |
| Laguna-S-2.1 | Q6_K | 89.93 GiB | bartowski | 94.70 GiB | -5.0% |
| Qwen2.5-72B-Instruct | Q6_K | 59.93 GiB | Qwen | 55.75 GiB | +7.5% |
| DeepSeek-R1-0528 | Q4_K_M | 381.12 GiB | unsloth | 377.13 GiB | +1.1% |
| Laguna-S-2.1 | Q8_0 | 116.44 GiB | poolside | 119.91 GiB | -2.9% |
| Qwen2.5-72B-Instruct | Q4_K_M | 44.16 GiB | Qwen | 40.99 GiB | +7.7% |
| Qwen3.6-35B-A3B | Q8_0 | 34.37 GiB | bartowski | 37.07 GiB | -7.3% |
| Qwen3.8-27B | Q8_0 | 27.05 GiB | ggml-org | 29.58 GiB | -8.5% |
| Qwen3.5-35B-A3B | Q6_K | 26.56 GiB | bartowski | 28.82 GiB | -7.9% |
| Qwen3.6-35B-A3B | Q6_K | 26.56 GiB | bartowski | 28.82 GiB | -7.9% |
| Llama-4-Scout-17B-16E-Instruct | Q4_K_M | 62.91 GiB | unsloth | 60.87 GiB | +3.3% |
| Qwen3.8-27B | Q4_K_M | 15.66 GiB | ggml-org | 17.67 GiB | -11.4% |
| NVIDIA-Nemotron-3-Nano-30B-A3B-BF16 | Q4_K_M | 22.83 GiB | ggml-org | 20.88 GiB | +9.3% |
| Laguna-XS-2.1 | Q6_K | 25.61 GiB | bartowski | 27.04 GiB | -5.3% |
| gemma-4-31B-it | Q6_K | 23.47 GiB | bartowski | 24.89 GiB | -5.7% |
| DeepSeek-R1-0528 | Q6_K | 514.51 GiB | unsloth | 513.20 GiB | +0.3% |
| Qwen3-30B-A3B-Thinking-2507 | Q3_K_L | 14.81 GiB | bartowski | 13.58 GiB | +9.0% |
| Qwen3-30B-A3B-Instruct-2507 | Q3_K_L | 14.81 GiB | bartowski | 13.58 GiB | +9.0% |
| Llama-4-Scout-17B-16E-Instruct | Q8_0 | 105.61 GiB | unsloth | 106.67 GiB | -1.0% |
| Qwen3.6-35B-A3B | Q4_K_M | 19.71 GiB | bartowski | 20.75 GiB | -5.0% |
| Qwen3.5-35B-A3B | Q4_K_M | 19.72 GiB | bartowski | 20.75 GiB | -5.0% |
| gemma-4-26B-A4B-it | Q8_0 | 25.02 GiB | ggml-org | 25.89 GiB | -3.4% |
| NVIDIA-Nemotron-3-Super-120B-A12B-BF16 | Q4_K_M | 80.14 GiB | bartowski | 81.01 GiB | -1.1% |
A quantization label describes a target, not a recipe. Publishers make different choices about which tensors to keep at higher precision, and some apply an importance matrix while others don't — so two files both honestly labelled the same thing can differ measurably in size and in quality.
Models they publish
| Model | Quantizations | Smallest |
|---|---|---|
| Qwen3-Coder-30B-A3B-Instruct | 4 | 13.58 GiB |
| Qwen3.6-27B | 3 | 15.41 GiB |
| Qwen3.6-35B-A3B | 3 | 19.71 GiB |
| Qwen3.8-27B | 3 | 15.66 GiB |
| Qwen3.5-9B | 3 | 5.24 GiB |
| gemma-4-26B-A4B-it | 3 | 15.64 GiB |
| gemma-4-12B-it | 3 | 6.87 GiB |
| Qwen3.5-4B | 3 | 2.52 GiB |
| gemma-4-12B-it-qat-q4_0-unquantized | 1 | 6.50 GiB |
| gemma-4-E4B-it | 3 | 4.97 GiB |
| Qwen3-30B-A3B-Thinking-2507 | 4 | 14.81 GiB |
| gemma-4-31B-it | 3 | 17.40 GiB |
| Qwen3-4B | 4 | 2.09 GiB |
| Qwen3-8B | 4 | 4.13 GiB |
| gemma-4-26B-A4B-it-qat-q4_0-unquantized | 1 | 13.45 GiB |
| Laguna-XS-2.1 | 3 | 18.88 GiB |
| DeepSeek-V4-Flash-0731 | 1 | 145.64 GiB |
| Qwen3.5-0.8B | 3 | 0.49 GiB |
| Llama-3.2-1B-Instruct | 4 | 0.68 GiB |
| gemma-4-E2B-it | 3 | 3.19 GiB |
| gemma-4-31B-it-qat-q4_0-unquantized | 1 | 16.44 GiB |
| gpt-oss-20b | 1 | 11.28 GiB |
| gemma-4-E4B-it-qat-q4_0-unquantized | 1 | 4.80 GiB |
| Qwen3-VL-30B-A3B-Instruct | 3 | 17.28 GiB |
| Laguna-S-2.1 | 3 | 66.28 GiB |
| Qwen3.5-35B-A3B | 3 | 19.72 GiB |
| Llama-3.1-8B-Instruct | 6 | 4.03 GiB |
| Qwen2.5-7B-Instruct | 4 | 3.81 GiB |
| gemma-3-1b-it | 4 | 0.70 GiB |
| gemma-4-E2B-it-qat-q4_0-unquantized | 1 | 3.12 GiB |