Alternatives to OmniAtlas-Qwen3-30B-A3B
OmniAtlas-Qwen3-30B-A3B alternatives
OmniAtlas-Qwen3-30B-A3B's smallest published quantization is 5.98 GiB. The models below do the same job with a different trade — less memory, comparable size, or a more permissive licence.
From the file· sizes from summed file bytes
Meaningfully smaller
under 70% of its smallest quantization
| Model | Params | Smallest | Quants | Licence |
|---|---|---|---|---|
| gemma-4-E4B-it | 8.0B | 3.30 GiB | 36 | apache-2.0 |
| Qwen3-4B | 4.0B | 1.01 GiB | 29 | — |
| Qwen3-8B | 8.2B | 2.12 GiB | 51 | apache-2.0 |
| HyperCLOVAX-SEED-Text-Instruct-1.5B | 1.6B | 1.06 GiB | 1 | other |
| Llama-3.2-1B-Instruct | 1.2B | 0.39 GiB | 39 | — |
| Qwen3-VL-8B-Instruct-abliterated-v1 | 8.8B | 1.97 GiB | 48 | apache-2.0 |
| Llama-3.1-8B-Instruct | 8.0B | 2.02 GiB | 45 | llama3.1 |
| ced-base | 86M | 0.12 GiB | 3 | apache-2.0 |
Comparable in size
within ±40%, so a like-for-like swap
| Model | Params | Smallest | Quants | Licence |
|---|---|---|---|---|
| Qwen3-Coder-30B-A3B-InstructMoE | 30.5B | 7.46 GiB | 46 | apache-2.0 |
| gemma-4-12B-it-qat-q4_0-unquantized | 12.0B | 6.50 GiB | 2 | apache-2.0 |
| Qwen3-30B-A3B-Thinking-2507MoE | 30.5B | 7.05 GiB | 51 | apache-2.0 |
| Qwen3-30B-A3BMoE | 30.5B | 7.59 GiB | 51 | apache-2.0 |
| llama-3-youko-8b | 8.0B | 5.34 GiB | 2 | llama3 |
| GLM-4.7-FlashMoE | 31.2B | 7.76 GiB | 26 | — |
| Bielik-11B-v3.0-Instruct | 11.2B | 6.26 GiB | 6 | apache-2.0 |
| gemma-4-26B-A4B-it-ultra-uncensored-hereticMoE | 25.8B | 7.72 GiB | 43 | apache-2.0 |
More permissively licensed
licences that allow commercial use
| Model | Params | Smallest | Quants | Licence |
|---|---|---|---|---|
| Qwen3-Coder-30B-A3B-InstructMoE | 30.5B | 7.46 GiB | 46 | apache-2.0 |
| Qwen3.6-27B | 27.8B | 8.74 GiB | 40 | apache-2.0 |
| Qwen3.8-27B | 27.8B | 8.39 GiB | 22 | apache-2.0 |
| DeepSeek-V4-FlashMoE | 291B | 76.87 GiB | 12 | mit |
| gemma-4-12B-it-qat-q4_0-unquantized | 12.0B | 6.50 GiB | 2 | apache-2.0 |
| gemma-4-E4B-it | 8.0B | 3.30 GiB | 36 | apache-2.0 |
| Qwen3-30B-A3B-Thinking-2507MoE | 30.5B | 7.05 GiB | 51 | apache-2.0 |
| Qwen3-8B | 8.2B | 2.12 GiB | 51 | apache-2.0 |