embeddinggemma-300m-qat-q8_0-unquantized alternatives

embeddinggemma-300m-qat-q8_0-unquantized's smallest published quantization is 0.31 GiB. The models below do the same job with a different trade — less memory, comparable size, or a more permissive licence.

From the file· sizes from summed file bytes

Meaningfully smaller

under 70% of its smallest quantization
ModelParamsSmallestQuantsLicence
jina-embeddings-v5-text-small596M0.19 GiB56cc-by-nc-4.0
nomic-embed-text-v1.5137M0.05 GiB39apache-2.0
jina-embeddings-v5-text-nano212M0.09 GiB42cc-by-nc-4.0
all-MiniLM-L6-v223M0.02 GiB28apache-2.0
mxbai-embed-xsmall-v124M0.03 GiB4apache-2.0
jina-reranker-v1-tiny-en33M0.03 GiB12apache-2.0
gte-small33M0.02 GiB12mit
nomic-embed-text-v1137M0.05 GiB14apache-2.0

Comparable in size

within ±40%, so a like-for-like swap
ModelParamsSmallestQuantsLicence
nomic-embed-text-v2-moeMoE475M0.25 GiB20apache-2.0
snowflake-arctic-embed-l-v2.0568M0.39 GiB15
Qwen3-Embedding-0.6B596M0.28 GiB22apache-2.0
bge-reranker-v2-m3568M0.41 GiB5apache-2.0

More permissively licensed

licences that allow commercial use
ModelParamsSmallestQuantsLicence
KaLM-embedding-multilingual-mini-instruct-v2.5494M0.49 GiB1apache-2.0
nomic-embed-text-v1.5137M0.05 GiB39apache-2.0
all-MiniLM-L6-v223M0.02 GiB28apache-2.0
mxbai-embed-xsmall-v124M0.03 GiB4apache-2.0
nomic-embed-text-v2-moeMoE475M0.25 GiB20apache-2.0
bge-m3567M0.59 GiB1mit
jina-reranker-v1-tiny-en33M0.03 GiB12apache-2.0
Qwen3-Embedding-0.6B596M0.28 GiB22apache-2.0