Alternatives to Octen-Embedding-4B

Octen-Embedding-4B alternatives

Octen-Embedding-4B's smallest published quantization is 1.55 GiB. The models below do the same job with a different trade — less memory, comparable size, or a more permissive licence.

From the file· sizes from summed file bytes

Meaningfully smaller

under 70% of its smallest quantization
ModelParamsSmallestQuantsLicence
jina-embeddings-v5-text-small596M0.19 GiB56cc-by-nc-4.0
embeddinggemma-300m-qat-q8_0-unquantized303M0.31 GiB1gemma
KaLM-embedding-multilingual-mini-instruct-v2.5494M0.49 GiB1apache-2.0
nomic-embed-text-v1.5137M0.05 GiB39apache-2.0
jina-embeddings-v5-text-nano212M0.09 GiB42cc-by-nc-4.0
all-MiniLM-L6-v223M0.02 GiB28apache-2.0
mxbai-embed-xsmall-v124M0.03 GiB4apache-2.0
nomic-embed-text-v2-moeMoE475M0.25 GiB20apache-2.0

Comparable in size

within ±40%, so a like-for-like swap
ModelParamsSmallestQuantsLicence
Qwen3-Embedding-4B4.0B1.55 GiB18apache-2.0
Nemotron-3-Embed-8B-BF168.0B1.81 GiB36openmdw-1.1
LCO-Embedding-Omni-3B-26054.7B1.96 GiB4apache-2.0
qwen-indic-v17.6B1.78 GiB36apache-2.0

More permissively licensed

licences that allow commercial use
ModelParamsSmallestQuantsLicence
KaLM-embedding-multilingual-mini-instruct-v2.5494M0.49 GiB1apache-2.0
nomic-embed-text-v1.5137M0.05 GiB39apache-2.0
all-MiniLM-L6-v223M0.02 GiB28apache-2.0
mxbai-embed-xsmall-v124M0.03 GiB4apache-2.0
nomic-embed-text-v2-moeMoE475M0.25 GiB20apache-2.0
bge-m3567M0.59 GiB1mit
jina-reranker-v1-tiny-en33M0.03 GiB12apache-2.0
Qwen3-Embedding-0.6B596M0.28 GiB22apache-2.0