xiabs · image

DreamOmni2

xiabs/DreamOmni2

DreamOmni2 at Q4_K_M is exactly 4,683,071,936 bytes (4.36 GiB / 4.68 GB) — an effective 4.919 bits per weight, not the nominal 4.

From the file· summed from 1 file(s)
Parameters
7.6B
Architecture
qwen2vl
Context
native (config.json)
License
apache-2.0

Shipped quantizations

exact bytes, summed from published files
QuantSizeExact bytesEffective bpwTensorsPublisher
Q2_K2.81 GiB3,015,938,4963.168rafacost
Q3_K_S3.25 GiB3,492,366,7843.669rafacost
Q3_K_M3.55 GiB3,808,389,5684.001rafacost
Q4_04.13 GiB4,431,389,1204.655rafacost
Q4_K_S4.15 GiB4,457,767,3604.683rafacost
Q4_K_M4.36 GiB4,683,071,9364.919rafacost
Q5_K_S4.95 GiB5,315,174,8485.583rafacost
Q5_04.95 GiB5,315,174,8485.583rafacost
Q5_K_M5.07 GiB5,444,829,6325.720rafacost
Q6_K5.82 GiB6,254,197,1846.570rafacost
Q8_07.54 GiB8,098,523,5848.507rafacost
F1614.19 GiB15,237,851,58416.007rafacost

Compare with

same modality, comparable size

Will it run on your card?

full quant x context sweep

Why other calculators give a different number

A parameters × bits ÷ 8 estimate puts Q4_K_M at roughly 3.99 GiB. The real file is 4.36 GiB, because a quantization is a mixture and some tensors are always kept at higher precision.

Architecture

Architecture unavailable — this repository is gated and no ungated mirror was found. Exact file sizes above are still authoritative; only the KV math needs the config.

Questions people ask

How much VRAM does DreamOmni2 need?
Q4_K_M is exactly 4,683,071,936 bytes (4.36 GiB) in weights. Add the KV cache, which depends on your context length, plus roughly half a gigabyte of runtime overhead.
Which quantization of DreamOmni2 should I use?
Q4_K_M is the usual default. Pick the largest quantization that fits your card at the context you actually need — the table above gives exact sizes for every one published.