Model comparison

Wan2.2-TI2V-5B vs Wan2.2-TI2V-5B-Turbo

At Q4_K_M, Wan2.2-TI2V-5B is the smaller download — 3,433,116,000 bytes against 3,437,927,136.

From the file· summed bytes, KV per layer

Side by side

Wan2.2-TI2V-5BWan2.2-TI2V-5B-Turbo
Parameters5.0B5.0B
Architecturewanwan
Layers
Native context
Mixture of expertsnono
Quantizations published3922
Smallest quantization1.73 GiB1.73 GiB
Q4_K_M3.20 GiB3.20 GiB
Licenceapache-2.0apache-2.0

KV cache by context

the term that decides long-context viability
ContextWan2.2-TI2V-5BWan2.2-TI2V-5B-TurboRatio
4,096
8,192
16,384
32,768
65,536
131,072