google · text

translategemma-27b-it

google/translategemma-27b-it

translategemma-27b-it at Q4_K_M is exactly 16,546,704,032 bytes (15.41 GiB / 16.55 GB) — an effective 4.590 bits per weight, not the nominal 4.

From the file· summed from 1 file(s)
Parameters
28.8B
Architecture
gemma3
Context
native (config.json)
License
gemma

Shipped quantizations

exact bytes, summed from published files
QuantSizeExact bytesEffective bpwTensorsPublisher
I1-IQ1_S5.83 GiB6,264,263,4241.738mradermacher
I1-IQ1_M6.33 GiB6,797,261,5681.885mradermacher
I1-IQ2_XXS7.16 GiB7,685,591,8082.132mradermacher
I1-IQ2_XS7.86 GiB8,438,919,9362.341mradermacher
I1-IQ2_S8.18 GiB8,782,424,8322.436mradermacher
I1-IQ2_M8.84 GiB9,493,089,0242.633mradermacher
I1-Q2_K_S9.09 GiB9,757,461,8882.707mradermacher
Q2_K9.78 GiB10,503,736,4162.913mradermacher
I1-Q2_K9.78 GiB10,503,736,7042.913mradermacher
I1-IQ3_XXS9.98 GiB10,716,494,5922.973mradermacher
I1-IQ3_XS10.77 GiB11,562,249,6003.207mradermacher
Q3_K_S11.33 GiB12,167,629,9203.375mradermacher
I1-Q3_K_S11.33 GiB12,167,630,2083.375mradermacher
I1-IQ3_S11.33 GiB12,167,630,2083.375mradermacher
I1-IQ3_M11.69 GiB12,547,089,7923.480mradermacher
Q3_K_M12.51 GiB13,437,656,1603.727mradermacher
I1-Q3_K_M12.51 GiB13,437,656,4483.727mradermacher
Q3_K_L13.54 GiB14,543,477,4084.034bullerwins
Q3_K_L13.54 GiB14,543,477,8564.034mradermacher
I1-Q3_K_L13.54 GiB14,543,478,1444.034mradermacher
I1-IQ4_XS13.75 GiB14,767,463,8084.096mradermacher
IQ4_XS13.87 GiB14,893,907,0404.131mradermacher
I1-Q4_014.55 GiB15,617,990,0164.332mradermacher
Q4_K_S14.60 GiB15,674,071,7124.348bullerwins
Q4_K_S14.60 GiB15,674,072,1604.348mradermacher
I1-Q4_K_S14.60 GiB15,674,072,4484.348mradermacher
Q4_K_M15.41 GiB16,546,704,0324.590bullerwins
Q4_K_M15.41 GiB16,546,704,4804.590mradermacher
I1-Q4_K_M15.41 GiB16,546,704,7684.590mradermacher
I1-Q4_115.99 GiB17,167,310,2084.762mradermacher
Q5_K_S17.48 GiB18,767,207,0725.205bullerwins
Q5_K_S17.48 GiB18,767,207,5205.205mradermacher
I1-Q5_K_S17.48 GiB18,767,207,8085.205mradermacher
Q5_K_M17.95 GiB19,271,690,9125.345bullerwins
Q5_K_M17.95 GiB19,271,691,3605.345mradermacher
I1-Q5_K_M17.95 GiB19,271,691,6485.345mradermacher
Q6_K20.64 GiB22,166,989,4726.149bullerwins
Q6_K20.64 GiB22,166,989,9206.149mradermacher
I1-Q6_K20.64 GiB22,166,990,2086.149mradermacher
Q8_026.74 GiB28,707,987,4887.963bullerwins

Compare with

same modality, comparable size

Will it run on your card?

full quant x context sweep

Why other calculators give a different number

A parameters × bits ÷ 8 estimate puts Q4_K_M at roughly 15.11 GiB. The real file is 15.41 GiB, because a quantization is a mixture and some tensors are always kept at higher precision.

Architecture

Architecture unavailable — this repository is gated and no ungated mirror was found. Exact file sizes above are still authoritative; only the KV math needs the config.

Questions people ask

How much VRAM does translategemma-27b-it need?
Q4_K_M is exactly 16,546,704,032 bytes (15.41 GiB) in weights. Add the KV cache, which depends on your context length, plus roughly half a gigabyte of runtime overhead.
Which quantization of translategemma-27b-it should I use?
Q4_K_M is the usual default. Pick the largest quantization that fits your card at the context you actually need — the table above gives exact sizes for every one published.