ZERO-POINT-AI · vision language

MR_MARTIN_31b_V1.0_Gemma-4-base-ftp

ZERO-POINT-AI/MR_MARTIN_31b_V1.0_Gemma-4-base-ftp

MR_MARTIN_31b_V1.0_Gemma-4-base-ftp at I1-IQ1_S is exactly 7,156,473,248 bytes (6.66 GiB / 7.16 GB) — an effective 1.865 bits per weight, not the nominal 1.

From the file· summed from 1 file(s)
Parameters
30.7B
Architecture
gemma4
Context
native (config.json)
License
other

Shipped quantizations

exact bytes, summed from published files
QuantSizeExact bytesEffective bpwTensorsPublisher
I1-IQ1_S6.66 GiB7,156,473,2481.865mradermacher
I1-IQ1_M7.20 GiB7,725,856,1602.013mradermacher
I1-IQ2_XXS8.08 GiB8,674,827,6802.261mradermacher
I1-IQ2_XS8.88 GiB9,530,342,8162.484mradermacher
I1-IQ2_S9.46 GiB10,157,872,5442.647mradermacher
I1-IQ2_M10.17 GiB10,917,049,7602.845mradermacher
I1-Q2_K_S10.22 GiB10,976,572,8322.861mradermacher
I1-Q2_K11.10 GiB11,916,297,6323.106mradermacher
I1-IQ3_XXS11.25 GiB12,077,491,6163.147mradermacher
I1-IQ3_XS12.17 GiB13,072,352,6723.407mradermacher
I1-Q3_K_S12.82 GiB13,761,340,8323.586mradermacher
I1-IQ3_S12.82 GiB13,761,340,8323.586mradermacher
I1-IQ3_M13.43 GiB14,424,481,1843.759mradermacher
I1-Q3_K_M14.24 GiB15,287,092,6403.984mradermacher
I1-Q3_K_L15.49 GiB16,628,254,1124.333mradermacher
I1-IQ4_XS15.59 GiB16,735,774,1124.362mradermacher
I1-Q4_016.49 GiB17,701,561,7604.613mradermacher
I1-Q4_K_S16.54 GiB17,763,149,2164.629mradermacher
I1-Q4_K_M17.40 GiB18,687,047,0724.870mradermacher
I1-Q4_118.14 GiB19,481,404,8325.077mradermacher
I1-Q5_K_S19.85 GiB21,311,825,3125.554mradermacher
I1-Q5_K_M20.35 GiB21,845,554,5925.693mradermacher
I1-Q6_K23.47 GiB25,201,468,8326.568mradermacher

Compare with

same modality, comparable size

Will it run on your card?

full quant x context sweep

Why other calculators give a different number

A parameters × bits ÷ 8 estimate puts I1-IQ1_S at roughly 16.08 GiB. The real file is 6.66 GiB, because a quantization is a mixture and some tensors are always kept at higher precision.

Architecture

Architecture unavailable — this repository is gated and no ungated mirror was found. Exact file sizes above are still authoritative; only the KV math needs the config.

Questions people ask

How much VRAM does MR_MARTIN_31b_V1.0_Gemma-4-base-ftp need?
I1-IQ1_S is exactly 7,156,473,248 bytes (6.66 GiB) in weights. Add the KV cache, which depends on your context length, plus roughly half a gigabyte of runtime overhead.
Which quantization of MR_MARTIN_31b_V1.0_Gemma-4-base-ftp should I use?
Q4_K_M is the usual default. Pick the largest quantization that fits your card at the context you actually need — the table above gives exact sizes for every one published.