0xA50C1A1 · text

Darkmere-14B-v0.1

0xA50C1A1/Darkmere-14B-v0.1

Darkmere-14B-v0.1 at I1-IQ1_S is exactly 3,251,941,984 bytes (3.03 GiB / 3.25 GB) — an effective 1.926 bits per weight, not the nominal 1.

From the file· summed from 1 file(s)
Parameters
13.5B
Architecture
mistral3
Context
native (config.json)
License
apache-2.0

Shipped quantizations

exact bytes, summed from published files
QuantSizeExact bytesEffective bpwTensorsPublisher
I1-IQ1_S3.03 GiB3,251,941,9841.926mradermacher
I1-IQ1_M3.26 GiB3,502,617,1842.075mradermacher
I1-IQ2_XXS3.65 GiB3,920,409,1842.322mradermacher
I1-IQ2_XS3.99 GiB4,280,857,1842.536mradermacher
I1-IQ2_S4.20 GiB4,509,577,8242.671mradermacher
I1-IQ2_M4.51 GiB4,843,811,4242.869mradermacher
I1-Q2_K_S4.58 GiB4,918,850,1442.914mradermacher
I1-Q2_K4.89 GiB5,246,530,1443.108mradermacher
I1-IQ3_XXS5.05 GiB5,427,081,8243.215mradermacher
I1-IQ3_XS5.42 GiB5,817,676,3843.446mradermacher
I1-Q3_K_S5.66 GiB6,074,905,1843.598mradermacher
I1-IQ3_S5.68 GiB6,102,757,9843.615mradermacher
I1-IQ3_M5.84 GiB6,269,874,7843.714mradermacher
I1-Q3_K_M6.22 GiB6,682,096,2243.958mradermacher
I1-Q3_K_L6.72 GiB7,210,316,3844.271mradermacher
I1-IQ4_XS6.90 GiB7,411,184,2244.390mradermacher
I1-Q4_07.27 GiB7,805,710,9444.623mradermacher
I1-IQ4_NL7.27 GiB7,805,710,9444.623mradermacher
I1-Q4_K_S7.30 GiB7,834,546,7844.641mradermacher
I1-Q4_K_M7.67 GiB8,239,067,7444.880mradermacher
I1-Q4_17.99 GiB8,581,657,1845.083mradermacher
I1-Q5_K_S8.74 GiB9,383,817,8245.558mradermacher
I1-Q5_K_M8.96 GiB9,620,566,6245.699mradermacher
I1-Q6_K10.33 GiB11,088,409,1846.568mradermacher

Compare with

same modality, comparable size

Will it run on your card?

full quant x context sweep

Why other calculators give a different number

A parameters × bits ÷ 8 estimate puts I1-IQ1_S at roughly 7.08 GiB. The real file is 3.03 GiB, because a quantization is a mixture and some tensors are always kept at higher precision.

Architecture

Architecture unavailable — this repository is gated and no ungated mirror was found. Exact file sizes above are still authoritative; only the KV math needs the config.

Questions people ask

How much VRAM does Darkmere-14B-v0.1 need?
I1-IQ1_S is exactly 3,251,941,984 bytes (3.03 GiB) in weights. Add the KV cache, which depends on your context length, plus roughly half a gigabyte of runtime overhead.
Which quantization of Darkmere-14B-v0.1 should I use?
Q4_K_M is the usual default. Pick the largest quantization that fits your card at the context you actually need — the table above gives exact sizes for every one published.