0xA50C1A1 · text

Darkmere-8B-v0.1

0xA50C1A1/Darkmere-8B-v0.1

Darkmere-8B-v0.1 at I1-IQ1_S is exactly 2,121,831,872 bytes (1.98 GiB / 2.12 GB) — an effective 2.000 bits per weight, not the nominal 1.

From the file· summed from 1 file(s)
Parameters
8.5B
Architecture
mistral3
Context
native (config.json)
License
apache-2.0

Shipped quantizations

exact bytes, summed from published files
QuantSizeExact bytesEffective bpwTensorsPublisher
I1-IQ1_S1.98 GiB2,121,831,8722.000mradermacher
I1-IQ1_M2.12 GiB2,273,416,6402.142mradermacher
I1-IQ2_XXS2.35 GiB2,526,057,9202.380mradermacher
I1-IQ2_XS2.56 GiB2,745,996,7362.588mradermacher
I1-IQ2_S2.71 GiB2,904,593,8562.737mradermacher
I1-IQ2_M2.89 GiB3,106,706,8802.928mradermacher
I1-Q2_K_S2.93 GiB3,147,273,6642.966mradermacher
I1-Q2_K3.12 GiB3,352,925,6323.160mradermacher
I1-IQ3_XXS3.22 GiB3,455,358,4003.256mradermacher
I1-IQ3_XS3.46 GiB3,714,356,6723.500mradermacher
I1-Q3_K_S3.60 GiB3,866,465,7283.643mradermacher
I1-IQ3_S3.62 GiB3,885,405,6323.661mradermacher
I1-IQ3_M3.72 GiB3,992,360,3843.762mradermacher
I1-Q3_K_M3.95 GiB4,242,052,5443.997mradermacher
I1-Q3_K_L4.25 GiB4,565,013,9524.302mradermacher
I1-IQ4_XS4.37 GiB4,696,413,6324.426mradermacher
I1-Q4_04.60 GiB4,937,323,9684.653mradermacher
I1-IQ4_NL4.60 GiB4,940,469,6964.656mradermacher
I1-Q4_K_S4.61 GiB4,954,101,1844.668mradermacher
I1-Q4_K_M4.84 GiB5,198,386,6244.899mradermacher
I1-Q4_15.05 GiB5,419,668,9285.107mradermacher
I1-Q5_K_S5.51 GiB5,916,693,9525.575mradermacher
I1-Q5_K_M5.64 GiB6,058,743,2325.709mradermacher
I1-Q6_K6.49 GiB6,972,872,1286.571mradermacher

Compare with

same modality, comparable size

Will it run on your card?

full quant x context sweep

Why other calculators give a different number

A parameters × bits ÷ 8 estimate puts I1-IQ1_S at roughly 4.45 GiB. The real file is 1.98 GiB, because a quantization is a mixture and some tensors are always kept at higher precision.

Architecture

Architecture unavailable — this repository is gated and no ungated mirror was found. Exact file sizes above are still authoritative; only the KV math needs the config.

Questions people ask

How much VRAM does Darkmere-8B-v0.1 need?
I1-IQ1_S is exactly 2,121,831,872 bytes (1.98 GiB) in weights. Add the KV cache, which depends on your context length, plus roughly half a gigabyte of runtime overhead.
Which quantization of Darkmere-8B-v0.1 should I use?
Q4_K_M is the usual default. Pick the largest quantization that fits your card at the context you actually need — the table above gives exact sizes for every one published.