GraySoft
Projects Models Compare Cloud benchmarks FAQ Download guIDE →
Model Intelligence Sheet

aimeri/spoomplesmaxx-mockingbird-36B-i1-GGUF overview

spoomplesmaxx mockingbird 36B — i1 GGUF weighted/imatrix Weighted/imatrix GGUF quants of spoomplesmaxx mockingbird 36B https://huggingface.co/aimeri/spoomplesm…

ggufroleplaycreative-writingtext-generationenbase_model:aimeri/spoomplesmaxx-mockingbird-36Bbase_model:quantized:aimeri/spoomplesmaxx-mockingbird-36Blicense:apache-2.0endpoints_compatibleregion:usimatrixconversational

Runs locally from ~13.15 GB disk (16 GB VRAM class GPUs with llama.cpp / guIDE).

Downloads
0
Likes
0
Pipeline
text-generation
Author

Repository Files & Downloads

5 GGUF files detected
Direct downloads for local inference
FileTypeQuantizationSizeLink
spoomplesmaxx-mockingbird-36B.i1-IQ3_XXS.ggufGGUFIQ3_XXS13.15 GBDownload
spoomplesmaxx-mockingbird-36B.i1-IQ4_XS.ggufGGUFIQ4_XS18.16 GBDownload
spoomplesmaxx-mockingbird-36B.i1-Q3_K_M.ggufGGUFQ3_K_M16.41 GBDownload
spoomplesmaxx-mockingbird-36B.i1-Q4_K_M.ggufGGUFQ4_K_M20.27 GBDownload
spoomplesmaxx-mockingbird-36B.i1-Q5_K_M.ggufGGUFQ5_K_M23.84 GBDownload

Model Details

Model IDaimeri/spoomplesmaxx-mockingbird-36B-i1-GGUF
Authoraimeri
Pipelinetext-generation
Licenseapache-2.0
Base modelaimeri/spoomplesmaxx-mockingbird-36B
Last modified2026-08-26T01:16:52.000Z

Model README

---

license: apache-2.0

base_model: aimeri/spoomplesmaxx-mockingbird-36B

library_name: gguf

pipeline_tag: text-generation

tags:

- roleplay

- creative-writing

language:

- en

---

spoomplesmaxx-mockingbird-36B — i1-GGUF (weighted/imatrix)

Weighted/imatrix GGUF quants of

spoomplesmaxx-mockingbird-36B.

The importance matrix was computed on a stratified sample of the model's

own training corpus — all ten lanes, rendered in the exact chat template

the model serves with — not a generic calibration set. At 3–4 bit these

should beat the static quants

noticeably; at Q5 the difference fades.

| Quant | Size | Notes |

|---|---|---|

| i1-IQ3_XXS | ~14 GB | smallest usable; VRAM-desperate only |

| i1-Q3_K_M | ~18 GB | the 18GB target, imatrix-weighted |

| i1-IQ4_XS | ~19 GB | best size/quality trade below Q4_K_M |

| i1-Q4_K_M | ~22 GB | recommended |

| i1-Q5_K_M | ~26 GB | closest to bf16 behavior |

The seed-native chat template is embedded in the GGUF metadata.

Sampling — read this part

temperature 1.0 · top_p 0.9 · repeat_penalty 1.0 (OFF)

> ⚠ Never use repetition, presence, or frequency penalties.

> The template ends every message with <seed:eos>; context-wide penalties

> suppress that token, the model stops ending its turns, and generation

> degenerates into the base model's untrained Chinese vocabulary. Many

> frontend presets default repeat_penalty to 1.05–1.1 — set it back to 1.0.

> Use DRY or XTC if you want extra anti-repetition; both leave special

> tokens alone.

Usable temperature window is ~0.95–1.05: lower loops verbatim, higher frays.

Full details on the

main model card.

mimids 01 · Apache 2.0

Run aimeri/spoomplesmaxx-mockingbird-36B-i1-GGUF with guIDE

Download guIDE — the AI-native code editor with local LLM inference and 69 built-in tools.

Download guIDE → · Browse 524k+ models · Compare models

Source: Hugging Face · Compare models